arXivDaily arXiv每日学术速递 周一至周五更新

视觉与机器人

3D 视觉

三维重建、NeRF、Gaussian Splatting、点云和空间智能。

2025-12-09 至 2025-12-09 共收录 34 信号源:cs.CV, cs.GR, cs.RO

1. 三维重建 6 篇

2509.15675 2025-12-09 cs.CV 79%

A PCA Based Model for Surface Reconstruction from Incomplete Point Clouds

基于PCA的不完整点云表面重建模型

Hao Liu

机构 * Department of Mathematics, Hong Kong Baptist University(香港 Baptist 大学数学系)

专题命中 三维重建 :point cloud(title,abstract);分类 cs.CV

AI总结 本文提出基于PCA的模型,用于从不完整点云数据中重建表面,通过估计法线信息和操作分裂方法提升重建效果。

详情

展开后加载摘要…

URL PDF HTML 收藏
2512.07504 2025-12-09 cs.CV 57%

ControlVP: Interactive Geometric Refinement of AI-Generated Images with Consistent Vanishing Points

ControlVP: 交互式几何校正AI生成图像以保持一致的消失点

Ryota Okumura, Kaede Shiohara, Toshihiko Yamasaki

机构 * The University of Tokyo(东京大学)

专题命中 三维重建 :3D reconstruction(abstract);分类 cs.CV

AI总结 ControlVP通过结合建筑轮廓的结构指导和几何约束,校正AI生成图像中的消失点不一致,提升空间结构真实性。

Comments Accepted to WACV 2026, 8 pages, supplementary included. Dataset and code: https://github.com/RyotaOkumura/ControlVP

详情

展开后加载摘要…

URL PDF HTML 收藏
2512.07381 2025-12-09 cs.CV 57%

Tessellation GS: Neural Mesh Gaussians for Robust Monocular Reconstruction of Dynamic Objects

Tessellation GS:基于神经网格高斯的鲁棒单目动态物体重建

Shuohan Tao, Boyao Zhou, Hanzhang Tu, Yuwang Wang, Yebin Liu

机构 * University of Cambridge(剑桥大学) Tsinghua University(清华大学)

专题命中 三维重建 :Gaussian Splatting(abstract);分类 cs.CV

AI总结 Tessellation GS通过基于网格面的神经高斯方法实现单目动态物体的鲁棒重建,显著提升了稀疏视角和动态场景下的重建性能。

详情

展开后加载摘要…

URL PDF HTML 收藏
2512.06024 2025-12-09 cs.CV physics.flu-dyn 57%

Neural reconstruction of 3D ocean wave hydrodynamics from camera sensing

基于摄像头传感的神经网络3D海洋波流体力学重建

Jiabin Liu, Zihao Zhou, Jialei Yan, Anxin Guo, Alvise Benetazzo, Hui Li

专题命中 三维重建 :3D reconstruction(abstract);分类 cs.CV

AI总结 该研究提出了一种基于神经网络的3D海洋波流体力学重建方法,通过多尺度注意力机制实现高精度、高效率的波浪自由表面和速度场重建。

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.17482 2025-12-09 cs.RO 57%

Variational Shape Inference for Grasp Diffusion on SE(3)

变分形状推断用于SE(3)上的抓取扩散

S. Talha Bukhari, Kaivalya Agrawal, Zachary Kingston, Aniket Bera

机构 * Department of Computer Science, Purdue University(计算机科学系,普渡大学)

专题命中 三维重建 :point cloud(abstract);分类 cs.RO

AI总结 本文提出一种基于变分形状推断的SE(3)抓取扩散框架,通过隐式神经表示训练自编码器并引导扩散模型,实现鲁棒的多模态抓取合成,实验显示在ACRONYM数据集上性能提升6.3%并具备零样本迁移能力。

Comments This work has been submitted to the IEEE for possible publication

详情

展开后加载摘要…

URL PDF HTML 收藏
2505.08013 2025-12-09 cs.CV 57%

RDD: Robust Feature Detector and Descriptor using Deformable Transformer

RDD: 基于变形Transformer的鲁棒特征检测与描述

Gonglin Chen, Tianwen Fu, Haiwei Chen, Wenbin Teng, Hanyuan Xiao, Yajie Zhao

机构 * Institute for Creative Technologies(创意技术研究所) University of Southern California(南加州大学)

专题命中 三维重建 :3D reconstruction(abstract);分类 cs.CV

AI总结 RDD通过变形Transformer实现鲁棒特征检测与描述,优于现有方法并在稀疏和半密集匹配中表现优异。

详情

展开后加载摘要…

URL PDF HTML 收藏

2. Gaussian Splatting 10 篇

2512.07197 2025-12-09 cs.CV 87%

SUCCESS-GS: Survey of Compactness and Compression for Efficient Static and Dynamic Gaussian Splatting

SUCCESS-GS:紧凑性与压缩的调查:为高效静态和动态高斯点划法服务

Seokhyun Youn, Soohyun Lee, Geonho Kim, Weeyoung Kwon, Sung-Ho Bae, Jihyong Oh

机构 * Chung-Ang University(Chung-Ang 大学) Kyung Hee University(Kyung Hee 大学)

专题命中 Gaussian Splatting :Gaussian Splatting(title,abstract);3DGS(abstract);3D reconstruction(abstract);novel view synthesis(abstract)

AI总结 本文综述了高效3D和4D高斯点划法技术,探讨了参数压缩与重构压缩两大方向,并讨论了其应用与未来研究方向。

Comments The first three authors contributed equally to this work. The last two authors are co-corresponding authors. Please visit our project page at https://cmlab-korea.github.io/Awesome-Efficient-GS/

详情

展开后加载摘要…

URL PDF HTML 收藏
2512.07230 2025-12-09 cs.CV 85%

STRinGS: Selective Text Refinement in Gaussian Splatting

STRinGS: 选择性文本细化在高斯点云中

Abhinav Raundhal, Gaurav Behera, P J Narayanan, Ravi Kiran Sarvadevabhatla, Makarand Tapaswi

机构 * CVIT, IIIT Hyderabad, India(计算机视觉研究所,印度海得拉巴印度理工学院)

专题命中 Gaussian Splatting :Gaussian Splatting(title,abstract);3DGS(abstract);3D reconstruction(abstract);分类 cs.CV

AI总结 STRinGS提出了一种文本感知的选性细化框架,通过分离文本和非文本区域并优化全场景,提升3DGS重建中文本的清晰度和可读性,同时引入OCR字符错误率评估和STRinGS-360数据集。

Comments Accepted to WACV 2026. Project Page, see https://STRinGS-official.github.io

详情

展开后加载摘要…

URL PDF HTML 收藏
2511.04283 2025-12-09 cs.CV 85%

FastGS: Training 3D Gaussian Splatting in 100 Seconds

FastGS:在100秒内训练3D高斯溅射

Shiwei Ren, Tianci Wen, Yongchun Fang, Biao Lu

机构 * NanKai University(南开大学)

专题命中 Gaussian Splatting :Gaussian Splatting(title,abstract);NeRF(abstract);3DGS(abstract);分类 cs.CV

AI总结 FastGS通过基于多视角一致性的致密化和修剪策略,实现了在100秒内高效训练3D高斯溅射,显著提升训练速度并保持渲染质量。

Comments Project page: https://fastgs.github.io/

详情

展开后加载摘要…

URL PDF HTML 收藏
2512.07247 2025-12-09 cs.CV cs.CR cs.LG 83%

AdLift: Lifting Adversarial Perturbations to Safeguard 3D Gaussian Splatting Assets Against Instruction-Driven Editing

AdLift: 将对抗扰动提升以保护3D高斯点云资产免受指令驱动编辑的影响

Ziming Hong, Tianyu Huang, Runnan Chen, Shanshan Ye, Mingming Gong, Bo Han, Tongliang Liu

机构 * Sydney AI Centre, The University of Sydney(悉尼人工智能中心,悉尼大学) University of Technology Sydney(技术大学悉尼) University of Melbourne(墨尔本大学) Hong Kong Baptist University(香港 Baptist 大学) Mohamed bin Zayed University of Artificial Intelligence(Mohamed bin Zayed 人工智能大学)

专题命中 Gaussian Splatting :Gaussian Splatting(title,abstract);3DGS(abstract);分类 cs.CV

AI总结 AdLift通过提升2D对抗扰动至3D高斯表示,有效保护3DGS资产免受指令驱动编辑攻击。

Comments 40 pages, 34 figures, 18 tables

详情

展开后加载摘要…

URL PDF HTML 收藏
2512.06269 2025-12-09 cs.CV 83%

TriaGS: Differentiable Triangulation-Guided Geometric Consistency for 3D Gaussian Splatting

TriaGS: 可微三角化引导的几何一致性用于3D高斯点扩散

Quan Tran, Tuan Dang

机构 * VinUniversity Ha Noi, Vietnam(越南河内Vin大学) University of Arkansas Fayetteville(阿肯色大学弗拉特维尔分校)

专题命中 Gaussian Splatting :Gaussian Splatting(title,abstract);novel view synthesis(abstract);分类 cs.CV

AI总结 TriaGS通过约束多视角三角化提升3D高斯点扩散的几何一致性,实现高保真表面重建。

Comments 10 pages

Journal ref WACV 2026

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.07118 2025-12-09 cs.RO 83%

DexFruit: Dexterous Manipulation and Gaussian Splatting Inspection of Fruit

DexFruit:灵巧操作与高分辨率3D高斯点云损伤检测

Aiden Swann, Alex Qiu, Matthew Strong, Angelina Zhang, Samuel Morstein, Kai Rayle, Monroe Kennedy

机构 * Department of Mechanical Engineering(机械工程系) Department of Computer Science(计算机科学系) Stanford University(斯坦福大学)

专题命中 Gaussian Splatting :Gaussian Splatting(title,abstract);3DGS(abstract);分类 cs.RO

AI总结 DexFruit通过光学触觉传感实现对水果的自主操作,减少损伤并提升抓取成功率,引入FruitSplat技术量化3D高斯点云中的视觉损伤。

Comments 8 pages, 5 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2501.17792 2025-12-09 cs.CV 79%

CrowdSplat: Exploring Gaussian Splatting For Crowd Rendering

CrowdSplat:探索高斯散射用于人群渲染

Xiaohan Sun, Yinghan Xu, John Dingliana, Carol O'Sullivan

专题命中 Gaussian Splatting :Gaussian Splatting(title,abstract);分类 cs.CV

AI总结 CrowdSplat通过3D高斯散射技术实现高质量实时人群渲染,结合LoD渲染和内存优化,提升计算效率与渲染质量。

Comments 4 pages, 4 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2512.07165 2025-12-09 cs.CV 70%

MuSASplat: Efficient Sparse-View 3D Gaussian Splats via Lightweight Multi-Scale Adaptation

MuSASplat: 通过轻量级多尺度适应实现高效的稀疏视角3D高斯喷射

Muyu Xu, Fangneng Zhan, Xiaoqin Zhang, Ling Shao, Shijian Lu

专题命中 Gaussian Splatting :Gaussian Splatting(abstract);novel view synthesis(abstract);分类 cs.CV

AI总结 MuSASplat通过轻量级多尺度适应技术,高效实现稀疏视角3D高斯喷射,显著降低计算成本并保持高质量渲染效果。

详情

展开后加载摘要…

URL PDF HTML 收藏
2512.05113 2025-12-09 cs.CV 57%

Splannequin: Freezing Monocular Mannequin-Challenge Footage with Dual-Detection Splatting

Splannequin: 通过双检测溅射实现单目人偶挑战视频的冻结3D场景

Hao-Jen Chien, Yi-Chuan Huang, Chung-Ho Wu, Wei-Lun Chao, Yu-Lun Liu

机构 * National Yang Ming Chiao Tung University The Ohio State University

专题命中 Gaussian Splatting :Gaussian Splatting(abstract);分类 cs.CV

AI总结 Splannequin通过双检测溅射技术,提升单目人偶挑战视频冻结3D场景的视觉质量,实现高保真用户可控的冻结时间渲染。

Comments WACV 2026. Project page: https://chien90190.github.io/splannequin/

详情

展开后加载摘要…

URL PDF HTML 收藏
2512.06158 2025-12-09 cs.CV 57%

Tracking-Guided 4D Generation: Foundation-Tracker Motion Priors for 3D Model Animation

基于跟踪的4D生成:用于3D模型动画的基础跟踪器运动先验

Su Sun, Cheng Zhao, Himangi Mittal, Gaurav Mittal, Rohith Kukkala, Yingjie Victor Chen, Mei Chen

机构 * Purdue University(普渡大学) Carnegie Mellon University(卡内基梅隆大学) Microsoft(微软)

专题命中 Gaussian Splatting :Gaussian Splatting(abstract);分类 cs.CV

AI总结 Track4DGen通过结合基础跟踪器和混合4D高斯点划法,提升多视角视频生成和4D生成的稳定性与精度。

Comments 15 pages, 11 figures

详情

展开后加载摘要…

URL PDF HTML 收藏

3. 点云 10 篇

2511.11210 2025-12-09 cs.CV 83%

STONE: Pioneering the One-to-N Universal Backdoor Threat in 3D Point Cloud

STONE:开创性地探索3D点云中的一对多通用后门威胁

Dongmei Shan, Wei Lian, Chongxia Wang

机构 * Changzhi University(长治大学)

专题命中 点云 :point cloud(title,abstract);3D vision(abstract);分类 cs.CV

AI总结 STONE首次提出3D点云中一对多通用后门威胁,通过可配置球形触发器设计实现高攻击成功率,为多目标后门威胁提供理论与实证基础。

Comments 15 pages, 5 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2511.07040 2025-12-09 cs.CV cs.CR 79%

3D-ANC: Adaptive Neural Collapse for Robust 3D Point Cloud Recognition

3D-ANC:适应性神经坍塌用于鲁棒的3D点云识别

Yuanmin Huang, Wenxuan Li, Mi Zhang, Xiaohan Zhang, Xiaoyu You, Min Yang

专题命中 点云 :point cloud(title,abstract);分类 cs.CV

AI总结 3D-ANC通过神经坍塌机制解决3D点云识别中的对抗攻击问题,结合ETF对齐和自适应训练框架提升模型鲁棒性。

Comments AAAI 2026

详情

展开后加载摘要…

URL PDF HTML 收藏
2512.06882 2025-12-09 cs.CV 79%

Hierarchical Image-Guided 3D Point Cloud Segmentation in Industrial Scenes via Multi-View Bayesian Fusion

基于多视角贝叶斯融合的工业场景分层图像引导3D点云分割

Yu Zhu, Naoya Chiba, Koichi Hashimoto

专题命中 点云 :point cloud(title,abstract);分类 cs.CV

AI总结 本文提出一种分层图像引导的3D点云分割框架,通过多视角贝叶斯融合解决工业场景中的遮挡和语义不一致问题,提升分割精度与鲁棒性。

Comments Accepted to BMVC 2025 (Sheffield, UK, Nov 24-27, 2025). Supplementary video and poster available upon request

详情

展开后加载摘要…

URL PDF HTML 收藏
2512.06058 2025-12-09 cs.CV 79%

Representation Learning for Point Cloud Understanding

点云理解的表示学习

Siming Yan

机构 * The University of Texas at Austin(德克萨斯大学奥斯汀分校)

专题命中 点云 :point cloud(title,abstract);分类 cs.CV

AI总结 本研究提出整合预训练2D模型以提升3D点云理解的方法,通过自监督学习和迁移学习改进点云表示学习。

Comments 181 pages

详情

展开后加载摘要…

URL PDF HTML 收藏
2512.07211 2025-12-09 cs.CV 74%

Object Pose Distribution Estimation for Determining Revolution and Reflection Uncertainty in Point Clouds

点云中确定旋转和反射不确定性时的物体姿态分布估计

Frederik Hagelskjær, Dimitrios Arapis, Steffen Madsen, Thorbjørn Mosekjær Iversen

机构 * SDU Robotics(SDU机器人研究所) The Mærsk Mc-Kinney Møller Institute(马士基麦金尼莫勒研究所) University of Southern Denmark(南部丹麦大学)

专题命中 点云 :point cloud(title);分类 cs.CV

AI总结 本文提出了一种基于深度学习的点云姿态分布估计方法,无需RGB输入,用于评估旋转和反射的不确定性。

Comments 8 pages, 8 figures, 5 tables, ICCR 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2503.19912 2025-12-09 cs.CV cs.LG cs.RO 62%

Enhanced Spatiotemporal Consistency for Image-to-LiDAR Data Pretraining

增强的时空一致性用于图像到LiDAR数据预训练

Xiang Xu, Lingdong Kong, Hui Shuai, Wenwei Zhang, Liang Pan, Kai Chen, Ziwei Liu, Qingshan Liu

机构 * College of Computer Science and Technology, Nanjing University of Aeronautics and Astronautics(南京航空航天大学计算机科学与技术学院) School of Computing, Department of Computer Science, National University of Singapore(新加坡国立大学计算机学院) School of Computer Science, Nanjing University of Posts and Telecommunications(南京邮电大学计算机学院) Shanghai AI Laboratory(上海人工智能实验室) S-Lab, Nanyang Technological University(南洋理工大学S实验室)

专题命中 点云 :point cloud(abstract);分类 cs.CV、cs.RO

AI总结 SuperFlow++通过整合时空线索提升图像到LiDAR数据预训练效果,实现更鲁棒的特征表示和更高效的自动驾驶感知。

Comments IEEE Transactions on Pattern Analysis and Machine Intelligence

详情

展开后加载摘要…

URL PDF HTML 收藏
2512.07599 2025-12-09 cs.CV 57%

Online Segment Any 3D Thing as Instance Tracking

在线分割三维物体作为实例跟踪

Hanshi Wang, Zijian Cai, Jin Gao, Yiwei Zhang, Weiming Hu, Ke Wang, Zhipeng Zhang

机构 * State Key Laboratory of Multimodal Artificial Intelligence Systems (MAIS), CASIA(多模态人工智能系统国家重点实验室(MAIS),中国科学院自动化所) School of Artificial Intelligence, University of Chinese Academy of Sciences(中国科学院大学人工智能学院) AutoLab, School of Artificial Intelligence, Shanghai Jiao Tong University(上海交通大学人工智能学院AutoLab) Anyverse Intelligence Beijing Key Laboratory of Super Intelligent Security of Multi-Modal Information(北京超智能多模态信息安全重点实验室) School of Information Science and Technology, ShanghaiTech University(上海科技大学信息科学与技术学院)

专题命中 点云 :point cloud(abstract);分类 cs.CV

AI总结 将在线3D分割重新定义为实例跟踪问题,通过时间信息传播和空间一致性学习提升具身智能体对环境的理解能力。

Comments NeurIPS 2025, Code is at https://github.com/AutoLab-SAI-SJTU/AutoSeg3D

详情

展开后加载摘要…

URL PDF HTML 收藏
2502.09620 2025-12-09 cs.CV cs.AI cs.CL 57%

Exploring the Potential of Encoder-free Architectures in 3D LMMs

探索无编码器架构在3D大语言模型中的潜力

Yiwen Tang, Zoey Guo, Zhuhao Wang, Ray Zhang, Qizhi Chen, Junli Liu, Delin Qu, Zhigang Wang, Dong Wang, Bin Zhao, Xuelong Li

机构 * Northwestern Polytechnical University(西北工业大学) Shanghai AI Laboratory(上海人工智能实验室) The Chinese University of Hong Kong(香港中文大学) Tsinghua University(清华大学) Tele AI

专题命中 点云 :point cloud(abstract);分类 cs.CV

AI总结 本文提出首个无编码器3D大语言模型ENEL,通过嵌入语义编码和分层几何聚合策略,在3D理解任务中达到与SOTA模型相当的性能。

详情

展开后加载摘要…

URL PDF HTML 收藏
2512.06517 2025-12-09 cs.RO 57%

Vision-Guided Grasp Planning for Prosthetic Hands in Unstructured Environments

面向无结构环境的假手视觉引导抓取规划

Shifa Sulaiman, Akash Bachhar, Ming Shen, Simon Bøgh

机构 * Control and Automation Section, Department of Electronics Systems, Aalborg University, Denmark(奥尔堡大学电子系统系自动化与自动化部门) Department of Mechanical Engineering, National Institute of Technology,Durgapur, India(印度德里加尔帕国家理工学院机械工程系) The Technical Faculty of IT(信息技术技术学院) Millimeter-Wave Systems, Department of Electronics Systems, Aalborg University, Denmark(毫米波系统,电子系统系,奥尔堡大学,丹麦)

专题命中 点云 :point cloud(abstract);分类 cs.RO

AI总结 本文提出一种基于视觉引导的假手抓取算法,通过整合感知、规划和控制,实现灵活的抓取操作,并在仿真和实验中验证其在无结构环境中的适应性。

详情

展开后加载摘要…

URL PDF HTML 收藏
2502.09274 2025-12-09 cs.CV 57%

FLARES: Fast and Accurate LiDAR Multi-Range Semantic Segmentation

FLARES: 快速且准确的LiDAR多范围语义分割

Bin Yang, Alexandru Paul Condurache

机构 * Automated Driving Research, Robert Bosch GmbH(罗伯特博世集团自动化驾驶研究部) Institute for Signal Processing, University of Lübeck(吕贝克大学信号处理研究所)

专题命中 点云 :point cloud(abstract);分类 cs.CV

AI总结 FLARES通过多范围图像训练和定制的数据增强技术,提升LiDAR语义分割的准确性和效率,实现mIoU提升和推理速度提升。

Comments The paper was accepted by WACV2026

详情

展开后加载摘要…

URL PDF HTML 收藏

4. 新视角合成 1 篇

2512.06818 2025-12-09 cs.CV 70%

MeshSplatting: Differentiable Rendering with Opaque Meshes

基于网格的可微渲染:使用不透明网格的可微渲染

Jan Held, Sanghyun Son, Renaud Vandeghen, Daniel Rebain, Matheus Gadelha, Yi Zhou, Anthony Cioppa, Ming C. Lin, Marc Van Droogenbroeck, Andrea Tagliasacchi

机构 * University of Liège(里耶克斯大学) Simon Fraser University(西蒙弗雷泽大学) University of Maryland(马里兰大学) University of British Columbia(不列颠哥伦比亚大学) University of Toronto(多伦多大学) Adobe Research(Adobe研究院)

专题命中 新视角合成 :Gaussian Splatting(abstract);novel view synthesis(abstract);分类 cs.CV

AI总结 MeshSplatting通过可微渲染方法,实现了基于网格的高效重建,提升了实时3D场景交互的质量和性能。

详情

展开后加载摘要…

URL PDF HTML 收藏

5. 3D生成 2 篇

2512.07628 2025-12-09 cs.CV 79%

MoCA: Mixture-of-Components Attention for Scalable Compositional 3D Generation

MoCA:基于组件的注意力机制用于可扩展的组合3D生成

Zhiqi Li, Wenhuan Li, Tengfei Wang, Zhenwei Wang, Junta Wu, Haoyuan Wang, Yunhan Yang, Zehuan Huang, Yang Li, Peidong Liu, Chunchao Guo

机构 * Zhejiang University(浙江大学) Westlake University(西湖大学) Tencent Hunyuan(腾讯文恩)

专题命中 3D生成 :3D generation(title,abstract);分类 cs.CV

AI总结 MoCA通过重要性路由和组件压缩技术,实现高效可扩展的组合性3D生成,优于现有基线方法。

详情

展开后加载摘要…

URL PDF HTML 收藏
2512.06424 2025-12-09 cs.CV 74%

DragMesh: Interactive 3D Generation Made Easy

DragMesh: 交互式3D生成变得容易

Tianshan Zhang, Zeyu Zhang, Hao Tang

机构 * School of Computer Science, Peking University(北京大学计算机学院)

专题命中 3D生成 :3D generation(title);分类 cs.CV

AI总结 DragMesh通过解耦运动学推理和运动生成,实现实时交互式3D生成,无需重新训练即可在新对象上生成合理运动。

详情

展开后加载摘要…

URL PDF HTML 收藏

6. 空间理解 3 篇

2512.07472 2025-12-09 cs.RO cs.LG 57%

Affordance Field Intervention: Enabling VLAs to Escape Memory Traps in Robotic Manipulation

可及场干预:使VLAs在机器人操作中摆脱记忆陷阱

Siyu Xu, Zijian Wang, Yunke Wang, Chenghao Xia, Tao Huang, Chang Xu

机构 * School of Computer Science, The University of Sydney(悉尼大学计算机科学学院) John Hopcropt Center for Computer Science, Shanghai Jiao Tong University(上海交通大学计算机科学中心)

专题命中 空间理解 :spatial understanding(abstract);分类 cs.RO

AI总结 本研究提出可及场干预(AFI)方法,通过引入3D空间可及场提升VLA在机器人操作中对分布变化的鲁棒性。

详情

展开后加载摘要…

URL PDF HTML 收藏