arXivDaily arXiv每日学术速递 周一至周五更新

视觉与机器人

机器人 / 具身智能

机器人、具身智能、机器人学习、操作、导航和具身世界模型。

2026-01-15 至 2026-01-15 共收录 40 信号源:cs.RO, cs.AI, cs.CV, cs.LG

1. 机器人操作 10 篇

2601.09104 2026-01-15 cs.RO 79%

Design Methodology of Hydraulically-driven Soft Robotic Gripper for a Large and Heavy Object

为大重量物体设计的液压驱动软机器人夹具方法学

Ko Yamamoto, Kyosuke Ishibashi, Hiroki Ishikawa, Osamu Azami

专题命中 机器人操作 :robotic(title,abstract);分类 cs.RO

AI总结 本文提出了一种液压驱动的软机器人夹具设计方法,通过数学模型和材料选择实现大重量物体的抓取与控制。

详情

展开后加载摘要…

URL PDF HTML 收藏
2601.08849 2026-01-15 cs.CL 78%

Gaming the Answer Matcher: Examining the Impact of Text Manipulation on Automated Judgment

利用答案匹配器:考察文本操纵对自动判断的影响

Manas Khatore, Sumana Sridharan, Kevork Sulahian, Benjamin J. Smith, Shi Feng

专题命中 机器人操作 :manipulation(title,abstract)

AI总结 研究通过测试文本操纵对自动答案匹配的影响,发现其对模型评分鲁棒,且在有参考答案时可作为替代评估方法。

Comments Accepted to the AAAI 2026 Workshop on AI Governance (AIGOV)

详情

展开后加载摘要…

URL PDF HTML 收藏
2506.13957 2026-01-15 cs.RO 77%

A Cooperative Contactless Object Transport with Acoustic Robots

基于声学机器人的协同非接触式物体运输

Narsimlu Kemsaram, Akin Delibasi, James Hardwick, Bonot Gautam, Diego Martinez Plasencia, Sriram Subramanian

机构 * University College London(伦敦大学学院)

专题命中 机器人操作 :robotics(abstract);manipulation(abstract);robotic(abstract);分类 cs.RO

AI总结 本文提出了一种基于声学机器人的非接触式空中物体运输系统,通过相位超声波换能器和机器人控制系统实现物体的精确操控,并验证了其在独立和协同运输中的可行性。

Comments This paper has been accepted for publication in the Proceedings of the 2025 IEEE/RSJ International Conference on Intelligent Robots and Systems (IROS 2025) as oral presentation, 8 pages with 8 figures

Journal ref In Proc. 2025 IEEE/RSJ International Conference on Intelligent Robots and Systems (IROS), Hangzhou, China, Oct. 2025, pp. 18043-18050

详情

展开后加载摘要…

URL PDF HTML 收藏
2506.02955 2026-01-15 cs.RO cs.AI cs.LG 67%

UniConFlow: A Unified Constrained Flow-Matching Framework for Certified Motion Planning

UniConFlow: 一种统一的约束流匹配框架用于认证运动规划

Zewen Yang, Xiaobing Dai, Dian Yu, Zhijun Li, Majid Khadiv, Sandra Hirche, Sami Haddadin

机构 * Technical University of Munich (TUM)(技术大学慕尼黑) School of Mechanical Engineering, Translational Research Center, Tongji University(机械工程学院、转化研究中心,同济大学) Mohamed Bin Zayed University of Artificial Intelligence(穆罕默德·本·扎耶德人工智能大学)

专题命中 机器人操作 :manipulation(abstract);分类 cs.RO、cs.AI、cs.LG

AI总结 UniConFlow提出一种统一约束流匹配框架,通过整合等式和不等式约束,提升轨迹生成在安全性和动态一致性上的性能。

详情

展开后加载摘要…

URL PDF HTML 收藏
2601.08873 2026-01-15 cs.CV cs.AI cs.LG cs.MM 67%

ForensicFormer: Hierarchical Multi-Scale Reasoning for Cross-Domain Image Forgery Detection

ForensicFormer: 基于层次多尺度推理的跨领域图像伪造检测

Hema Hariharan Samson

机构 * Independent Researcher(独立研究者)

专题命中 机器人操作 :manipulation(abstract);分类 cs.AI、cs.CV、cs.LG

AI总结 ForensicFormer通过层次多尺度推理框架,实现跨领域图像伪造检测的高准确率与鲁棒性,显著提升对AI生成图像的检测能力。

Comments 9 pages, 4 figures, 5 tables. Technical report on hierarchical multi-scale image forgery detection

详情

展开后加载摘要…

URL PDF HTML 收藏
2506.12335 2026-01-15 cs.CV cs.AI cs.DC 62%

GroupNL: Low-Resource and Robust CNN Design over Cloud and Device

GroupNL: 低资源和鲁棒的云和设备上的CNN设计

Chuntao Ding, Jianhang Xie, Junna Zhang, Salman Raza, Shangguang Wang, Jiannong Cao

机构 * School of Artificial Intelligence, Beijing Normal University(人工智能学院,北京师范大学) School of Computer Science and Technology, Beijing Jiaotong University(计算机科学与技术学院,北京交通大学) Department of Computer Science, City University of Hong Kong(计算机科学系,城市大学) School of Computer and Information Engineering, Henan Normal University(计算机与信息工程学院,河南师范大学) Department of Computer Science, National Textile University Faisalabad(计算机科学系,国立纺织大学费萨尔巴德分校) State Key Laboratory of Networking and Switching Technology, Beijing University of Posts and Telecommunications(网络与交换技术国家重点实验室,北京邮电大学) Department of Computing, The Hong Kong Polytechnic University(计算系,香港理工大学)

专题命中 机器人操作 :manipulation(abstract);分类 cs.AI、cs.CV

AI总结 GroupNL通过轻量级非线性变换函数生成多样化特征图,降低资源消耗并提高CNN鲁棒性,实验显示其在多个数据集上均优于传统卷积层。

Comments IEEE Transactions on Mobile Computing, accepted manuscript

详情

展开后加载摘要…

URL PDF HTML 收藏
2407.10570 2026-01-15 cs.RO 61%

Virtual-force Based Visual Servo for Multiple Peg-in-Hole Assembly with Tightly Coupled Multi-Manipulator

基于虚拟力的多机械臂多插孔装配视觉伺服控制

Jiawei Zhang, Chengchao Bai, Wei Pan, Jifeng Guo

机构 * Harbin Institute of Technology, China(哈尔滨工业大学)

专题命中 机器人操作 :robotic(abstract);分类 cs.RO;robotics(comments)

AI总结 本文提出基于虚拟力的多机械臂视觉伺服控制方法,通过整合多机械臂的视觉特征提升多插孔装配任务的精度与鲁棒性。

Comments 8 pages, 11 figures, this paper has been published by IEEE Robotics and Automation Letters

详情

展开后加载摘要…

URL PDF HTML 收藏
1804.08063 2026-01-15 physics.optics 50%

Adhesion-assisted nanoscale rotary locomotor in non-liquid environments

非液体环境中基于粘附的纳米尺度旋转运动

Jinsheng Lu, Qiang Li, Cheng-Wei Qiu, Min Qiu

专题命中 机器人操作 :manipulation(abstract)

AI总结 本研究在非液体环境中实现基于粘附的纳米级旋转运动,利用光热效应驱动微米级金属纳米板绕微纤维旋转,展示高精度光驱动微镜扫描技术。

Comments 13 pages, 5 figures

Journal ref Science Advances, 5 (2019), eaau8271

详情

展开后加载摘要…

URL PDF HTML 收藏
2601.09140 2026-01-15 physics.optics 50%

Space and space-time topologies in a type-II hyperbolic lattice

二维双曲晶格中的空间与时空拓扑

Jingming Chen, Zebin Zhu, Minqi Cheng, Linyun Yang, Yuxin Zhong, Zhen Gao

专题命中 机器人操作 :manipulation(abstract)

AI总结 本文提出了一种新型双曲晶格,通过实验实现了双曲陈绝缘体,并观测到内外边缘的反向 chirality 边缘态,展示了反时间奇偶性相变,为双曲时空拓扑态的动态操控提供了新范式。

详情

展开后加载摘要…

URL PDF HTML 收藏
2410.21135 2026-01-15 cond-mat.quant-gas cond-mat.str-el quant-ph 50%

Interplay of Unidirectional Quantum Strings in Kagome Rydberg Atom Array

Kagome Rydberg 原子阵列中单向量子弦的相互作用

Wei Xu, Xue-Feng Zhang

专题命中 机器人操作 :manipulation(abstract)

AI总结 该研究通过新开发的量子蒙特卡洛方法,研究Kagome Rydberg原子阵列中量子弦的单向相互作用,揭示了其几何约束下的物理现象。

Comments 9 pages, 8 figures, almost published version, comments are welcome, and more information at http://cqutp.org/users/xfzhang/

Journal ref Published in Phys. Rev. B 113, L020408 (2026)

详情

展开后加载摘要…

URL PDF HTML 收藏

2. 具身导航 13 篇

2601.08868 2026-01-15 cs.CV cs.AI cs.RO 85%

Residual Cross-Modal Fusion Networks for Audio-Visual Navigation

残差跨模态融合网络用于音频视觉导航

Yi Wang, Yinfeng Yu, Bin Ren

机构 * School of Computer Science and Technology, Xinjiang University, Urumqi, China(新疆大学计算机科学与技术学院) Joint International Research Laboratory of Silk Road Multilingual Cognitive(丝绸之路多语认知联合国际实验室) School of Mechatronic Engineering and Automation Shanghai University, Shanghai, China(上海大学机电工程与自动化学院)

专题命中 具身导航 :navigation(title,abstract);embodied agent(abstract);分类 cs.RO、cs.AI、cs.CV

AI总结 本文提出残差跨模态融合网络,通过双向残差交互实现音频视觉信息互补建模,提升跨域导航性能。

Comments Main paper (10 pages). Accepted for publication by the 14th international conference on Computational Visual Media (CVM 2026)

详情

展开后加载摘要…

URL PDF HTML 收藏
2601.08953 2026-01-15 cs.RO cs.AI 84%

Fairness risk and its privacy-enabled solution in AI-driven robotic applications

人工智能驱动机器人应用中的公平性风险及其隐私增强解决方案

Le Liu, Bangguo Yu, Nynke Vellinga, Ming Cao

专题命中 具身导航 :robotic(title,abstract);navigation(abstract);分类 cs.RO、cs.AI

AI总结 本文提出了一种基于效用的公平性度量标准,用于机器人决策,并分析公平性与隐私的联合关系,通过机器人导航任务验证了隐私预算对公平性的联合影响。

详情

展开后加载摘要…

URL PDF HTML 收藏
2601.09318 2026-01-15 cs.RO cs.SY eess.SY 79%

Feedback-Based Mobile Robot Navigation in 3-D Environments Using Artificial Potential Functions Technical Report

基于反馈的三维环境移动机器人导航人工势场方法技术报告

Ro'i Lang, Elon Rimon

专题命中 具身导航 :navigation(title,abstract);分类 cs.RO

AI总结 本文提出了一种用于三维环境中移动机器人导航的人工势场方法,通过多项式函数构造并验证了在存在障碍物时的导航性能。

详情

展开后加载摘要…

URL PDF HTML 收藏
2601.09248 2026-01-15 cs.CV cs.AI 73%

Hybrid guided variational autoencoder for visual place recognition

混合引导变分自编码器用于视觉地点识别

Ni Wang, Zihan You, Emre Neftci, Thorben Schoepe

机构 * Amazon Development Center Germany GmbH, Berlin, Germany(亚马逊德国开发中心) Southeast University, Nanjing, China(东南大学) Forschungszentrum Jülich GmbH, Aachen, Germany(尤利希研究中心) imec, Luven, Belgium(imec)

专题命中 具身导航 :robotics(abstract);navigation(abstract);分类 cs.AI、cs.CV

AI总结 本文提出一种混合引导变分自编码器用于视觉地点识别,结合事件视觉传感器和脉冲神经网络,实现紧凑、鲁棒且具有泛化能力的模型,提升移动机器人导航性能。

详情

展开后加载摘要…

URL PDF HTML 收藏
2601.09578 2026-01-15 cs.RO cs.CV 62%

Multimodal Signal Processing For Thermo-Visible-Lidar Fusion In Real-time 3D Semantic Mapping

多模态信号处理用于热-可见-激光雷达融合的实时3D语义制图

Jiajun Sun, Yangyi Ou, Haoyuan Zheng, Chao yang, Yue Ma

机构 * College of Mechatronics and Control Engineering, Shenzhen University(深圳大学机械与控制工程学院) School of Robotics, Xi’an-Jiaotong Liverpool University(西安交通大学利物浦大学机器人学院)

专题命中 具身导航 :navigation(abstract);分类 cs.RO、cs.CV

AI总结 本文提出通过多模态信号处理融合热、可见和激光雷达数据,提升实时3D语义制图的精度与语义理解能力,适用于灾害评估和工业维护等场景。

Comments 5 pages,7 figures. Under review

详情

展开后加载摘要…

URL PDF HTML 收藏
2505.07808 2026-01-15 cs.RO 61%

AcoustoBots: A swarm of robots for acoustophoretic multimodal interactions

AcoustoBots:用于声学电势多模交互的机器人群

Narsimlu Kemsaram, James Hardwick, Jincheng Wang, Bonot Gautam, Ceylan Besevli, Giorgos Christopoulos, Sourabh Dogra, Lei Gao, Akin Delibasi, Diego Martinez Plasencia, Orestis Georgiou, Marianna Obrist, Ryuji Hirayama, Sriram Subramanian

专题命中 具身导航 :robotic(abstract);分类 cs.RO;robotics(journal_ref)

AI总结 AcoustoBots通过可移动相位阵列和BeadDispenserBot实现声学电势多模交互,提升机器人群的灵活性和交互能力。

Journal ref Frontiers in Robotics and AI, 12:1537101, 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2601.09665 2026-01-15 cs.CV 57%

SCE-SLAM: Scale-Consistent Monocular SLAM via Scene Coordinate Embeddings

SCE-SLAM:通过场景坐标嵌入实现尺度一致的单目SLAM

Yuchen Wu, Jiahe Li, Xiaohan Yu, Lina Yu, Jin Zheng, Xiao Bai

机构 * School of Computer Science and Engineering, State Key Laboratory of Complex & Critical Software Environment, Jiangxi Research Institute, Beihang University(计算机科学与工程学院、复杂与关键软件环境国家重点实验室、江西研究院、北航) Macquarie University(麦考瑞大学) Beijing Key Laboratory of Semiconductor Neural Network Intelligent Sensing and Computing Technology(北京半导体神经网络智能感知与计算技术重点实验室)

专题命中 具身导航 :navigation(abstract);分类 cs.CV

AI总结 SCE-SLAM通过场景坐标嵌入实现单目SLAM的尺度一致性,实验显示在KITTI上轨迹误差减少8.36米,同时保持36 FPS。

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.03547 2026-01-15 cs.RO 57%

Shape-Space Graphs: Fast and Collision-Free Path Planning for Soft Robots

形状空间图:用于软机器人快速且无碰撞路径规划

Carina Veil, Moritz Flaschel, Ellen Kuhl

机构 * Department of Mechanical Engineering, Stanford University(机械工程系,斯坦福大学) Institute of Applied Mechanics, Friedrich-Alexander-Universität Erlangen–Nürnberg(应用力学研究所,弗赖堡-埃尔兰根-纽伦堡大学)

专题命中 具身导航 :robotics(abstract);分类 cs.RO

AI总结 本文提出了一种基于形状空间图的路径规划方法,用于软机器人在复杂环境中的快速且无碰撞导航,通过预计算形状库和多目标边成本实现高效路径生成。

Comments revised version

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.20612 2026-01-15 cs.LG 57%

Policy Compatible Skill Incremental Learning via Lazy Learning Interface

通过懒惰学习接口实现策略兼容的技能增量学习

Daehee Lee, Dongsu Lee, TaeYoon Kwack, Wonje Choi, Honguk Woo

机构 * Sungkyunkwan University(成均馆大学) University of Texas at Austin(德克萨斯大学奥斯汀分校)

专题命中 具身导航 :embodied agent(abstract);分类 cs.LG

AI总结 本文提出SIL-C框架,通过懒惰学习接口实现技能与策略的兼容性,提升下游任务性能无需重新训练策略。

Comments NeurIPS 2025 Spotlight

详情

展开后加载摘要…

URL PDF HTML 收藏
2601.09230 2026-01-15 cs.CV 57%

CLIDD: Cross-Layer Independent Deformable Description for Efficient and Discriminative Local Feature Representation

CLIDD: 跨层独立可变形描述用于高效且判别性的局部特征表示

Haodi Yao, Fenghua He, Ning Hao, Yao Su

机构 * School of Astronautics, Harbin Institute of Technology, China(航天学院,哈尔滨工业大学,中国) State Key Laboratory of General Artificial Intelligence, Beijing Institute for General Artificial Intelligence (BIGAI), China(通用人工智能国家重点实验室,北京通用人工智能研究院(BIGAI),中国)

专题命中 具身导航 :navigation(abstract);分类 cs.CV

AI总结 CLIDD通过跨层独立可变形描述实现高效且判别性的局部特征表示,提供高精度匹配与低计算开销的解决方案。

详情

展开后加载摘要…

URL PDF HTML 收藏
2507.23459 2026-01-15 cs.IR cs.AI 57%

KLAN: Kuaishou Landing-page Adaptive Navigator

快手落地页自适应导航

Fan Li, Chang Meng, Jiaqi Fu, Shuchang Liu, Jiashuo Zhang, Tianke Zhang, Xueliang Wang, Xiaoqiang Feng

机构 * Duke University(杜克大学)

专题命中 具身导航 :navigation(abstract);分类 cs.AI

AI总结 KLAN通过分层框架实现个性化落地页导航,提升用户活跃度和满意度。

Comments We propose PLPM, a new task for selecting optimal landing pages upon user entry. Our solution, KLAN, models static and dynamic user interests and is successfully deployed on Kuaishou, improving DAU and user lifetime

详情

展开后加载摘要…

URL PDF HTML 收藏
2601.09632 2026-01-15 cs.HC 50%

Perceptually-Guided Adjusted Teleporting: Perceptual Thresholds for Teleport Displacements in Virtual Environments

基于感知的调整式传送:虚拟环境中传送位移的感知阈值

Rose Connolly, Victor Zordan, Rachel McDonnell

专题命中 具身导航 :navigation(abstract)

AI总结 本研究探讨了虚拟环境中传送位移的感知阈值,发现传送目的地可被无察觉调整,且向后调整容忍度更高,为适应性VR移动系统提供了新机遇。

Comments 9 pages. to be published in IEEE VR conference proceedings 2026

详情

展开后加载摘要…

URL PDF HTML 收藏
2601.09359 2026-01-15 cond-mat.mtrl-sci 50%

Field report from Collaborative Research Center 1625: Heterogeneous research data management using ontology representations

协同研究中心1625的现场报告:使用本体表示进行异构研究数据管理

Doaa Mohamed, Samuel García Vázquez, Behnam Mardani, Victor Dudarev, Alfred Ludwig, Maribel Acosta, Markus Stricker

专题命中 具身导航 :navigation(abstract)

AI总结 本研究提出了一种基于本体表示的异构研究数据管理系统,用于管理多学科复杂材料数据,通过知识图谱和数据库结合实现数据的有效整合与应用。

Comments 22 pages, 11 figures

详情

展开后加载摘要…

URL PDF HTML 收藏

3. 具身推理 1 篇

2601.09452 2026-01-15 cs.CV 79%

MAD: Motion Appearance Decoupling for efficient Driving World Models

MAD:用于高效驾驶世界模型的运动外观解耦

Ahmad Rahimi, Valentin Gerard, Eloi Zablocki, Matthieu Cord, Alexandre Alahi

机构 * Sorbonne Université(索邦大学)

专题命中 具身推理 :world model(title,abstract);分类 cs.CV

AI总结 MAD通过解耦运动学习与外观合成,高效地将通用视频扩散模型转化为可控的驾驶世界模型,实现低计算成本和高性能

详情

展开后加载摘要…

URL PDF HTML 收藏

4. 模仿学习与强化学习 3 篇

2509.25756 2026-01-15 cs.RO cs.LG 73%

SAC Flow: Sample-Efficient Reinforcement Learning of Flow-Based Policies via Velocity-Reparameterized Sequential Modeling

SAC Flow: 通过速度重参数化序列建模高效训练基于流的策略

Yixian Zhang, Shu'ang Yu, Tonghe Zhang, Mo Guang, Haojia Hui, Kaiwen Long, Yu Wang, Chao Yu, Wenbo Ding

机构 * Tsinghua University(清华大学) Carnegie Mellon University(卡内基梅隆大学) Li Auto(利汽车) Shanghai AI Laboratory(上海人工智能实验室)

专题命中 模仿学习与强化学习 :manipulation(abstract);robotic(abstract);分类 cs.RO、cs.LG

AI总结 SAC Flow通过速度重参数化序列建模,高效训练基于流的策略,实现连续控制和机器人操作的高性能表现。

详情

展开后加载摘要…

URL PDF HTML 收藏
2505.19789 2026-01-15 cs.LG 57%

What Can RL Bring to VLA Generalization? An Empirical Study

RL能为VLA泛化带来什么?一项实证研究

Jijia Liu, Feng Gao, Bingwen Wei, Xinlei Chen, Qingmin Liao, Yi Wu, Chao Yu, Yu Wang

机构 * Shenzhen International Graduate School, Tsinghua University(清华大学深圳国际研究生院) Institute for Interdisciplinary Information Sciences, Tsinghua University(清华大学交叉信息研究院) Department of Electronic Engineering, Tsinghua University(清华大学电子工程系)

专题命中 模仿学习与强化学习 :embodied AI(abstract);分类 cs.LG

AI总结 本研究通过实证分析发现,PPO在提升VLA的语义理解和执行鲁棒性方面优于SFT,同时保持视觉鲁棒性,为VLA泛化提供了有效的方法。

Comments Accepted by NeurIPS 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2601.09215 2026-01-15 cs.CL 50%

UserLM-R1: Modeling Human Reasoning in User Language Models with Multi-Reward Reinforcement Learning

UserLM-R1: 通过多奖励强化学习建模人类推理在用户语言模型中的应用

Feng Zhang, Shijia Li, Chunmao Zhang, Zhanyu Ma, Jun Xu, Jiuchong Gao, Jinghua Hao, Renqing He, Jingwen Xu, Han Liu

机构 * Meituan(美团) Peking University(北京大学) Beijing University of Posts and Telecommunications(北京邮电大学) University of Chinese Academy of Sciences(中国科学院大学) Dalian University of Technology(大连理工大学)

专题命中 模仿学习与强化学习 :manipulation(abstract)

AI总结 UserLM-R1通过多奖励强化学习和动态目标驱动策略,提升用户语言模型在跨领域和对抗性场景中的推理与策略能力。

详情

展开后加载摘要…

URL PDF HTML 收藏

5. 机器人数据与评测 8 篇

2601.09031 2026-01-15 cs.RO cs.AI 84%

Generalizable Geometric Prior and Recurrent Spiking Feature Learning for Humanoid Robot Manipulation

可推广的几何先验与递归脉冲特征学习用于双足机器人操作

Xuetao Li, Wenke Huang, Mang Ye, Jifeng Xuan, Bo Du, Sheng Liu, Miao Li

机构 * School of Computer Science, School of Robotics, Institute of Technological Sciences, Wuhan University(计算机学院、机器人学院、技术科学研究院、武汉大学)

专题命中 机器人数据与评测 :manipulation(title,abstract);robotic(abstract);分类 cs.RO、cs.AI

AI总结 本文提出RGMP-S方法,通过几何先验和递归脉冲网络提升双足机器人操作的泛化能力与数据效率。

详情

展开后加载摘要…

URL PDF HTML 收藏
2601.08876 2026-01-15 cs.CV 83%

The Semantic Lifecycle in Embodied AI: Acquisition, Representation and Storage via Foundation Models

具身AI中的语义生命周期:通过基础模型实现获取、表示与存储

Shuai Chen, Hao Chen, Yuanchen Bei, Tianyang Zhao, Zhibo Zhou, Feiran Huang

机构 * College of Information Science and Technology, Jinan University(信息科学与技术学院,暨南大学) Faculty of Data Science, City University of Macau(数据科学学院,澳门城市大学) University of Illinois Urbana-Champaign(伊利诺伊大学厄巴纳-香槟分校) Zhongguancun Laboratory(中关村实验室) College of Cyber Security, Jinan University(网络安全学院,暨南大学)

专题命中 机器人数据与评测 :embodied AI(title,abstract);embodied agent(abstract);分类 cs.CV

AI总结 本文提出语义生命周期框架,通过基础模型在具身AI中实现语义信息的获取、表示与存储,探讨了语义处理的连续流动与维护,并总结了当前挑战与未来研究方向。

详情

展开后加载摘要…

URL PDF HTML 收藏
2601.09163 2026-01-15 cs.RO 70%

CEI: A Unified Interface for Cross-Embodiment Visuomotor Policy Learning in 3D Space

CEI:一种用于3D空间跨身体视觉运动策略学习的统一接口

Tong Wu, Shoujie Li, Junhao Gong, Changqing Guo, Xingting Li, Shilong Mu, Wenbo Ding

机构 * Shenzhen International Graduate School, Tsinghua University(清华大学深圳国际研究生院)

专题命中 机器人数据与评测 :manipulation(abstract);robotic(abstract);分类 cs.RO

AI总结 CEI提出了一种跨身体视觉运动策略学习框架,通过功能相似性量化和梯度优化实现不同机械臂和末端执行器间的演示转移,实验显示在模拟和现实任务中均取得高转移效率。

详情

展开后加载摘要…

URL PDF HTML 收藏