arXivDaily arXiv每日学术速递 周一至周五更新

视觉与机器人

机器人 / 具身智能

机器人、具身智能、机器人学习、操作、导航和具身世界模型。

共收录 14074 信号源:cs.RO, cs.AI, cs.CV, cs.LG

1. 具身导航 14074 篇

2605.17327 2026-05-19 cs.RO cs.AI cs.CV 67%

Efficient Feature-Free Initialization for Monocular Visual-Inertial Systems Using a Feed-Forward 3D Model

为单目视觉-惯性系统使用前馈3D模型实现高效的特征-free初始化

Yuantai Zhang, Jiaqi Yang, Huajian Zeng, Changhao Chen, Haoang Li, Liang Li, Dezhen Song, Xingxing Zuo

机构 * MBZUAI(马克斯·普朗克人工智能研究所) HKUST (GZ)(香港科技大学(广州)) Zhejiang University(浙江大学)

专题命中 具身导航 :navigation(abstract);分类 cs.RO、cs.AI、cs.CV

AI总结 本文提出了一种无需视觉特征跟踪的初始化框架,利用前馈3D模型预测的点云,从而提高了单目视觉-惯性导航系统的初始化可靠性与效率,实验表明其初始化成功率超过90%且数据需求显著减少。

详情

展开后加载摘要…

URL PDF HTML 收藏
2605.16391 2026-05-19 eess.SP cs.AI cs.LG cs.RO 67%

Overcoming the Intrinsic Performance Limitations of MEMS IMU via Diffusion-Based Generative Learning

通过扩散生成学习克服MEMS惯性测量单元的固有性能限制

Jiarui Lv, Feng Zhu, Xiaohong Zhang

机构 * School of Geodesy and Geomatics, Wuhan University(武汉大学测绘学院) Hubei Luojia Laboratory, Wuhan University(湖北珞珈实验室) Chinese Antarctic Center of Surveying and Mapping, Wuhan University(中国极地测绘南极科考中心)

专题命中 具身导航 :navigation(abstract);分类 cs.RO、cs.AI、cs.LG

AI总结 本文提出基于扩散的生成学习框架,利用低成本IMU数据生成高保真虚拟IMU数据,提升定位和姿态估计性能,并在空中测绘中验证了其有效性。

详情

展开后加载摘要…

URL PDF HTML 收藏
2605.12674 2026-05-14 cs.AI cs.LG cs.RO 67%

Revealing Interpretable Failure Modes of VLMs

揭示视觉语言模型的可解释性故障模式

Isha Chaudhary, Vedaant V Jain, Kavya Sachdeva, Sayan Ranu, Gagandeep Singh

机构 * UIUC(伊利诺伊大学香槟分校) Kumo AI IIT Delhi(德里印度理工学院)

专题命中 具身导航 :robotics(abstract);分类 cs.RO、cs.AI、cs.LG

AI总结 本文提出REVELIO框架,通过系统性探索揭示VLMs在自动驾驶和室内机器人领域中的可解释性故障模式,发现模型在空间定位和安全威胁识别上的缺陷。

详情

展开后加载摘要…

URL PDF HTML 收藏
2605.08086 2026-05-12 cs.GR 67%

Representations of 3D Rotations: Mathematical Foundations and Comparative Analysis

三维旋转的表示:数学基础与比较分析

Aizierjiang Aiersilan, Haochen Liu, James Hahn

专题命中 具身导航 :robotics(abstract);navigation(abstract)

AI总结 本文比较了SO(3)的不同表示方法,分析其数学性质、连续性、计算效率等,指出四元数在紧凑性和效率上的优势,同时探讨了新兴的连续概率方法。

Comments Keywords: Rotation representations, $SO(3)$, quaternions, continuity, gimbal lock, 3D shape registration

详情

展开后加载摘要…

URL PDF HTML 收藏
2605.05717 2026-05-08 eess.SY cs.SY 67%

Space-Time Diversity in Observability and Estimation on Product Lie Groups

时空多样性在乘积李群上的可观测性与估计

Somasundhar Venkatasubramanian, Anirudh Venkat, Advaidh Venkat

专题命中 具身导航 :robotics(abstract);navigation(abstract)

AI总结 本文研究了在乘积李群上动态状态估计中的时空多样性,提出耦合条件、空间多样性饱和定理和时空多样性分解,为冗余和非冗余传感器架构提供精确可观测性保证。

Comments 6 Pages (two columns), 1 figure 2 tables and an alogorithm. This work has been submitted to the IEEE for possible publication

详情

展开后加载摘要…

URL PDF HTML 收藏
2603.28032 2026-04-23 cs.RO cs.AI cs.CV cs.HC 67%

CARLA-Air: Fly Drones Inside a CARLA World -- A Unified Infrastructure for Air-Ground Embodied Intelligence

CARLA-Air:在CARLA世界中飞行无人机——一个统一的空地具身智能基础设施

Tianle Zeng, Yanci Wen, Hong Zhang

机构 * Hong Zhang 1,†(香港理工大学)

专题命中 具身导航 :navigation(abstract);分类 cs.RO、cs.AI、cs.CV

AI总结 CARLA-Air通过统一高保真城市驾驶与物理准确多旋翼飞行,解决了空地协同模拟中的同步与一致性问题,支持多种具身智能任务。

Comments Prebuilt binaries, project page, full source code, and community discussion group are all available at: https://github.com/louiszengCN/CarlaAir

详情

展开后加载摘要…

URL PDF HTML 收藏
2602.06400 2026-04-22 cs.CV cs.AI cs.RO 67%

TFusionOcc: T-Primitive Based Object-Centric Multi-Sensor Fusion Framework for 3D Occupancy Prediction

TFusionOcc:基于T-原语的对象中心多传感器融合框架用于3D占用预测

Zhenxing Ming, Yaoqi Huang, Julie Stephany Berrio, Mao Shan, Stewart Worrall

机构 * Australian Centre for Robotics(澳大利亚机器人中心)

专题命中 具身导航 :navigation(abstract);分类 cs.RO、cs.AI、cs.CV

AI总结 本文提出TFusionOcc框架,通过T-原语实现3D语义占用预测,结合概率模型和多阶段融合架构,实验显示其在nuScenes数据集上表现优异。

详情

展开后加载摘要…

URL PDF HTML 收藏
2604.11538 2026-04-14 cs.HC 67%

ResearchCube: Multi-Dimensional Trade-off Exploration for Research Ideation

ResearchCube: 多维权衡探索用于研究构想

Zijian Ding, Fenghai Li, Ziyi Wang, Joel Chan

专题命中 具身导航 :manipulation(abstract);navigation(abstract)

AI总结 ResearchCube通过多维权衡谱帮助研究者探索和优化构想,提供三维空间交互,提升认知和控制感。

详情

展开后加载摘要…

URL PDF HTML 收藏
2604.03277 2026-04-07 cs.CV cs.AI cs.LG 67%

Event-Driven Neuromorphic Vision Enables Energy-Efficient Visual Place Recognition

基于事件驱动的类脑视觉实现高效视觉位置识别

Geoffroy Keime, Nicolas Cuperlier, Benoit R. Cottereau

机构 * IPAL, CNRS IRL 2955, Singapore(IPAL,CNRS IRL 2955,新加坡) Univ Toulouse, CNRS, CerCo UMR 5549, Toulouse, France(图卢兹大学,CNRS,CerCo UMR 5549,法国图卢兹) Laboratoire ETIS UMR 8051, CY Cergy-Paris Université, ENSEA, CNRS, Cergy, France(ETIS实验室 UMR 8051,CY塞尔吉-巴黎大学,ENSEA,CNRS,法国塞尔吉)

专题命中 具身导航 :navigation(abstract);分类 cs.AI、cs.CV、cs.LG

AI总结 本文提出SpikeVPR,结合事件相机与脉冲神经网络,实现高效鲁棒的视觉位置识别,在动态环境下使用更少参数和能量。

Comments 40 pages single column, v1

详情

展开后加载摘要…

URL PDF HTML 收藏
2303.09744 2026-04-02 cs.MA 67%

Inferring Occluded Agent Behavior in Dynamic Games from Noise Corrupted Observations

从噪声受损观测中推断动态博弈中被遮挡行为

Tianyu Qiu, David Fridovich-Keil

专题命中 具身导航 :robotics(abstract);navigation(abstract)

AI总结 本文提出一种考虑遮挡的博弈推理方法,用于估计可能被遮挡的智能体位置,并推断可见和被遮挡智能体的意图,同时验证了在噪声观测下的导航安全性。

详情

展开后加载摘要…

URL PDF HTML 收藏
2511.20823 2026-03-13 cs.CV cs.AI cs.LG 67%

RefTr: Recurrent Refinement of Confluent Trajectories for 3D Vascular Tree Centerlines

RefTr: 基于连通轨迹的3D血管树中心线递归细化

Roman Naeem, David Hagerman, Jennifer Alvén, Fredrik Kahl

专题命中 具身导航 :navigation(abstract);分类 cs.AI、cs.CV、cs.LG

AI总结 RefTr通过递归细化连通轨迹生成3D血管树中心线,采用Transformer架构提升精度并减少参数,实验显示其在性能、速度和参数数量上均优于现有方法。

详情

展开后加载摘要…

URL PDF HTML 收藏
2601.03869 2026-03-09 cs.CV cs.GR cs.LG cs.RO 67%

Bayesian Monocular Depth Refinement via Neural Radiance Fields

基于神经辐射场的贝叶斯单目深度细化

Arun Muthukkumar

机构 * Department of Computer Science Illinois Mathematics Science Academy Aurora, United States

专题命中 具身导航 :navigation(abstract);分类 cs.RO、cs.CV、cs.LG

AI总结 本文提出MDENeRF,通过神经辐射场和贝叶斯融合改进单目深度估计,提升场景理解的精细几何细节。

Comments IEEE 8th International Conference on Algorithms, Computing and Artificial Intelligence (ACAI 2025)

Journal ref Proc. IEEE 8th International Conference on Algorithms, Computing and Artificial Intelligence (ACAI), pp. 488-492, 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2503.13568 2026-03-03 cs.RO cs.AI cs.LG 67%

WMINet: A Wheel-Mounted Inertial Learning Approach For Mobile-Robot Positioning

WMINet: 一种轮式惯性学习方法用于移动机器人定位

Gal Versano, Itzik Klein

机构 * The Hatter Department of Marine Technologies, Charney School of Marine Sciences, University of Haifa(哈伊夫大学海洋技术系、查内海洋科学学院)

专题命中 具身导航 :navigation(abstract);分类 cs.RO、cs.AI、cs.LG

AI总结 WMINet通过结合轮式方法和周期性轨迹运动,利用深度学习减少惯性漂移,实现移动机器人仅依赖惯性传感器的高精度定位。

详情

展开后加载摘要…

URL PDF HTML 收藏
2602.09082 2026-02-25 cs.CV cs.AI cs.CL cs.LG 67%

UI-Venus-1.5 Technical Report

UI-Venus-1.5 技术报告

Venus Team, Changlong Gao, Zhangxuan Gu, Yulin Liu, Xinyu Qiu, Shuheng Shen, Yue Wen, Tianyu Xia, Zhenyu Xu, Zhengwen Zeng, Beitong Zhou, Xingran Zhou, Weizhi Chen, Sunhao Dai, Jingya Dou, Yichen Gong, Yuan Guo, Zhenlin Guo, Feng Li, Qian Li, Jinzhen Lin, Yuqi Zhou, Linchao Zhu, Liang Chen, Zhenyu Guo, Changhua Meng, Weiqiang Wang

机构 * Ant Group(蚂蚁集团)

专题命中 具身导航 :navigation(abstract);分类 cs.AI、cs.CV、cs.LG

AI总结 UI-Venus-1.5通过统一的GUI代理和关键技术改进,在多个基准测试中实现了最先进的性能,展示了在现实世界应用中的强大导航能力。

详情

展开后加载摘要…

URL PDF HTML 收藏
2505.11409 2026-02-23 cs.LG cs.AI cs.CL cs.CV 67%

Visual Planning: Let's Think Only with Images

视觉规划:只用图像来思考

Yi Xu, Chengzu Li, Han Zhou, Xingchen Wan, Caiqi Zhang, Anna Korhonen, Ivan Vulić

机构 * Language Technology Lab, University of Cambridge(剑桥大学语言技术实验室) Google(谷歌公司)

专题命中 具身导航 :navigation(abstract);分类 cs.AI、cs.CV、cs.LG

AI总结 本文提出视觉规划范式,通过图像进行推理,优于纯文本推理,在视觉导航任务中表现更优。

Comments ICLR 2026 (Oral)

详情

展开后加载摘要…

URL PDF HTML 收藏
2602.13972 2026-02-17 physics.optics 67%

Polarization-Multiplexed Chaotic LiDAR Based on a VCSEL with Delayed Orthogonal Feedback

基于VCSEL的极化多路复用混沌LiDAR

T. Wang, Z. Li, H. Shen, Y. Ma, Y. Li, S. Xiang, S. Baland, Y. Hao

专题命中 具身导航 :robotics(abstract);navigation(abstract)

AI总结 本文提出基于VCSEL的极化多路复用混沌LiDAR系统,通过延迟正交极化反馈实现高精度测距,具备可调谐和抗干扰优势,适用于自动驾驶等应用。

详情

展开后加载摘要…

URL PDF HTML 收藏
2602.13327 2026-02-17 physics.gen-ph 67%

Uncertainty in space, time, and motion on the special Galilean group

时空与运动中的不确定性在特殊伽利略群中

Jonathan Kelly, Matthew Giamou

专题命中 具身导航 :robotics(abstract);navigation(abstract)

AI总结 本文提出在伽利略群上直接表达和传播不确定性的方法,通过闭式雅可比表达式提升时空估计的统计一致性。

Comments 21 pages, 3 figures. Submitted to Philosophical Transactions of the Royal Society A

详情

展开后加载摘要…

URL PDF HTML 收藏
2406.09260 2026-02-03 cs.CV cs.AI cs.RO eess.IV 67%

Deep Transformer Network for Monocular Pose Estimation of Shipborne Unmanned Aerial Vehicle

用于船载无人机的深度变换器网络姿态估计

Maneesha Wickramasuriya, Taeyoung Lee, Murray Snyder

专题命中 具身导航 :navigation(abstract);分类 cs.RO、cs.AI、cs.CV

AI总结 本文提出了一种基于深度变换器网络的单目图像姿态估计方法,用于船载无人机的6D姿态估计,通过贝叶斯融合提升精度,并在合成数据和实地实验中验证了其鲁棒性和准确性。

Comments 23 pages, 25 figures, 3 tables

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.10029 2026-01-28 eess.SY cs.SY 67%

Ruggedized Ultrasound Sensing in Harsh Conditions: eRTIS in the wild

在恶劣条件下增强的超声传感:eRTIS在野外的应用

Dennis Laurijssen, Wouter Jansen, Arne Aerts, Walter Daems, Jan Steckel

专题命中 具身导航 :robotics(abstract);navigation(abstract)

AI总结 eRTIS是一种用于恶劣工业环境的坚固超声传感系统,通过模块化硬件架构和GPU加速信号处理,在复杂环境中实现稳健的传感性能。

详情

展开后加载摘要…

URL PDF HTML 收藏
2309.14025 2026-01-26 physics.flu-dyn 67%

Optimum control strategies for maximum thrust production in underwater undulatory swimming

在水下摆动游泳中实现最大推力生产的最优控制策略

L. fu, S. Israilov, J. Sanchez Rodriguez, C. Brouzet, G. Allibert, C. Raufaste, M. Argentina

专题命中 具身导航 :robotics(abstract);robotic(abstract)

AI总结 本文通过仿生机器人研究了水下游泳中推力生成的最优控制策略,结合机器学习和直观模型,发现最优尾部拍打频率和幅度,为自主游泳机器人提供实用方案。

Journal ref Phys. Rev. Fluids 10, 043101 (2025)

详情

展开后加载摘要…

URL PDF HTML 收藏
2601.15578 2026-01-23 cs.NI cs.AI cs.LG cs.RO 67%

MapViT: A Two-Stage ViT-Based Framework for Real-Time Radio Quality Map Prediction in Dynamic Environments

MapViT:一种基于视觉Transformer的两阶段框架,用于动态环境中实时无线电质量地图预测

Cyril Shih-Huan Hsu, Xi Li, Lanfranco Zanzi, Zhiheng Yang, Chrysa Papagianni, Xavier Costa Pérez

机构 * Informatics Institute, University of Amsterdam(阿姆斯特丹大学信息学院) NEC Laboratories Europe(NEC欧洲实验室) i2CAT Foundation(i2CAT基金会) Catalan Institution for Research and Advanced Studies (ICREA)(加泰罗尼亚研究与高级研究机构(ICREA))

专题命中 具身导航 :robotic(abstract);分类 cs.RO、cs.AI、cs.LG

AI总结 MapViT通过两阶段ViT框架实现实时动态环境中无线信号质量预测,结合预训练和微调方法提升准确性和效率,适用于资源受限的移动机器人场景。

Comments This paper has been accepted for publication at IEEE International Conference on Communications (ICC) 2026

详情

展开后加载摘要…

URL PDF HTML 收藏
2503.10331 2026-01-22 cs.CV cs.AI cs.CL cs.RO 67%

OSMa-Bench: Evaluating Open Semantic Mapping Under Varying Lighting Conditions

OSMa-Bench: 评估开放语义映射在不同光照条件下的性能

Maxim Popov, Regina Kurkova, Mikhail Iumanov, Jaafar Mahmoud, Sergey Kolyubin

机构 * Biomechatronics and Energy-Efficient Robotics (BE2R) Lab, ITMO University(生物机电学与节能机器人实验室,ITMO大学)

专题命中 具身导航 :robotic(abstract);分类 cs.RO、cs.AI、cs.CV

AI总结 OSMa-Bench通过评估不同光照条件下语义映射算法的性能,推动了更鲁棒的机器人系统发展。

Comments Project page: https://be2rlab.github.io/OSMa-Bench/

Journal ref 2025 IEEE/RSJ International Conference on Intelligent Robots and Systems (IROS), Hangzhou, China, 2025, pp. 10786-10791

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.25482 2026-01-21 cs.AI cs.LG cs.RO cs.SY eess.SY stat.ML 67%

Message passing-based inference in an autoregressive active inference agent

基于消息传递的自回归主动推断代理推理

Wouter M. Kouw, Tim N. Nisslbeck, Wouter L. N. Nuijten

机构 * Eindhoven University of Technology(埃因霍温理工大学) Lazy Dynamics B.V.(Lazy Dynamics公司)

专题命中 具身导航 :navigation(abstract);分类 cs.RO、cs.AI、cs.LG

AI总结 本文提出一种基于消息传递的自回归主动推断代理,通过在因子图上进行推理,在机器人导航任务中展示了在连续观测空间中实现探索与利用的能力,并在动态建模方面优于传统控制器。

Comments 14 pages, 4 figures, proceedings of the International Workshop on Active Inference 2025. Erratum v1: in Eq. (50), $p(y_t, Θ, u_t \mid y_{*}, \mathcal{D}_k)$ should have been $p(y_t, Θ\mid u_t, y_{*}, \mathcal{D}_k)$

详情

展开后加载摘要…

URL PDF HTML 收藏
2601.08229 2026-01-14 cs.CR 67%

A Survey of Security Challenges and Solutions for UAS Traffic Management (UTM) and small Unmanned Aerial Systems (sUAS)

无人机交通管理(UTM)和小型无人驾驶航空系统(sUAS)的安全挑战与解决方案综述

Iman Sharifi, Mahyar Ghazanfari, Abenezer Taye, Peng Wei, Maheed H. Ahmed, Hyeong Tae Kim, Mahsa Ghasemi, Vijay Gupta, Noah Dahle, Robert Canady, Abel Diaz Gonzalez, Austin Coursey, Bryce Bjorkman, Cailani Lemieux-Mack, Bryan C. Ward, Xenofon Koutsoukos, Gautam Biswas, Heber Herencia-Zapana, Saqib Hasan, Isaac Amundson, Filippos Fotiadis, Ufuk Topcu, Junchi Lu, Qi Alfred Chen, Nischal Aryal, Amer Ibrahim, Abdul Karim Ras, Amir Shirkhodaie

专题命中 具身导航 :manipulation(abstract);navigation(abstract)

AI总结 本文综述了sUAS和UTM生态系统的网络安全漏洞与防御措施,旨在建立统一的分类法并识别未来UTM环境中的安全挑战。

Comments 26 pages, 3 figures, 5 tables

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.17117 2026-01-05 cond-mat.soft 67%

Digitization Can Stall Swarm Transport: Commensurability Locking in Quantized-Sensing Chains

数字化可能阻碍群体运输:量化感知链中的公约数锁定

Caroline N. Cappetto, Penelope Messinger, Kaitlyn S. Yasumura, Miro Rothman, Tuan K. Do, Gao Wang, Liyu Liu, Robert H. Austin, Shengkai Li, Trung V. Phan

专题命中 具身导航 :robotics(abstract);robotic(abstract)

AI总结 研究发现机器人群体在均匀梯度下因传感器分离与间距比例满足数论条件而出现集体运输停滞现象,揭示了数论与群体机器人运输之间的联系。

详情

展开后加载摘要…

URL PDF HTML 收藏
2503.07020 2025-12-15 cs.RO cs.AI cs.LG 67%

Driving Through Uncertainty: Risk-Averse Control with LLM Commonsense for Autonomous Driving under Perception Deficits

在不确定性中驾驶:利用大语言模型常识实现风险厌恶控制以应对自动驾驶中的感知缺陷

Yuting Hu, Chenhui Xu, Ruiyang Qin, Dancheng Liu, Amir Nassereldine, Yiyu Shi, Jinjun Xiong

机构 * University at Buffalo(布法罗大学) University of Notre Dame(诺特尔大学)

专题命中 具身导航 :navigation(abstract);分类 cs.RO、cs.AI、cs.LG

AI总结 本文提出LLM-RCO框架,利用大语言模型整合人类驾驶常识,以提升自动驾驶在感知缺陷下的风险规避与适应性控制能力。

详情

展开后加载摘要…

URL PDF HTML 收藏
2512.09898 2025-12-11 cs.RO cs.AI cs.CV cs.MA cs.SY eess.SY 67%

Visual Heading Prediction for Autonomous Aerial Vehicles

自主航空器的视觉航向预测

Reza Ahmari, Ahmad Mohammadi, Vahid Hemmati, Mohammed Mynuddin, Parham Kebria, Mahmoud Nabil Mahmoud, Xiaohong Yuan, Abdollah Homaifar

机构 * Department of Computer Science at North Carolina A&T State University(北卡罗来纳A&T州立大学计算机科学系) Department of Electrical and Computer Engineering at North Carolina A&T State University(北卡罗来纳A&T州立大学电气与计算机工程系)

专题命中 具身导航 :navigation(abstract);分类 cs.RO、cs.AI、cs.CV

AI总结 本文提出基于视觉的数据驱动框架,实现无人机与无人地面车辆的实时整合,通过YOLOv5检测UGV并利用轻量级ANN预测航向角,实现高精度的导航与协调。

详情

展开后加载摘要…

URL PDF HTML 收藏
2511.19768 2025-11-26 cs.CV cs.AI cs.RO 67%

Prune-Then-Plan: Step-Level Calibration for Stable Frontier Exploration in Embodied Question Answering

剪枝后再规划:用于具身问答中稳定前沿探索的分步校准

Noah Frahm, Prakrut Patel, Yue Zhang, Shoubin Yu, Mohit Bansal, Roni Sengupta

机构 * University of North Carolina at Chapel Hill(北卡罗来纳大学教堂山分校)

专题命中 具身导航 :navigation(abstract);分类 cs.RO、cs.AI、cs.CV

AI总结 Prune-Then-Plan通过分步校准稳定具身问答中的前沿探索,提升导航效率和回答质量。

Comments webpage: https://noahfrahm.github.io/Prune-Then-Plan-project-page/

详情

展开后加载摘要…

URL PDF HTML 收藏
2504.08057 2025-11-21 cs.NE cs.AI cs.LG cs.RO 67%

Vector Quantized-Elites: Unsupervised and Problem-Agnostic Quality-Diversity Optimization

向量量化-精英:无监督和问题无关的质量-多样性优化

Constantinos Tsakonas, Konstantinos Chatzilygeroudis

机构 * Computational Intelligence Laboratory (CILab), Department of Mathematics, University of Patras(数学系计算智能实验室(CILab),帕特拉大学) Laboratory of Automation and Robotics (LAR) in the Department of Electrical & Computer Engineering, University of Patras(电气与计算机工程系自动化与机器人实验室(LAR),帕特拉大学)

专题命中 具身导航 :robotic(abstract);分类 cs.RO、cs.AI、cs.LG

AI总结 VQ-Elites通过无监督学习构建结构化行为空间,实现灵活且任务无关的质量-多样性优化。

Comments 15 pages (+4 supplementary), 14 (+1) figures, 1 algorithm, 1 (+8) table(s), accepted at IEEE Transactions on Evolutionary Computation

详情

展开后加载摘要…

URL PDF HTML 收藏
2511.11169 2025-11-17 cs.CV cs.AI cs.LG 67%

Refine and Align: Confidence Calibration through Multi-Agent Interaction in VQA

Ayush Pandey, Jai Bardhan, Ishita Jain, Ramya S Hebbalaguppe, Rohan Raju Dhanakshirur, Lovekesh Vig

机构 * TCS Research(塔塔咨询研究)

专题命中 具身导航 :navigation(abstract);分类 cs.AI、cs.CV、cs.LG

Comments 17 pages, 6 figures, 5 tables. Accepted to Special Track on AI Alignment, AAAI 2026. Project Page- https://refine-align.github.io/

详情

展开后加载摘要…

URL PDF HTML 收藏