arXivDaily arXiv每日学术速递 周一至周五更新
arXiv周末暂无论文更新,休息一下吧,周末愉快~~

视觉与机器人

机器人 / 具身智能

机器人、具身智能、机器人学习、操作、导航和具身世界模型。

共收录 14074 信号源:cs.RO, cs.AI, cs.CV, cs.LG

1. 具身导航 14074 篇

2509.17287 2026-03-10 cs.RO cs.CV 62%

Event-Based Visual Teach-and-Repeat via Fast Fourier-Domain Cross-Correlation

基于快速傅里叶域交叉相关性的事件视觉教与重复

Gokul B. Nair, Alejandro Fontan, Michael Milford, Tobias Fischer

机构 * Queensland University of Technology(昆士兰理工大学)

专题命中 具身导航 :navigation(abstract);分类 cs.RO、cs.CV

AI总结 本文提出基于事件相机的快速VT&R导航系统,通过频域交叉相关技术实现高效事件流匹配,实验验证其在复杂环境下的高精度导航能力。

Comments 8 Pages, 5 Figures, Under Review

详情

展开后加载摘要…

URL PDF HTML 收藏
2603.05511 2026-03-09 cs.HC cs.AI cs.GR cs.RO 62%

An Embodied Companion for Visual Storytelling

具身化伴侣:视觉叙事中的交互伙伴

Patrick Tresset, Markus Wulfmeier

机构 * Goldsmiths, University of London(伦敦大学金史密斯学院) DeepMind(深度Mind)

专题命中 具身导航 :robotic(abstract);分类 cs.RO、cs.AI

AI总结 Companion通过结合绘图机器人与大语言模型,实现人机协同创作,推动视觉叙事的创新与美学探索。

Comments 35 pages, 18 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2601.14895 2026-03-09 cs.CV cs.AI 62%

SpatialMem: Metric-Aligned Long-Horizon Video Memory for Language Grounding and QA

SpatialMem: 用于语言定位和问答的度量对齐长周期视频记忆

Xinyi Zheng, Yunze Liu, Chi-Hao Wu, Fan Zhang, Hao Zheng, Wenqi Zhou, Walterio W. Mayol-Cuevas, Junxiao Shen

机构 * University of Bristol(布里斯托大学) Memories.ai Research(Memories.ai研究院)

专题命中 具身导航 :navigation(abstract);分类 cs.AI、cs.CV

AI总结 SpatialMem通过构建度量对齐的空间支架,实现长周期视频记忆的可解释检索与问答,支持语言引导的检索和离线导航任务。

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.20739 2026-03-09 cs.RO cs.CV 62%

Decision-Driven Semantic Object Exploration for Legged Robots via Confidence-Calibrated Perception and Topological Subgoal Selection

通过置信度校准感知与拓扑子目标选择实现决策驱动的语义对象探索

Guoyang Zhao, Yudong Li, Weiqing Qi, Kai Zhang, Bonan Liu, Kai Chen, Haoang Li, Jun Ma

机构 * Robotics and Autonomous Systems Thrust, The Hong Kong University of Science and Technology (Guangzhou), China(机器人与自主系统方向,香港科技大学(广州)) Department of Mechanical and Energy Engineering, Southern University of Science and Technology, China(机械与能源工程系,南方科技大学)

专题命中 具身导航 :navigation(abstract);分类 cs.RO、cs.CV

AI总结 本文提出了一种基于视觉的决策驱动语义对象探索方法,通过置信度校准感知与拓扑子目标选择机制提升腿部机器人在开放环境中的探索性能。

详情

展开后加载摘要…

URL PDF HTML 收藏
2603.04659 2026-03-06 cs.RO cs.AI 62%

GIANT - Global Path Integration and Attentive Graph Networks for Multi-Agent Trajectory Planning

GIANT - 全局路径整合与注意力图网络用于多智能体轨迹规划

Jonas le Fevre Sejersen, Toyotaro Suzumura, Erdal Kayacan

机构 * Artificial Intelligence in Robotics Laboratory (AiR Lab), Department of Electrical and Computer Engineering, Aarhus University(人工智能机器人实验室(AiR实验室),电气与计算机工程系,奥胡斯大学) Foundation Models for Artificial Intelligence group, Department of Information and Communication Engineering, Tokyo University(人工智能基础模型组,信息与通信工程系,东京大学) Automatic Control Group, Department of Electrical Engineering and Information Technology, Paderborn University(自动控制组,电气工程与信息科技系,波德恩大学)

专题命中 具身导航 :navigation(abstract);分类 cs.RO、cs.AI

AI总结 GIANT通过结合全局路径规划与注意力图网络,提升多智能体在复杂动态环境中的避障与导航性能。

Comments Published in: 2025 IEEE/RSJ International Conference on Intelligent Robots and Systems (IROS)

详情

展开后加载摘要…

URL PDF HTML 收藏
2603.04180 2026-03-05 cs.LG cs.AI 62%

Architectural Proprioception in State Space Models: Thermodynamic Training Induces Anticipatory Halt Detection

状态空间模型中的建筑本体感知:热力学训练诱导预见性停止检测

Jay Noon

专题命中 具身导航 :navigation(abstract);分类 cs.AI、cs.LG

AI总结 本研究提出通过热力学训练使状态空间模型具备预见性停止检测能力,揭示其在元认知和计算自我意识方面的独特优势。

Comments 17 pages, 15 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2603.03935 2026-03-05 cs.CV cs.RO 62%

DISC: Dense Integrated Semantic Context for Large-Scale Open-Set Semantic Mapping

DISC: 密集集成语义上下文用于大规模开放集语义映射

Felix Igelbrink, Lennart Niecksch, Martin Atzmueller, Joachim Hertzberg

机构 * German Research Center for Artificial Intelligence (DFKI)(德国人工智能研究中心)

专题命中 具身导航 :robotic(abstract);分类 cs.RO、cs.CV

AI总结 DISC通过密集集成语义上下文方法,提升大规模开放集语义映射的语义准确性和实时性,适用于复杂场景的机器人部署。

详情

展开后加载摘要…

URL PDF HTML 收藏
2506.03317 2026-03-05 physics.optics cs.CV cs.LG physics.app-ph 62%

Structural Vibration Monitoring with Diffractive Optical Processors

基于衍射光学处理器的结构振动监测

Yuntian Wang, Zafer Yilmaz, Yuhang Li, Edward Liu, Eric Ahlberg, Farid Ghahari, Ertugrul Taciroglu, Aydogan Ozcan

专题命中 具身导航 :navigation(abstract);分类 cs.CV、cs.LG

AI总结 本文提出一种基于衍射光学处理器的结构振动监测系统,通过联合优化的衍射层和浅层神经网络实现低功耗、高精度的三维振动谱提取,适用于结构健康监测及灾害韧性等应用。

Comments 33 Pages, 8 Figures, 1 Table

Journal ref Science Advances (2026)

详情

展开后加载摘要…

URL PDF HTML 收藏
2603.03546 2026-03-05 cs.RO cs.LG cs.SY eess.SY 62%

Real-time loosely coupled GNSS and IMU integration via Factor Graph Optimization

基于因子图优化的实时松散耦合GNSS与IMU融合

Radu-Andrei Cioaca, Cristian Rusu, Paul Irofti, Gianluca Caparra, Andrei-Alexandru Marinache, Florin Stoican

机构 * Three Tensors S.R.L.(Three Tensors公司) Navigation Systems Definition Section (TEC-SEN)(欧洲航天局导航系统定义部门) Romanian InSpace Engineering S.R.L. (RISE)(罗马尼亚InSpace工程公司)

专题命中 具身导航 :navigation(abstract);分类 cs.RO、cs.LG

AI总结 本文提出基于因子图优化的实时松散耦合GNSS与IMU融合方法,实现实时操作并提高服务可用性,但牺牲部分定位精度。

详情

展开后加载摘要…

URL PDF HTML 收藏
2603.03075 2026-03-04 cs.CV cs.AI cs.AR 62%

TinyIceNet: Low-Power SAR Sea Ice Segmentation for On-Board FPGA Inference

TinyIceNet:低功耗SAR海冰分割用于机载FPGA推断

Mhd Rashed Al Koutayni, Mohamed Selim, Gerd Reis, Alain Pagani, Didier Stricker

机构 * German Research Center for Artificial Intelligence, DFKI(德国人工智能研究中心,DFKI)

专题命中 具身导航 :navigation(abstract);分类 cs.AI、cs.CV

AI总结 TinyIceNet是一种专为机载FPGA设计的低功耗SAR海冰分割网络,通过架构简化和低精度量化实现高效准确的海冰制图。

Comments undergoing publication at CVC 2026

详情

展开后加载摘要…

URL PDF HTML 收藏
2603.02609 2026-03-04 cs.CV cs.RO 62%

VLMFusionOcc3D: VLM Assisted Multi-Modal 3D Semantic Occupancy Prediction

VLMFusionOcc3D: 基于VLM的多模态3D语义占位预测

A. Enes Doruk, Hasan F. Ates

机构 * Department of Artificial Intelligence and Data Engineering, Ozyegin University(人工智能与数据工程系,奥克辛大学)

专题命中 具身导航 :navigation(abstract);分类 cs.RO、cs.CV

AI总结 VLMFusionOcc3D通过融合视觉语言模型的语义先验,提升多模态3D语义占位预测的鲁棒性和适应性。

详情

展开后加载摘要…

URL PDF HTML 收藏
2602.19534 2026-03-03 cs.RO cs.AI 62%

Large Language Model-Assisted UAV Operations and Communications: A Multifaceted Survey and Tutorial

大型语言模型辅助的无人机操作与通信:多方面的综述与教程

Yousef Emami, Hao Zhou, Radha Reddy, Atefeh Hajijamali Arani, Biliang Wang, Kai Li, Luis Almeida, Zhu Han

机构 * IEEE

专题命中 具身导航 :navigation(abstract);分类 cs.RO、cs.AI

AI总结 本文综述了大型语言模型在无人机操作与通信中的应用,探讨了LLMs在提升UAV智能方面的多方面技术与未来研究方向。

Comments 40 pages, 10 figures, 13 tables

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.16607 2026-03-02 cs.LG cs.AI 62%

Asymptotically Stable Quaternion-valued Hopfield-structured Neural Network with Periodic Projection-based Supervised Learning Rules

渐近稳定的四元数值Hopfield结构神经网络及其基于周期投影的监督学习规则

Tianwei Wang, Xinhui Ma, Wei Pang

机构 * University of Edinburgh(爱丁堡大学) University of Hull(赫尔大学) Heriot-Watt University(赫瑞-沃德大学)

专题命中 具身导航 :robotic(abstract);分类 cs.AI、cs.LG

AI总结 本文提出了一种基于四元数的Hopfield结构神经网络,通过周期投影策略实现监督学习,具有高精度、快速收敛和强可靠性,适用于机器人控制等需要四元数参数化的场景。

Comments Preprint. Accepted by NeurIPS 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2602.23235 2026-02-27 cs.CV cs.AI 62%

Spatio-Temporal Token Pruning for Efficient High-Resolution GUI Agents

时空令牌剪枝用于高效高分辨率GUI代理

Zhou Xu, Bowen Zhou, Qi Wang, Shuwen Feng, Jingyu Xiao

机构 * Tsinghua University(清华大学) Xidian University(西安电子科技大学) The Chinese University of Hong Kong(香港中文大学)

专题命中 具身导航 :navigation(abstract);分类 cs.AI、cs.CV

AI总结 GUIPruner通过时空令牌剪枝技术,实现高分辨率GUI导航的高效处理,显著提升性能并降低资源消耗。

详情

展开后加载摘要…

URL PDF HTML 收藏
2602.22923 2026-02-27 cs.CV cs.RO 62%

WaterVideoQA: ASV-Centric Perception and Rule-Compliant Reasoning via Multi-Modal Agents

WaterVideoQA: 以ASV为中心的感知与符合规则的推理 via 多模态智能体

Runwei Guan, Shaofeng Liang, Ningwei Ouyang, Weichen Fei, Shanliang Yao, Wei Dai, Chenhao Ge, Penglei Sun, Xiaohui Zhu, Tao Huang, Ryan Wen Liu, Hui Xiong

机构 * Thrust of Artificial Intelligence, The Hong Kong University of Science and Technology (Guangzhou)(香港理工大学(广州)人工智能研究所) Hubei Key Laboratory of Inland Shipping Technology (Wuhan University of Technology)(湖北内河航运技术重点实验室(武汉理工大学)) School of Navigation, Wuhan University of Technology(武汉理工大学航海学院) School of Advanced Technology, Xi’an Jiaotong-Liverpool University(西安交通大学利物浦大学先进科技学院) School of Artificial Intelligence, Nanjing University(南京大学人工智能学院) School of Information Engineering, Yancheng Institute of Technology(盐城职业技术学院信息工程学院) School of Engineering, Stanford University(斯坦福大学工程学院) Centre for AI and Data Science Innovation and the School of Science and Engineering, James Cook University(詹姆斯库克大学人工智能与数据科学创新中心及科学与工程学院)

专题命中 具身导航 :navigation(abstract);分类 cs.RO、cs.CV

AI总结 WaterVideoQA通过多模态智能体系统,实现ASV在复杂水域环境中的感知与规则合规推理,提升自主航行的安全性和精确性。

Comments 11 pages,8 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2602.21904 2026-02-26 cs.CV cs.RO 62%

UNet-Based Keypoint Regression for 3D Cone Localization in Autonomous Racing

基于UNet的关键点回归用于自动驾驶赛车中3D圆锥定位

Mariia Baidachna, James Carty, Aidan Ferguson, Joseph Agrane, Varad Kulkarni, Aubrey Agub, Michael Baxendale, Aaron David, Rachel Horton, Elliott Atkinson

机构 * School of Computer Science, University of Glasgow(计算机科学学院,格拉斯哥大学) Amazon(亚马逊)

专题命中 具身导航 :navigation(abstract);分类 cs.RO、cs.CV

AI总结 本文提出基于UNet的关键点回归方法,用于自动驾驶赛车中3D圆锥定位,通过大规模数据集提升定位精度并实现颜色预测。

Comments 8 pages, 9 figures. Accepted to ICCV End-to-End 3D Learning Workshop 2025 and presented as a poster; not included in the final proceedings due to a conference administrative error

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.12846 2026-02-24 cs.RO cs.CV 62%

Unleashing the Power of Discrete-Time State Representation: Ultrafast Target-based IMU-Camera Spatial-Temporal Calibration

释放离散时间状态表示的潜力:超快基于目标的IMU-相机空间-时间校准

Junlin Song, Antoine Richard, Miguel Olivares-Mendez

专题命中 具身导航 :navigation(abstract);分类 cs.RO、cs.CV

AI总结 本文提出了一种基于离散时间状态表示的高效IMU-相机空间-时间校准方法,以提高校准效率并解决连续时间表示的计算成本问题。

Comments Accepted by ICRA 2026

详情

展开后加载摘要…

URL PDF HTML 收藏
2602.19698 2026-02-24 cs.DL cs.AI cs.CV cs.IR 62%

Iconographic Classification and Content-Based Recommendation for Digitized Artworks

图标分类与基于内容的数字艺术作品推荐

Krzysztof Kutt, Maciej Baczyński

机构 * Department of Human-Centered Artificial Intelligence(人中心人工智能系) Institute of Applied Computer Science(应用计算机科学研究所) Faculty of Physics, Astronomy and Applied Computer Science(物理、天文学与应用计算机科学学院) Jagiellonian University(雅盖隆大学)

专题命中 具身导航 :navigation(abstract);分类 cs.AI、cs.CV

AI总结 本文提出利用Iconclass词汇和AI方法,实现数字化艺术作品的图标分类与基于内容的推荐,通过四阶段工作流提升文化遗产库的导航效率。

Comments 14 pages, 7 figures; submitted to ICCS 2026 conference

详情

展开后加载摘要…

URL PDF HTML 收藏
2602.19304 2026-02-24 cs.RO cs.AI cs.HC cs.MA 62%

Safe and Interpretable Multimodal Path Planning for Multi-Agent Cooperation

安全且可解释的多模态路径规划用于多智能体协作

Haojun Shi, Suyu Ye, Katherine M. Guerrerio, Jianzhi Shen, Yifan Yin, Daniel Khashabi, Chien-Ming Huang, Tianmin Shu

机构 * Johns Hopkins University(约翰霍普金斯大学)

专题命中 具身导航 :robotic(abstract);分类 cs.RO、cs.AI

AI总结 CaPE通过多模态路径规划实现安全且可解释的多智能体协作,利用视觉-语言模型和模型基于规划器确保路径调整的安全性和可解释性。

详情

展开后加载摘要…

URL PDF HTML 收藏
2507.16874 2026-02-24 cs.MA cs.AI cs.RO 62%

Budget Allocation Policies for Real-Time Multi-Agent Path Finding

实时多智能体路径寻找中的预算分配策略

Raz Beck, Roni Stern

专题命中 具身导航 :robotics(abstract);分类 cs.RO、cs.AI

AI总结 本文提出了一种实时多智能体路径寻找中的预算分配策略,通过智能分配规划预算提升求解效率。

Comments 11 pages, 4 figures, 4 tables

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.16650 2026-02-24 eess.SY cs.LG cs.RO cs.SY math.DS math.OC 62%

Safe and Near-Optimal Control with Online Dynamics Learning

安全且近最优的控制与在线动力学学习

Manish Prajapat, Johannes Köhler, Melanie N. Zeilinger, Andreas Krause

专题命中 具身导航 :navigation(abstract);分类 cs.RO、cs.LG

AI总结 本文提出了一种安全且近最优的在线控制方法,通过最大安全动态学习在有限时间内实现高精度动态建模,同时确保全程安全运行,适用于自动驾驶和无人机等对安全要求高的领域。

详情

展开后加载摘要…

URL PDF HTML 收藏
2602.17182 2026-02-20 cs.CV cs.RO 62%

NRGS-SLAM: Monocular Non-Rigid SLAM for Endoscopy via Deformation-Aware 3D Gaussian Splatting

NRGS-SLAM:通过变形感知的3D高斯点云实现内窥镜单目非刚性SLAM

Jiwei Shan, Zeyu Cai, Yirui Li, Yongbo Chen, Lijun Han, Yun-hui Liu, Hesheng Wang, Shing Shin Cheng

机构 * Department of Mechanical and Automation Engineering and T Stone Robotics Institute, The Chinese University of Hong Kong(机械与自动化工程系和T Stone机器人研究所,香港中文大学) School of Automation and Intelligent Sensing, the State Key Laboratory of Avionics Integration and Aviation System of-Systems Synthesis, and Shanghai Key Laboratory of Navigation and Location Based Services, Shanghai Jiao Tong University(自动化与智能感知学院、航空系统集成与航空系统综合国家重点实验室、上海导航与定位服务重点实验室,上海交通大学)

专题命中 具身导航 :navigation(abstract);分类 cs.RO、cs.CV

AI总结 NRGS-SLAM通过变形感知的3D高斯点云实现内窥镜单目非刚性SLAM,有效解决刚性假设与变形耦合问题,提升定位精度与重建质量。

详情

展开后加载摘要…

URL PDF HTML 收藏
2602.16826 2026-02-20 cs.LG cs.AI 62%

HiVAE: Hierarchical Latent Variables for Scalable Theory of Mind

HiVAE:用于可扩展理论之心的层次潜在变量

Nigel Doering, Rahath Malladi, Arshia Sangwan, David Danks, Tauhidur Rahman

专题命中 具身导航 :navigation(abstract);分类 cs.AI、cs.LG

AI总结 HiVAE通过层次化变分架构提升理论之心推理能力,解决现实时空领域中的心理状态推断问题,并提出自监督对齐策略以增强潜在表示的参照性。

Comments Accepted at the Workshop on Theory of Mind for AI (ToM4AI) at the 40th AAAI Conference on Artificial Intelligence (AAAI-26), Singapore, 2026

详情

展开后加载摘要…

URL PDF HTML 收藏
2602.16764 2026-02-20 cs.LG cs.RO cs.SY eess.SY 62%

Machine Learning Argument of Latitude Error Model for LEO Satellite Orbit and Covariance Correction

低地球轨道卫星轨道和协方差校正的机器学习论证

Alex Moody, Penina Axelrad, Rebecca Russell

机构 * Draper Scholar University of Colorado Boulder(Draper学者 大学科罗拉多大学博尔德分校) University of Colorado Boulder(大学科罗拉多大学博尔德分校) The Charles Stark Draper Laboratory Inc.(查尔斯·斯泰克·德拉珀实验室公司)

专题命中 具身导航 :navigation(abstract);分类 cs.RO、cs.LG

AI总结 本文提出了一种基于机器学习的误差校正方法,用于改进低地球轨道卫星的纬度参数误差建模,通过高斯分布预测和反向传播误差来提高轨道传播精度。

Comments Appearing in 2026 IEEE Aerospace Conference

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.00074 2026-02-19 cs.AI cs.CL cs.LG 62%

Language and Experience: A Computational Model of Social Learning in Complex Tasks

语言与经验:复杂任务中的社会学习计算模型

Cédric Colas, Tracey Mills, Ben Prystawski, Michael Henry Tessler, Noah Goodman, Jacob Andreas, Joshua Tenenbaum

机构 * MIT(麻省理工学院) Stanford University(斯坦福大学) Google DeepMind(谷歌DeepMind)

专题命中 具身导航 :world model(abstract);分类 cs.AI、cs.LG

AI总结 本文提出了一种计算模型,通过结合语言指导与直接经验,模拟复杂任务中的社会学习过程,并展示了人类与模型之间的知识传递机制。

Comments Code: github.com/ccolas/language_and_experience Demo: cedriccolas.com/demos/language_and_experience

Journal ref ICLR 2026; CogSci 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2602.15926 2026-02-19 cs.CV cs.LG 62%

A Study on Real-time Object Detection using Deep Learning

基于深度学习的实时目标检测研究

Ankita Bose, Jayasravani Bhumireddy, Naveen N

机构 * Department of Computer Science and Engineering(计算机科学与工程系) GITAM University(GITAM大学) India(印度)

专题命中 具身导航 :navigation(abstract);分类 cs.CV、cs.LG

AI总结 本文研究了基于深度学习的实时目标检测方法,分析了多种算法及其在不同领域的应用,并通过实验比较了不同策略的效果。

Comments 34 pages, 18 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2602.15274 2026-02-18 cs.AI cs.LG 62%

When Remembering and Planning are Worth it: Navigating under Change

记忆与规划的价值:在变化中导航

Omid Madani, J. Brian Burns, Reza Eghbali, Thomas L. Dean

机构 * Brown University(布朗大学)

专题命中 具身导航 :navigation(abstract);分类 cs.AI、cs.LG

AI总结 研究探讨了在变动环境中,利用记忆和概率学习提升空间导航效率的方法。

详情

展开后加载摘要…

URL PDF HTML 收藏
2602.11541 2026-02-13 cs.AI cs.LG 62%

Budget-Constrained Agentic Large Language Models: Intention-Based Planning for Costly Tool Use

预算约束下的代理型大语言模型:基于意图的规划以应对昂贵的工具使用

Hanbing Liu, Chunhao Tian, Nan An, Ziyuan Wang, Pinyan Lu, Changyuan Yu, Qi Qi

机构 * Gaoling School of Artificial Intelligence, Renmin University of China(中国人民大学人工智能学院) Shanghai University of Finance(上海财经大学) Baidu Inc.(百度公司)

专题命中 具身导航 :world model(abstract);分类 cs.AI、cs.LG

AI总结 本文提出INTENT框架,通过意图感知的分层世界模型,在预算约束下提升工具使用任务的成功率并保持鲁棒性。

详情

展开后加载摘要…

URL PDF HTML 收藏
2602.09255 2026-02-13 cs.RO cs.AI 62%

STaR: Scalable Task-Conditioned Retrieval for Long-Horizon Multimodal Robot Memory

STaR:可扩展的任务条件检索用于长时多模态机器人记忆

Mingfeng Yuan, Hao Zhang, Mahan Mohammadi, Runhao Li, Jinjun Shan, Steven L. Waslander

机构 * University of Toronto Institute for Aerospace Studies(多伦多大学航空航天研究所) University of Toronto Robotics Institute(多伦多大学机器人研究所) Department of Earth and Space Science(地球与空间科学系) Lassonde School of Engineering(拉索nde工程学院) York University(约克大学)

专题命中 具身导航 :navigation(abstract);分类 cs.RO、cs.AI

AI总结 STaR通过任务条件检索算法提升机器人长时多模态记忆的可扩展性和上下文推理能力

详情

展开后加载摘要…

URL PDF HTML 收藏
2602.09949 2026-02-11 cs.CV cs.AI 62%

Bladder Vessel Segmentation using a Hybrid Attention-Convolution Framework

利用混合注意力-卷积框架进行膀胱血管分割

Franziska Krauß, Matthias Ege, Zoltan Lovasz, Albrecht Bartz-Schmidt, Igor Tsaur, Oliver Sawodny, Carina Veil

机构 * Institute for System Dynamics in the University of Stuttgart(斯图加特大学系统动力学研究所) University Hospital Tübingen(图宾根大学医院) Department of Mechanical Engineering, Stanford University(斯坦福大学机械工程系)

专题命中 具身导航 :navigation(abstract);分类 cs.AI、cs.CV

AI总结 本文提出混合注意力-卷积框架,通过结合Transformer和CNN实现高精度膀胱血管分割,解决内窥镜数据中复杂变形和黏膜褶皱等挑战。

详情

展开后加载摘要…

URL PDF HTML 收藏