arXivDaily arXiv每日学术速递 周一至周五更新
arXiv周末暂无论文更新,休息一下吧,周末愉快~~

视觉与机器人

机器人 / 具身智能

机器人、具身智能、机器人学习、操作、导航和具身世界模型。

2026-02-25 至 2026-02-25 共收录 68 信号源:cs.RO, cs.AI, cs.CV, cs.LG

1. 具身导航 11 篇

2411.06144 2026-02-25 astro-ph.IM 50%

First Use of GPS Satellites for Beam Calibration of Radio Dish Telescopes

利用GPS卫星进行射电天线抛物面镜校准的首次应用

Sabrina Berger, Arianna Lasinski, Vincent MacKay, Eamon Egan, Dallas Wulf, Aman Chokshi, Jonathan Sievers

专题命中 具身导航 :navigation(abstract)

AI总结 利用GPS卫星首次实现射电天线抛物面镜校准,通过高频次卫星观测提升校准精度,为射电天文应用提供新方法。

Comments Accepted to PASA. 17 pages, 12 figures

Journal ref Publ. Astron. Soc. Aust. 43 (2026) e016

详情

展开后加载摘要…

URL PDF HTML 收藏
2602.20565 2026-02-25 physics.flu-dyn 50%

Hydrodynamic modulation via cupping in a crustacean-inspired propulsor

通过蟹足的杯状调节实现流体动力学调节

Sara Oliveira Santos, Maggie Brown, Minki Kim, Nils Tack, Monica M. Wilhelmus

专题命中 具身导航 :robotic(abstract)

AI总结 研究通过调节蟹足的杯状角度,发现适度角度可优化推力-升力平衡,揭示了虾足作为混合推进器的机制。

详情

展开后加载摘要…

URL PDF HTML 收藏

2. 具身推理 3 篇

2602.18639 2026-02-25 cs.LG math.OC 83%

Learning Invariant Visual Representations for Planning with Joint-Embedding Predictive World Models

学习具有联合嵌入预测的世界模型以实现规划

Leonardo F. Toso, Davit Shadunts, Yunyang Lu, Nihal Sharma, Donglin Zhan, Nam H. Nguyen, James Anderson

机构 * Department of Electrical Engineering, Columbia University, New York, USA.(电气工程系,哥伦比亚大学,纽约,美国)

专题命中 具身推理 :world model(title,abstract);navigation(abstract);分类 cs.LG

AI总结 本文提出了一种联合嵌入预测世界模型,通过引入双模拟编码器增强对慢特征的鲁棒性,提升在减少潜在空间中的导航任务性能。

详情

展开后加载摘要…

URL PDF HTML 收藏
2507.04002 2026-02-25 cs.CV cs.RO eess.IV 81%

NRSeg: Noise-Resilient Learning for BEV Semantic Segmentation via Driving World Models

NRSeg: 通过驾驶世界模型实现噪声鲁棒的BEV语义分割学习

Siyu Li, Fei Teng, Yihong Cao, Kailun Yang, Zhiyong Li, Yaonan Wang

机构 * School of Artificial Intelligence and Robotics and the National Engineering Research Center of Robot Visual Perception and Control Technology, Hunan University(人工智能与机器人学院和机器人视觉感知与控制技术国家工程研究中心,湖南大学)

专题命中 具身推理 :world model(title,abstract);分类 cs.RO、cs.CV

AI总结 NRSeg通过驾驶世界模型生成的合成数据增强BEV语义分割学习,提出PGCM、BiDPP和HLSE模块以提升模型鲁棒性和分割性能。

Comments Accepted to IEEE Transactions on Image Processing (TIP). The source code will be made publicly available at https://github.com/lynn-yu/NRSeg

详情

展开后加载摘要…

URL PDF HTML 收藏
2602.20550 2026-02-25 cs.CV 57%

The Finite Primitive Basis Theorem for Computational Imaging: Formal Foundations of the OperatorGraph Representation

计算成像的有限原始基定理:操作图表示法的正式基础

Chengshuai Yang

机构 * NextGen PlatformAI C Corp, USA(NextGen平台AI公司)

专题命中 具身推理 :world model(abstract);分类 cs.CV

AI总结 该研究提出有限原始基定理,证明所有成像前向模型可表示为由11个标准原语构成的DAG,并通过实验验证其在多种成像模态中的有效性。

详情

展开后加载摘要…

URL PDF HTML 收藏

3. 模仿学习与强化学习 6 篇

2602.20715 2026-02-25 cs.RO 88%

IG-RFT: An Interaction-Guided RL Framework for VLA Models in Long-Horizon Robotic Manipulation

IG-RFT: 一种基于交互引导的强化学习框架用于长时域机器人操作中的VLA模型

Zhian Su, Weijie Kong, Haonan Dong, Huixu Dong

机构 * Grasp Laboratory, Mechanical Engineering Department, Zhejiang University(浙江大学机械工程学院抓取实验室) Torch Kernel Co., Ltd.(火炬核科技有限公司)

专题命中 模仿学习与强化学习 :manipulation(title,abstract);robotic(title,abstract);分类 cs.RO

AI总结 IG-RFT通过交互引导强化学习框架提升VLA模型在长时域机器人任务中的性能,实现85%的成功率,优于传统方法。

详情

展开后加载摘要…

URL PDF HTML 收藏
2511.23055 2026-02-25 cs.AI 79%

MindPower: Enabling Theory-of-Mind Reasoning in VLM-based Embodied Agents

MindPower: 使基于视觉语言的具身智能体具备心智理论推理能力

Ruoxuan Zhang, Qiyun Zheng, Zhiyu Zhou, Ziqi Liao, Siyu Wu, Jian-Yu Jiang-Lin, Bin Wen, Hongxia Xie, Jianlong Fu, Wen-Huang Cheng

机构 * Jilin University(吉林大学) National Taiwan University(台湾大学) Microsoft Research Asia(微软亚洲研究院)

专题命中 模仿学习与强化学习 :embodied agent(title,abstract);分类 cs.AI

AI总结 MindPower通过整合感知、心理推理、决策和行动,使基于视觉语言的具身智能体具备心智理论推理能力,并在决策和行动生成上超越GPT-4o。

Comments Accepted by CVPR 2026

详情

展开后加载摘要…

URL PDF HTML 收藏
2602.20517 2026-02-25 cs.AI cs.CL cs.LG 73%

Inner Speech as Behavior Guides: Steerable Imitation of Diverse Behaviors for Human-AI coordination

内部言语作为行为引导:为人类-人工智能协调的可操控模仿

Rakshit Trivedi, Kartik Sharma, David C Parkes

机构 * Massachusetts Institute of Technology(麻省理工学院) Georgia Institute of Technology(佐治亚理工学院) Harvard University(哈佛大学)

专题命中 模仿学习与强化学习 :manipulation(abstract);robotic(abstract);分类 cs.AI、cs.LG

AI总结 MIMIC通过语言作为内部表示,利用视觉语言模型和扩散策略实现人类-人工智能协调中的行为引导与多样化的模仿控制。

Comments Spotlight paper at NeurIPS 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2505.19698 2026-02-25 cs.LG cs.AI cs.RO 67%

Performance Asymmetry in Model-Based Reinforcement Learning

基于模型的强化学习中的性能不对称性

Jing Yu Lim, Rushi Shah, Zarif Ikram, Samson Yu, Haozhe Ma, Tze-Yun Leong, Dianbo Liu

机构 * Department of XXX, University of YYY, Location, Country(XXX系,YYY大学,地点,国家) School of ZZZ, Institute of WWW, Location, Country(ZZZ学院,WWW研究所,地点,国家) National University of Singapore(新加坡国立大学)

专题命中 模仿学习与强化学习 :world model(abstract);分类 cs.RO、cs.AI、cs.LG

AI总结 本文提出JEDI世界模型,通过解决基于模型的强化学习中的性能不对称问题,在Human-Optimal任务和Breakout上取得最优成绩,同时提升计算效率。

Comments Preprint

详情

展开后加载摘要…

URL PDF HTML 收藏
2507.15444 2026-02-25 cs.RO cs.CV 66%

Low-Latency Event-Based Velocimetry for Quadrotor Control in a Narrow Pipe

窄管道中四旋翼无人机低延迟事件基速度测量控制

Leonard Bauersfeld, Davide Scaramuzza

机构 * Robotics and Perception Group, University of Zurich(苏黎世大学机器人与感知组)

专题命中 模仿学习与强化学习 :robotics(abstract,journal_ref);分类 cs.RO、cs.CV

AI总结 本文提出了一种基于实时流场测量的四旋翼无人机闭环控制方法,用于在狭窄管道中实现稳定悬停,通过低延迟事件基烟雾速度测量和强化学习控制器有效对抗气动扰动。

Comments 19 pages

Journal ref in IEEE Transactions on Robotics, vol. 42, pp. 1-19, 2026

详情

展开后加载摘要…

URL PDF HTML 收藏
2505.03356 2026-02-25 cs.RO 57%

Effective Reinforcement Learning Control using Conservative Soft Actor-Critic

通过保守的软演员-评论家实现有效的强化学习控制

Zhiwei Shang, Xinyi Yuan, Wenjun Huang, Yunduan Cui, Di Chen, Meixin Zhu

机构 * School of Data Science, The Chinese University of Hong Kong, Shenzhen, China(香港中文大学(深圳)数据科学学院) Graduate School of Engineering Science, Osaka University, Japan(大阪大学工学研究科) Shenzhen Institute of Advanced Technology, Chinese Academy of Sciences, China(中国科学院深圳先进技术研究院) Department of Electrical and Electronic Engineering, The Hong Kong Polytechnic University, Hong Kong SAR, China(香港理工大学电子与电气工程系) School of Transportation, Southeast University, China(东南大学交通运输学院)

专题命中 模仿学习与强化学习 :robotic(abstract);分类 cs.RO

AI总结 本文提出保守的软演员-评论家算法,通过熵和相对熵正则化提升强化学习控制的稳定性与效率。

Comments 8 pages, 9 figures

详情

展开后加载摘要…

URL PDF HTML 收藏

4. 人机交互与遥操作 4 篇

2401.00959 2026-02-25 cs.HC 78%

Creating an Intelligent Dementia-Friendly Living Space: A Feasibility Study Integrating Assistive Robotics, Wearable Sensors, and Spatial Technology

构建智能痴呆友好型生活环境:整合辅助机器人、可穿戴传感器和空间技术的可行性研究

Arshia A Khan, Rupak Kumar Das, Anna Martin, Dale Dowling, Rana Imtiaz

专题命中 人机交互与遥操作 :robotics(title,abstract)

AI总结 本研究通过整合辅助机器人、可穿戴传感器和空间技术,探索智能环境中提升痴呆症护理质量的可行性。

Journal ref IEEE Healthcom, 2024

详情

展开后加载摘要…

URL PDF HTML 收藏
2602.20219 2026-02-25 cs.RO cs.AI 73%

An Approach to Combining Video and Speech with Large Language Models in Human-Robot Interaction

一种结合视频和语音的大型语言模型在人机交互中的应用方法

Guanting Shen, Zi Tian

机构 * Dalian University of Technology(大连理工大学)

专题命中 人机交互与遥操作 :manipulation(abstract);robotic(abstract);分类 cs.RO、cs.AI

AI总结 本文提出结合视觉-语言模型、语音处理和模糊逻辑的多模态人机交互框架,通过语音指令实现机械臂的精准操控,实验显示命令执行准确率达75%,为未来人机协作提供灵活的基础。

Comments Preprint currently under revision

详情

展开后加载摘要…

URL PDF HTML 收藏
2602.13932 2026-02-25 cs.RO 57%

Joint Task Assistance Planning via Nested Branch and Bound (Extended Version)

通过嵌套分支限界法进行联合任务协助规划(扩展版)

Omer Daube, Oren Salzman

专题命中 人机交互与遥操作 :robotic(abstract);分类 cs.RO

AI总结 本文提出嵌套分支限界法解决机器人联合任务协助规划问题,通过分层探索路径空间,实现高效计算并显著提升性能。

详情

展开后加载摘要…

URL PDF HTML 收藏
2602.20180 2026-02-25 cs.CY cs.HC cs.RO 57%

Is Robot Labor Labor? Delivery Robots and the Politics of Work in Public Space

机器人劳动是劳动吗?配送机器人与公共空间中的劳动政治

EunJeong Cheon, Do Yeon Shin

机构 * Syracuse University(Syracuse大学) University of Illinois Chicago(伊利诺伊大学香槟分校)

专题命中 人机交互与遥操作 :robotic(abstract);分类 cs.RO

AI总结 本文探讨配送机器人在公共空间中如何重新配置劳动,揭示机器人特权及人类与机器人的互动差异。

详情

展开后加载摘要…

URL PDF HTML 收藏

5. 机器人数据与评测 19 篇

2602.21161 2026-02-25 cs.RO 86%

ActionReasoning: Robot Action Reasoning in 3D Space with LLM for Robotic Brick Stacking

动作推理:利用LLM进行三维空间中的机器人动作推理以实现积木堆叠

Guangming Wang, Qizhen Ying, Yixiong Jing, Olaf Wysocki, Brian Sheil

机构 * CV4DT, CSIC, Department of Engineering, University of Cambridge(CV4DT、CSIC、工程系、剑桥大学)

专题命中 机器人数据与评测 :robotic(title,abstract);embodied AI(abstract);manipulation(abstract);分类 cs.RO

AI总结 本文提出ActionReasoning框架,利用LLM进行三维空间中的动作推理,实现稳定积木堆叠,提升机器人操作的泛化能力。

Comments 8 pages, 5 figures, accepted by the 2026 IEEE International Conference on Robotics and Automation

详情

展开后加载摘要…

URL PDF HTML 收藏
2602.21203 2026-02-25 cs.RO cs.CV cs.LG 85%

Squint: Fast Visual Reinforcement Learning for Sim-to-Real Robotics

Squint:用于机器人仿真到现实的高效视觉强化学习

Abdulaziz Almuzairee, Henrik I. Christensen

专题命中 机器人数据与评测 :robotics(title,abstract);manipulation(abstract);分类 cs.RO、cs.CV、cs.LG

AI总结 Squint通过并行仿真、分布性批评者等方法,实现了在实际时间中比现有视觉强化学习方法更快的训练速度,成功实现了从仿真到现实的机器人任务迁移。

Comments For website and code, see https://aalmuzairee.github.io/squint

详情

展开后加载摘要…

URL PDF HTML 收藏
2507.02602 2026-02-25 cs.CV cs.AI 81%

Addressing Camera Sensors Faults in Vision-Based Navigation: Simulation and Dataset Development

解决基于视觉的导航中摄像头传感器故障:仿真与数据集开发

Riccardo Gallon, Fabian Schiemenz, Alessandra Menicucci, Eberhard Gill

机构 * Department of Space Systems Engineering, Faculty of Aerospace Engineering(航天系统工程系,航空航天工程学院) Airbus Defence and Space GmbH(空客防御与空间公司)

专题命中 机器人数据与评测 :navigation(title,abstract);分类 cs.AI、cs.CV

AI总结 本文通过仿真与数据集开发,解决基于视觉导航中摄像头传感器故障的检测问题,提供故障图像数据集用于训练AI故障检测算法。

Comments Submitted to Acta Astronautica

详情

展开后加载摘要…

URL PDF HTML 收藏
2602.20687 2026-02-25 cs.AI 79%

How Foundational Skills Influence VLM-based Embodied Agents:A Native Perspective

基础技能如何影响基于VLM的具身体验代理:一种原生视角

Bo Peng, Pi Bu, Keyu Pan, Xinrun Xu, Yinxiu Zhao, Miao Chen, Yang Du, Lin Li, Jun Song, Tong Xu

机构 * Alibaba-inc(阿里巴巴集团)

专题命中 机器人数据与评测 :embodied agent(title,abstract);分类 cs.AI

AI总结 本文提出NativeEmbodied基准测试,通过统一的原生低级动作空间评估VLM驱动具身体验代理的综合性能,并揭示基础技能缺陷对高级任务性能的限制。

详情

展开后加载摘要…

URL PDF HTML 收藏
2602.20566 2026-02-25 cs.RO cs.CV 73%

BFA++: Hierarchical Best-Feature-Aware Token Prune for Multi-View Vision Language Action Model

BFA++: 多视角视觉语言动作模型的分层最佳特征感知令牌剪枝

Haosheng Li, Weixin Mao, Zihan Lan, Hongwei Xiong, Hongan Wang, Chenyang Si, Ziwei Liu, Xiaoming Deng, Hua Chen

机构 * Institute of Software, Chinese Academy of Sciences(中国科学院软件研究所) University of Chinese Academy of Sciences(中国科学院大学) LimX Dynamic(LimX动态) Nanjing University(南京大学) Nanyang Technological University(南洋理工大学) Zhejiang University(浙江大学)

专题命中 机器人数据与评测 :manipulation(abstract);robotic(abstract);分类 cs.RO、cs.CV

AI总结 BFA++通过分层最佳特征感知令牌剪枝提升多视角视觉语言动作模型的效率和成功率

Comments 9 pages, 10 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2602.21015 2026-02-25 cs.CV 70%

From Perception to Action: An Interactive Benchmark for Vision Reasoning

从感知到行动:一个交互式基准测试用于视觉推理

Yuhao Wu, Maojia Song, Yihuai Lan, Lei Wang, Zhiqiang Hu, Yao Xiao, Heng Zhou, Weihua Zheng, Dylan Raharja, Soujanya Poria, Roy Ka-Wei Lee

机构 * Singapore University of Technology(新加坡科技设计大学) Singapore Management University (SMU), Singapore(新加坡管理学院) Nanyang Technological University (NTU), Singapore(南洋理工大学) University of Science(科学大学)

专题命中 机器人数据与评测 :embodied agent(abstract);manipulation(abstract);分类 cs.CV

AI总结 CHAIN基准测试通过交互式3D物理驱动环境,评估模型在动态环境中理解、规划和执行基于物理约束的结构化动作序列的能力。

Comments Work in processing. Website: https://social-ai-studio.github.io/CHAIN/

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.13394 2026-02-25 cs.CV 70%

Spatial-DISE: A Unified Benchmark for Evaluating Spatial Reasoning in Vision-Language Models

空间-DisE:评估视觉语言模型空间推理能力的统一基准

Xinmiao Huang, Qisong He, Zhenglin Huang, Boxuan Wang, Zhuoyun Li, Guangliang Cheng, Yi Dong, Xiaowei Huang

机构 * School of Computer Science & informatics, University of Liverpool(计算机科学与信息学系,利物浦大学)

专题命中 机器人数据与评测 :robotics(abstract);navigation(abstract);分类 cs.CV

AI总结 本文提出Spatial-DISE基准,用于评估视觉语言模型的空间推理能力,通过四个象限分类和大规模数据集,揭示当前模型在多步骤多视角推理上的不足。

Comments ICLR 2026 Accepted Project Page: https://shinmohuang.github.io/spatialdise_page/

详情

展开后加载摘要…

URL PDF HTML 收藏
2506.04867 2026-02-25 cs.AI cs.HC cs.LG cs.RO 69%

Sensory-Motor Control with Large Language Models via Iterative Policy Refinement

通过迭代策略优化实现大语言模型的感官-运动控制

Jônata Tyska Carvalho, Stefano Nolfi

机构 * Federal University of Santa Catarina (UFSC)(圣卡塔琳娜联邦大学) Institute of Cognitive Sciences and Technologies (ISTC-CNR)(认知科学与技术研究所)

专题命中 机器人数据与评测 :embodied agent(abstract);分类 cs.RO、cs.AI、cs.LG;robotics(comments)

AI总结 本文提出通过迭代策略优化,利用大语言模型控制具身代理,结合符号知识与子符号感官运动数据实现高效控制。

Comments Final version of the article accepted for publication on Scientific Reports. 29 pages (13 pages are from appendix), 8 figures, 2 tables, code for experiments replication and supplementary material provided at https://github.com/jtyska/llm-robotics-article/

详情

展开后加载摘要…

URL PDF HTML 收藏
2602.21119 2026-02-25 cs.RO cs.AI 66%

Cooperative-Competitive Team Play of Real-World Craft Robots

现实世界工艺机器人的协作-竞争团队合作

Rui Zhao, Xihui Li, Yizheng Zhang, Yuzhen Liu, Zhong Zhang, Yufeng Zhang, Cheng Zhou, Zhengyou Zhang, Lei Han

机构 * Tencent Robotics X Laboratory(腾讯机器人实验室) Tsinghua University(清华大学)

专题命中 机器人数据与评测 :robotic(abstract);分类 cs.RO、cs.AI;robotics(comments)

AI总结 本文提出了一种针对现实世界工艺机器人协作与竞争任务的强化学习方法,通过引入OODSI技术提升仿真到现实的迁移性能。

Comments Accepted by 2026 IEEE International Conference on Robotics and Automation (ICRA 2026), Vienna, Austria

详情

展开后加载摘要…

URL PDF HTML 收藏
2602.20709 2026-02-25 cs.CV cs.AI 62%

Onboard-Targeted Segmentation of Straylight in Space Camera Sensors

空间相机传感器中 straylight 的 onboard 目标分割

Riccardo Gallon, Fabian Schiemenz, Alessandra Menicucci, Eberhard Gill

机构 * Department of Space Systems Engineering, Faculty of Aerospace Engineering(航天系统工程系,航空航天工程学院) Airbus Defence and Space GmbH(Airbus防务与空间有限公司)

专题命中 机器人数据与评测 :navigation(abstract);分类 cs.AI、cs.CV

AI总结 本研究提出了一种基于 AI 的方法,用于在航天器资源受限硬件上实现空间相机中 straylight 的语义分割,并通过自定义指标评估其系统性能。

Comments Submitted to Aerospace Science and Technology

详情

展开后加载摘要…

URL PDF HTML 收藏
2602.20636 2026-02-25 cs.CV cs.AI 62%

SurgAtt-Tracker: Online Surgical Attention Tracking via Temporal Proposal Reranking and Motion-Aware Refinement

SurgAtt-Tracker: 通过时间提案重排和运动感知细化实现在线手术注意力跟踪

Rulin Zhou, Guankun Wang, An Wang, Yujie Ma, Lixin Ouyang, Bolin Cui, Junyan Li, Chaowei Zhu, Mingyang Li, Ming Chen, Xiaopin Zhong, Peng Lu, Jiankun Wang, Xianming Liu, Hongliang Ren

机构 * Department of Electronic Engineering, The Chinese University of Hong Kong, Hong Kong SAR, China(香港中文大学电子工程系) Division of Gastrointestinal Surgery, Shenzhen People's Hospital, China(深圳人民医院胃肠外科) College of Mechatronics and Control Engineering, Shenzhen University, China(深圳大学机电与控制工程学院) Department of Mechanical Engineering, The University of Hong Kong, Hong Kong SAR, China(香港大学机械工程系) Department of Electronic and Electrical Engineering, Southern University of Science and Technology(南方科技大学电子与电气工程系)

专题命中 机器人数据与评测 :robotic(abstract);分类 cs.AI、cs.CV

AI总结 SurgAtt-Tracker通过时间提案重排和运动感知细化实现手术注意力跟踪,提供帧级视野指导以支持机器人规划和自动摄像头控制。

详情

展开后加载摘要…

URL PDF HTML 收藏
2602.20867 2026-02-25 cs.CR cs.AI cs.CE cs.ET 57%

SoK: Agentic Skills -- Beyond Tool Use in LLM Agents

SoK: 代理技能——超越LLM代理中的工具使用

Yanna Jiang, Delong Li, Haiyu Deng, Baihe Ma, Xu Wang, Qin Wang, Guangsheng Yu

机构 * University of Technology Sydney(悉尼科技大学) CSIRO Data61(CSIRO数据61)

专题命中 机器人数据与评测 :robotics(abstract);分类 cs.AI

AI总结 本文探讨了代理技能在LLM代理中的应用,分析了技能的生命周期管理、分类方法及其安全影响,并提出了未来研究挑战。

详情

展开后加载摘要…

URL PDF HTML 收藏
2602.20813 2026-02-25 cs.AI 57%

Pressure Reveals Character: Behavioural Alignment Evaluation at Depth

压力揭示特性:深度下的行为对齐评估

Nora Petrova, John Burden

机构 * Prolific

专题命中 机器人数据与评测 :manipulation(abstract);分类 cs.AI

AI总结 本文提出一个深度行为对齐评估基准,通过多轮场景揭示模型行为倾向,发现多数模型在多个类别存在弱点,且对齐行为呈现统一构念特性。

Comments Preprint

详情

展开后加载摘要…

URL PDF HTML 收藏
2602.20773 2026-02-25 cs.CV 57%

Federated Learning for Cross-Modality Medical Image Segmentation via Augmentation-Driven Generalization

通过增强驱动泛化实现跨模态医学图像分割的联邦学习

Sachin Dudda Nagaraju, Ashkan Moradi, Bendik Skarre Abrahamsen, Mattijs Elschot

机构 * Department of Circulation and Medical Imaging, Norwegian University of Science and Technology(循环医学成像系,挪威科学与技术大学)

专题命中 机器人数据与评测 :manipulation(abstract);分类 cs.CV

AI总结 本文提出一种联邦学习方法,通过增强驱动泛化实现跨模态医学图像分割,提升模型泛化能力的同时保护数据隐私。

Comments Submitted to IEEE JBHI

详情

展开后加载摘要…

URL PDF HTML 收藏
2602.20639 2026-02-25 cs.AI 57%

Grounding LLMs in Scientific Discovery via Embodied Actions

通过具身动作 grounding LLMs 在科学发现中

Bo Zhang, Jinfeng Zhou, Yuxuan Chen, Jianing Yin, Minlie Huang, Hongning Wang

机构 * Tsinghua University, Beijing, China(清华大学)

专题命中 机器人数据与评测 :embodied agent(abstract);分类 cs.AI

AI总结 通过具身动作框架,LLMs在科学发现中实现更稳定的物理模拟和更高的建模准确性。

Comments 24 pages, 7 figures, 7 tables. Preprint

详情

展开后加载摘要…

URL PDF HTML 收藏