arXivDaily arXiv每日学术速递 周一至周五更新

视觉与机器人

世界模型

面向环境建模、时序预测、仿真规划、具身智能和自动驾驶的世界模型方法与应用。

2026-03-02 至 2026-03-02 共收录 8 信号源:cs.AI, cs.LG, cs.CV, cs.RO, cs.MA

1. 通用世界模型 6 篇

2602.23997 2026-03-02 cs.LG cs.AI 94%

Foundation World Models for Agents that Learn, Verify, and Adapt Reliably Beyond Static Environments

为能够学习、验证和在动态环境中适应的智能体构建基础世界模型

Florent Delgrange

机构 * AI Lab, Vrije Universiteit Brussel \& Flanders Make Brussels Belgium AI Lab, Vrije Universiteit Brussel \& Flanders Make

专题命中 通用世界模型 :world model(title,abstract);world models(title,abstract);world model(title,abstract);world models(title,abstract)

AI总结 本文提出基础世界模型,通过可学习奖励模型、自适应形式验证、在线抽象校准和测试时合成,使智能体在动态环境中可靠学习、验证和适应。

Comments AAMAS 2026, Blue Sky Idea Track. 4 pages, 1 Figure

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.22353 2026-03-02 cs.LG cs.AI 94%

Context and Diversity Matter: The Emergence of In-Context Learning in World Models

上下文与多样性至关重要:世界模型中情境学习的出现

Fan Wang, Zhiyuan Chen, Yuxuan Zhong, Sunjian Zheng, Pengtao Shao, Bo Yu, Shaoshan Liu, Jianan Wang, Ning Ding, Yang Cao, Yu Kang

机构 * Shenzhen Institute of Artificial Intelligence and Robotics for Society(深圳人工智能与机器人社会研究院) University of Science and Technology of China(中国科学技术大学) Anhui Province Key Laboratory of Intelligent Low-Carbon Information Technology and Equipment(安徽省智能低碳信息技术与设备重点实验室)

专题命中 通用世界模型 :world model(title,abstract);world models(title,abstract);world model(title,abstract);world models(title,abstract)

AI总结 本文研究了世界模型中情境学习的机制,揭示了环境识别和学习的核心作用,并探讨了长上下文和多样化环境对学习效果的影响。

Journal ref 2026 International Conference on Learning Representations (ICLR)

详情

展开后加载摘要…

URL PDF HTML 收藏
2601.10553 2026-03-02 cs.CV 93%

Inference-time Physics Alignment of Video Generative Models with Latent World Models

视频生成模型的推理时间物理对齐:基于潜在世界模型

Jianhao Yuan, Xiaofeng Zhang, Felix Friedrich, Nicolas Beltran-Velez, Melissa Hall, Reyhane Askari-Hemmat, Xiaochuang Han, Nicolas Ballas, Michal Drozdzal, Adriana Romero-Soriano

机构 * FAIR, Meta Superintelligence Labs University of Oxford Mila - Qu\' e bec AI Institute Columbia University McGill University Canada CIFAR AI Chair

专题命中 通用世界模型 :world model(title,abstract);world models(title,abstract);world model(title,abstract);world models(title,abstract)

AI总结 本文提出WMReward方法,通过潜在世界模型提升视频生成的物理合理性,实验表明在多个生成设置中显著提高物理合理性,并在ICCV 2025挑战中取得第一名。

Comments 22 pages, 10 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2512.00306 2026-03-02 q-bio.CB cs.AI cs.LG 88%

VCWorld: A Biological World Model for Virtual Cell Simulation

VCWorld:一种用于虚拟细胞模拟的生物世界模型

Zhijian Wei, Runze Ma, Zichen Wang, Zhongmin Li, Shuotong Song, Shuangjia Zheng

专题命中 通用世界模型 :world model(title,abstract);world model(title,abstract);分类 cs.AI、cs.LG

AI总结 VCWorld通过整合生物学知识与大语言模型的推理能力,构建了可解释的虚拟细胞模拟模型,实现了对细胞扰动的高效预测和机理阐释。

Comments Accepted at ICLR 2026

详情

展开后加载摘要…

URL PDF HTML 收藏
2602.23969 2026-03-02 cs.MM cs.CV 81%

MSVBench: Towards Human-Level Evaluation of Multi-Shot Video Generation

MSVBench: 向多镜头视频生成的人机水平评估迈进

Haoyuan Shi, Yunxin Li, Nanhao Deng, Zhenran Xu, Xinyu Chen, Longyue Wang, Baotian Hu, Min Zhang

机构 * Harbin Institute of Technology (Shenzhen)(哈尔滨工业大学(深圳)) Alibaba International Digital Commerce(阿里巴巴国际数字商业)

专题命中 通用世界模型 :world model(abstract);world models(abstract);world model(abstract);world models(abstract)

AI总结 MSVBench通过引入分层脚本和参考图像,提出混合评估框架,验证了视频生成模型的连贯性和吸引力,并展示了其在多镜头视频生成中的有效性。

详情

展开后加载摘要…

URL PDF HTML 收藏
2511.05256 2026-03-02 hep-th 80%

Entanglement, defects, and $T\bar{T}$ on a black hole background

纠缠、缺陷与黑洞背景上的$T\bar{T}$

Ankur Dey

专题命中 通用世界模型 :world model(abstract);world models(abstract);world model(abstract);world models(abstract)

AI总结 本文研究了在黑洞背景上纠缠、缺陷与$T\bar{T}$变形之间的关系,通过计算纠缠熵验证了岛屿和缺陷极值表面方案的对偶性,并讨论了$T\bar{T}$变形对Page曲线的影响。

Comments 26 pages, 8 figures, v2 matches with published version

Journal ref J. High Energ. Phys. 2026, 242 (2026)

详情

展开后加载摘要…

URL PDF HTML 收藏

2. 具身与机器人 1 篇

2602.23706 2026-03-02 cs.RO cs.CV 56%

A Reliable Indoor Navigation System for Humans Using AR-based Technique

基于AR技术的可靠室内导航系统

Vijay U. Rathod, Manav S. Sharma, Shambhavi Verma, Aadi Joshi, Sachin Aage, Sujal Shahane

专题命中 具身与机器人 :environment model(abstract);分类 cs.CV、cs.RO

AI总结 本文提出基于AR技术的室内导航系统,利用Vuforia Area Target和A*算法提升导航精度与用户体验,适用于校园等有限空间,但需进一步优化NavMesh以适应大规模动态环境。

Comments 6 pages, 6 figures, 2 tables, Presented at 7th International Conference on Advances in Science and Technology (ICAST 2024-25)

详情

展开后加载摘要…

URL PDF HTML 收藏

3. 模型式强化学习 1 篇

2602.23285 2026-03-02 cs.AI 50%

ODEBrain: Continuous-Time EEG Graph for Modeling Dynamic Brain Networks

ODEBrain: 基于连续时间EEG图的动态脑网络建模

Haohui Jia, Zheng Chen, Lingwei Zhu, Rikuto Kotoge, Jathurshan Pradeepkumar, Yasuko Matsubara, Jimeng Sun, Yasushi Sakurai, Takashi Matsubara

机构 * Information Science and Techinology, Hokkaido University, Japan(信息科学与技术,北海道大学,日本) SANKEN, The University of Osaka, Japan(SANKEN,大阪大学,日本) Great Bay University, China(大湾大学,中国) Department of Computer Science, University of Illinois Urbana-Champaign, USA(计算机科学系,伊利诺伊大学厄巴纳-香槟分校,美国)

专题命中 模型式强化学习 :latent dynamics(abstract);分类 cs.AI

AI总结 ODEBrain通过整合时空频特征和神经ODE,提升EEG动态预测的鲁棒性和泛化能力。

详情

展开后加载摘要…

URL PDF HTML 收藏