arXivDaily arXiv每日学术速递 周一至周五更新

视觉与机器人

世界模型

面向环境建模、时序预测、仿真规划、具身智能和自动驾驶的世界模型方法与应用。

2026-04-28 至 2026-04-28 共收录 12 信号源:cs.AI, cs.LG, cs.CV, cs.RO, cs.MA

1. 通用世界模型 11 篇

2602.10298 2026-04-28 cs.CL 93%

On Emergent Social World Models -- Evidence for Functional Integration of Theory of Mind and Pragmatic Reasoning in Language Models

关于涌现的社会世界模型——证据显示语言模型中理论思维与语用推理的功能整合

Polina Tsvilodub, Jan-Felix Klumpp, Amir Mohammadpour, Jennifer Hu, Michael Franke

机构 * Department of Linguistics University of Tübingen(语言学系图宾根大学) Department of Cognitive Science Johns Hopkins University(认知科学系约翰霍普金斯大学)

专题命中 通用世界模型 :world model(title,abstract);world models(title,abstract);world model(title,abstract);world models(title,abstract)

AI总结 本文研究语言模型是否通过共享的计算机制整合一般理论思维与语言特定的语用推理,以支持语言模型是否具备涌现的社会世界模型。通过行为评估和因果机制实验,分析语言模型在七种理论思维能力子类别上的表现,发现支持功能整合假说。

Comments 39 pages, 20 figures, accepted to ACL 2026 Main Conference

详情

展开后加载摘要…

URL PDF HTML 收藏
2504.18576 2026-04-28 cs.RO 92%

DriVerse: Navigation World Model for Driving Simulation via Multimodal Trajectory Prompting and Motion Alignment

DriVerse:通过多模态轨迹提示和运动对齐实现驾驶模拟的导航世界模型

Xiaofan Li, Chenming Wu, Zhao Yang, Zhihao Xu, Dingkang Liang, Yumeng Zhang, Ji Wan, Jun Wang

机构 * Baidu Inc.(百度公司)

专题命中 通用世界模型 :world model(title,abstract);world model(title,abstract);world models(abstract);driving world model(abstract)

AI总结 DriVerse通过多模态轨迹提示和运动对齐技术,实现从单张图像和未来轨迹生成导航驱动的驾驶场景,提升了动态对象的生成精度和时间一致性。

Comments 13 pages, 5 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2501.02548 2026-04-28 cs.LG cs.AI 91%

Planning Under Observation Mismatch for Traffic Signal Control via Adaptive Modular World Models

基于观测不匹配的交通信号控制的适应性模块化世界模型规划

Zherui Huang, Yicheng Liu, Chumeng Liang, Guanjie Zheng

机构 * Shanghai Jiao Tong University(上海交通大学)

专题命中 通用世界模型 :world model(title);world models(title);world model(title);world models(title)

AI总结 本文提出适应性模块化模型,用于解决观测不匹配下的交通信号控制问题,通过模块化规划架构提升性能和数据效率。

Comments Accepted by ICAPS 2026

详情

展开后加载摘要…

URL PDF HTML 收藏
2604.23532 2026-04-28 cs.CV cs.AI 88%

Emotion-Conditioned Short-Horizon Human Pose Forecasting with a Lightweight Predictive World Model

基于轻量预测世界模型的情感条件短周期人体姿态预测

Jingni Huang, Peter Bloodsworth

专题命中 通用世界模型 :world model(title,abstract);world model(title,abstract);分类 cs.AI、cs.CV

AI总结 本文研究了情感嵌入对短周期姿态预测的辅助作用,提出轻量级自回归预测世界模型,结合姿态关键点与情感嵌入,提升情感驱动动作序列的预测性能。

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.24948 2026-04-28 cs.RO 88%

World-Env: Leveraging World Model as a Virtual Environment for VLA Post-Training

World-Env: 利用世界模型作为虚拟环境进行VLA后训练

Junjin Xiao, Yandan Yang, Xinyuan Chang, Ronghan Chen, Feng Xiong, Mu Xu, Wei-Shi Zheng, Qing Zhang

机构 * School of Computer Science and Engineering, Sun Yat-sen University, China(中山大学计算机科学与工程学院) AMap, Alibaba Group(阿里巴巴集团高德地图) Key Laboratory of Machine Intelligence and Advanced Computing, Ministry of Education, China(教育部机器智能与先进计算重点实验室)

专题命中 通用世界模型 :world model(title,abstract);world model(title,abstract);分类 cs.RO

AI总结 针对VLA模型在数据稀缺场景下的性能下降问题,提出World-Env框架,通过虚拟仿真器替代真实环境,提升安全性和效率,实现任务完成检测与持续奖励。

详情

展开后加载摘要…

URL PDF HTML 收藏
2604.22879 2026-04-28 cs.MA cs.AI cs.CR cs.LG 83%

Beyond Single-Agent Alignment: Preventing Context-Fragmented Violations in Multi-Agent Systems

超越单体对齐:防止多智能体系统中的上下文碎片化违规

Jie Wu, Ming Gong

机构 * Atlassian(Atlassian公司)

专题命中 通用世界模型 :world model(abstract);world models(abstract);world model(abstract);world models(abstract)

AI总结 本文提出Distributed Sentinel架构,通过Semantic Taint Token协议解决多智能体系统中因上下文碎片化导致的政策违规问题,实验证明其在跨域政策验证中的高效性与可靠性。

Comments 34 pages, 3 figures, 20 tables

详情

展开后加载摘要…

URL PDF HTML 收藏
2604.23278 2026-04-28 cs.AI 69%

Active Inference: A method for Phenotyping Agency in AI systems?

主动推断:一种用于AI系统中表型代理的方法?

Philip Wilson, Axel Constant, Mahault Albarracin, Nicolás Hinrichs, Jasmine Moore, Daniel Polani, Karl Friston

机构 * Independent Researcher Laboratoire d'Analyse Cognitive de l'Information, Universit\' e du Qu\' e bec \` a Montr\' e al, Qu\' e bec, Canada Centre of Excellence for AI Robotics, Sheffield Hallam University, Sheffield, UK Methods Statistical Computing, Max Planck Institute for Human Cognitive Brain Sciences, Leipzig, Germany Department of Computer Science, School of Physics, Engineering Computer Science, University of Hertfordshire, Hatfield, UK Wellcome Centre for Human Neuroimaging, University College London, London, UK Department of Engineering Informatics, University of Sussex, Falmer, Brighton, BN1 9RH, UK

专题命中 通用世界模型 :world model(abstract);world model(abstract);分类 cs.AI

AI总结 本文提出基于主动推断的表型代理方法,通过变分框架将信念、偏好和自由能最小化结合,以区分不同代理表型。

详情

展开后加载摘要…

URL PDF HTML 收藏
2506.19089 2026-04-28 cs.CL cs.AI 69%

Language Models Might Not Understand You: Evaluating Theory of Mind via Story Prompting

语言模型可能并不理解你:通过故事提示评估理论自我意识

Nathaniel Getachew, Abulhair Saparov

专题命中 通用世界模型 :world model(abstract);world model(abstract);分类 cs.AI

AI总结 通过故事提示框架评估大语言模型的理论自我意识和世界建模能力,发现模型在世界建模任务中表现优于理论自我意识任务,且对人物推理更准确。

Comments 21 pages, 17 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2604.22812 2026-04-28 cs.CY cs.LG stat.AP 67%

Cross-Course Generalizability of SRL-Aligned Predictive Models Using Digital Learning Traces

基于数字学习轨迹的SRL对齐预测模型的跨课程泛化能力

Jakob Schwerter, Loreen Sabel, Judith Bose, Matthew L. Bernacki, Di Xu, Marko Schmellenkamp, Thomas Zeume, Philipp Doebler

机构 * Hector Research Institute of Education Sciences and Psychology, University of Tübingen(教育科学与心理学赫克托研究 institute,图宾根大学) TU Dortmund University(多特蒙德技术大学) University of North-Carolina, Chapel Hill(北卡罗来纳大学教堂山分校) University of California, Irvine(加州大学尔湾分校) Ruhr University Bochum(波鸿鲁尔大学)

专题命中 通用世界模型 :predictive model(title,abstract);predictive models(title,abstract);分类 cs.LG

AI总结 研究通过分析多模态数字轨迹数据,评估SRL对齐模型在不同课程和机构中的预测性能,发现Elastic Net在跨情境泛化中表现更稳健,但模型泛化需谨慎考虑不同情境下的风险率差异。

详情

展开后加载摘要…

URL PDF HTML 收藏
2604.03768 2026-04-28 cs.AI cs.LG 56%

RL-Driven Sustainable Land-Use Allocation for the Lake Malawi Basin

基于强化学习的可持续土地利用分配:马拉维盆地

Ying Yao

机构 * Georgia Institute of Technology(佐治亚理工学院)

专题命中 通用世界模型 :environment model(abstract);分类 cs.AI、cs.LG

AI总结 本文提出基于深度强化学习的框架,用于优化马拉维盆地的土地利用分配以最大化生态系统服务价值,通过空间奖励塑造引导生态友好的土地利用模式。

Comments 9 pages, 11 figures; added baseline comparison under "Result" section; revised limitation and discussion

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.01635 2026-04-28 cs.LG cs.AI cs.DC cs.PF 50%

Reliable Microservice Tail Latency Prediction via Decoupled Dual-Stream Learning and Gradient Modulation

通过解耦双流学习和梯度调制实现可靠微服务尾延迟预测

Wenzhuo Qian, Hailiang Zhao, Jiayi Chen, Ziqi Wang, Tianlv Chen, Zhiwei Ling, Xinkui Zhao, Kingsum Chow, Albert Y. Zomaya, Shuiguang Deng

机构 * College of Computer Science and Technology, Zhejiang University(浙江大学计算机科学与技术学院) School of Software Technology, Zhejiang University(浙江大学软件学院) Polytechnic Institute, Zhejiang University(浙江大学 polytechnic 院) School of Computer Science, University of Sydney(悉尼大学计算机科学学院)

专题命中 通用世界模型 :分类 cs.AI、cs.LG;predictive model(abstract);predictive models(abstract)

AI总结 本文提出USRFNet框架,通过解耦需求与容量建模,利用图神经网络和门控MLP分别建模流量和资源动态,结合层次张量融合提升预测精度,实验显示在三个真实世界基准上MAPE降低15.62%-26.11%。

详情

展开后加载摘要…

URL PDF HTML 收藏

2. 模型式强化学习 1 篇

2506.05480 2026-04-28 cs.GR cs.CV cs.LG 56%

ODE-GS: Latent ODEs for Dynamic Scene Extrapolation with 3D Gaussian Splatting

ODE-GS:基于动态场景外推的隐式ODE与3D高斯点划法

Daniel Wang, Patrick Rim, Tian Tian, Dong Lao, Alex Wong, Ganesh Sundaramoorthi

机构 * Yale University(耶鲁大学) TU Delft(代尔夫特理工大学) Louisiana State University(路易斯安那州立大学) RTX(RTX公司)

专题命中 模型式强化学习 :latent dynamics(abstract);分类 cs.LG、cs.CV

AI总结 ODE-GS结合隐式ODE与3D高斯点划法,通过连续时间隐式动态建模实现动态3D场景的未来外推,优于传统时间条件变形网络方法,提升外推性能19.8%。

Comments Accepted at ICLR 2026

详情

展开后加载摘要…

URL PDF HTML 收藏