arXivDaily arXiv每日学术速递 周一至周五更新

视觉与机器人

世界模型

面向环境建模、时序预测、仿真规划、具身智能和自动驾驶的世界模型方法与应用。

2026-02-26 至 2026-02-26 共收录 6 信号源:cs.AI, cs.LG, cs.CV, cs.RO, cs.MA
2602.21467 2026-02-26 cs.LG 93%

Geometric Priors for Generalizable World Models via Vector Symbolic Architecture

基于向量符号架构的几何先验用于通用世界模型

William Youngwoo Chung, Calvin Yeung, Hansen Jin Lillemark, Zhuowen Zou, Xiangjian Liu, Mohsen Imani

机构 * University of California Irvine(加州大学伊文斯分校) University of California San Diego(加州大学圣地亚哥分校)

专题命中 通用世界模型 :world model(title,abstract);world models(title,abstract);world model(title,abstract);world models(title,abstract)

AI总结 本文提出基于向量符号架构的几何先验方法,通过学习复数空间的群结构,实现高效且可解释的世界模型,提升多步组合和泛化能力。

Comments 9 pages, accepted to Neurips 2025 Workshop Symmetry and Geometry in Neural Representations

详情

展开后加载摘要…

URL PDF HTML 收藏
2501.16443 2026-02-26 cs.LG cs.CV 93%

Object-Centric World Models from Few-Shot Annotations for Sample-Efficient Reinforcement Learning

基于少量标注的面向对象世界模型用于样本高效强化学习

Weipu Zhang, Adam Jelley, Trevor McInroe, Amos Storkey, Gang Wang

机构 * University of Edinburgh(爱丁堡大学) Beijing Institute of Technology(北京理工大学)

专题命中 通用世界模型 :world model(title,abstract);world models(title);world model(title,abstract);world models(title)

AI总结 本文提出OC-STORM框架,通过面向对象表示提升MBRL在复杂视觉领域的样本效率,实验证明其在Atari和Hollow Knight中的优越性能。

详情

展开后加载摘要…

URL PDF HTML 收藏
2602.22010 2026-02-26 cs.RO cs.CV 88%

World Guidance: World Modeling in Condition Space for Action Generation

世界引导:在条件空间中进行世界建模以生成动作

Yue Su, Sijin Chen, Haixin Shi, Mingyu Liu, Zhengshen Zhang, Ningyuan Huang, Weiheng Zhong, Zhengbang Zhu, Yuxiao Liu, Xihui Liu

机构 * The University of Hong Kong(香港大学)

专题命中 通用世界模型 :world model(title,abstract);world model(title,abstract);分类 cs.CV、cs.RO

AI总结 WoG通过在条件空间中建模未来观测,提升视觉-语言-动作模型的动作生成能力与泛化性能。

Comments Project Page: https://selen-suyue.github.io/WoGNet/

详情

展开后加载摘要…

URL PDF HTML 收藏
2602.20685 2026-02-26 cs.CV 88%

RAYNOVA: Scale-Temporal Autoregressive World Modeling in Ray Space

RAYNOVA:在射线空间中实现尺度-时间自回归世界建模

Yichen Xie, Chensheng Peng, Mazen Abdelfattah, Yihan Hu, Jiezhi Yang, Eric Higgins, Ryan Brigden, Masayoshi Tomizuka, Wei Zhan

机构 * Applied Intuition UC Berkeley(加州大学伯克利分校)

专题命中 通用世界模型 :world model(title,abstract);world model(title,abstract);分类 cs.CV

AI总结 RAYNOVA通过双因果自回归框架实现尺度-时间自回归世界建模,在驾驶场景中实现多视角视频生成,具有更高的吞吐量和可控性。

Comments Accepted by CVPR 2026; Project website: https://raynova-ai.github.io/

详情

展开后加载摘要…

URL PDF HTML 收藏
2602.20659 2026-02-26 cs.AI 69%

Recursive Belief Vision Language Action Models

递归信念视觉语言动作模型

Vaidehi Bagaria, Bijo Sebastian, Nirav Kumar Patel

机构 * Department of Engineering Design, Indian Institute of Technology, Madras, Chennai, India(工程设计系,印度理工学院,马德拉斯,钦奈,印度)

专题命中 通用世界模型 :world-model(abstract);world-model(abstract);分类 cs.AI

AI总结 RB-VLA通过信念与意图联合条件化扩散策略,实现长周期任务的高效执行,显著提升成功率并降低推理延迟。

详情

展开后加载摘要…

URL PDF HTML 收藏
2404.12097 2026-02-26 eess.SY cs.LG cs.SY 67%

MPC of Uncertain Nonlinear Systems with Meta-Learning for Fast Adaptation of Neural Predictive Models

具有元学习的不确定非线性系统MPC以实现神经预测模型的快速适应

Jiaqi Yan, Ankush Chakrabarty, Alisa Rupenyan, John Lygeros

机构 * Automatic Control Laboratory, ETH Zurich(瑞士苏黎世联邦理工学院自动控制实验室) Mitsubishi Electric Research Laboratories(三菱电机研究实验室) ZHAW Centre for Artificial Intelligence, Zurich University of Applied Sciences(瑞士应用科学大学人工智能中心)

专题命中 通用世界模型 :predictive model(title,abstract);predictive models(title,abstract);分类 cs.LG

AI总结 本文提出一种基于元学习的MPC方法,通过聚合神经状态空间模型快速适应非线性系统,提升控制性能。

详情

展开后加载摘要…

URL PDF HTML 收藏