arXivDaily arXiv每日学术速递 周一至周五更新

期刊&会议

International Conference on Computer Vision · 会议 · Computer Vision

2026-04-07 至 2026-04-07 共收录 2
2604.03334 2026-04-07 cs.CV

Bridging the Dimensionality Gap: A Taxonomy and Survey of 2D Vision Model Adaptation for 3D Analysis

弥合维度差距:2D视觉模型适应3D分析的分类与综述

Akshat Pandya, Bhavuk Jain

机构 * Independent Researcher(独立研究员)

AI总结 本文综述了将2D视觉模型适应3D分析的策略,分类为数据导向、架构导向和混合方法,探讨了计算复杂度、预训练依赖性和几何归纳偏置的权衡。

Comments VISAPP 2026

Journal ref Proceedings of the 21st International Conference on Computer Vision Theory and Applications - Volume 3: VISAPP 2026; ISBN 978-989-758-804-4; ISSN 2184-4321, SciTePress, pages 353-364

详情

展开后加载摘要…

URL PDF HTML 收藏
2503.08751 2026-04-07 cs.CV cs.LG

Disentangled World Models: Learning to Transfer Semantic Knowledge from Distracting Videos for Reinforcement Learning

解耦世界模型:从干扰视频中学习转移语义知识以用于强化学习

Qi Wang, Zhipeng Zhang, Baao Xie, Xin Jin, Yunbo Wang, Shiyu Wang, Liaomo Zheng, Xiaokang Yang, Wenjun Zeng

机构 * MoE Key Lab of Artificial Intelligence, AI Institute, Shanghai Jiao Tong University(上海交通大学人工智能研究院教育部人工智能重点实验室) Ningbo Institute of Digital Twin, Eastern Institute of Technology, Ningbo, China(东方理工高等研究院宁波数字孪生研究院) Ningbo Key Laboratory of Spatial Intelligence and Digital Derivative, Ningbo, China(宁波市空间智能与数字衍生重点实验室) University of Chinese Academy of Sciences(中国科学院大学) Shenyang Institute of Computing Technology, Chinese Academy of Sciences(中国科学院沈阳计算技术研究所) Shenyang CASNC Technology Co., Ltd(沈阳中科数控技术股份有限公司)

AI总结 本文提出了解耦世界模型,通过离线到在线的潜在蒸馏和灵活解耦约束,从干扰视频中学习语义知识,提升强化学习的样本效率。

Comments Accepted by ICCV 2025. Project page: https://qiwang067.github.io/diswm

详情

展开后加载摘要…

URL PDF HTML 收藏