PAGE-4D: Disentangled pose and geometry estimation for vggt-4d perception
PAGE-4D: 通过解耦姿态与几何估计实现VGGT-4D感知
机构 * Harvard AI and Robotics Lab, Harvard University(哈佛人工智能与机器人实验室,哈佛大学) ; Media Lab and Electrical Engineering and Computer Science, Massachusetts Institute of Technology(媒体实验室和电气工程与计算机科学,麻省理工学院) ; Department of Computing, Imperial College London(计算系,帝国理工学院) ; Kempner Institute for the Study of Natural and Artificial Intelligence, Harvard University(哈佛大学自然与人工智能研究学院)
AI总结 提出PAGE-4D,扩展VGGT到动态场景,通过动态感知聚合器解耦静态与动态信息,同时提升相机姿态估计、深度预测和点云重建性能。
Comments ICLR 2026, VGGT-4D, Dynamic VGGT