PhysOmni: Physics-Grounded Multi-Object Scene Generation from a Single Image with Real-Time Interaction
TelePhysics: 从单张图像生成物理一致的多物体场景 with 实时交互
机构 * Fudan University(复旦大学) ; Institute of Artificial Intelligence, China Telecom (TeleAI)(中国电信人工智能研究院(TeleAI))
AI总结 本文提出TelePhysics,一种无需训练的框架,通过整体场景级3D重建将单张图像转换为物理一致且可控的视频。该方法通过统一的空间坐标系统表示完整场景几何,解决物体穿透和对齐模糊问题,实现准确的多物体交互和更丰富的复杂控制类型,从而在保持逼真视觉保真度的同时实现实时物理交互预览。
Comments ACMMM 2026. Project page: https://physomni.github.io/