Bootstrapping Physics-Grounded Video Generation through VLM-Guided Iterative Self-Refinement
通过VLM引导的迭代自优化提升物理导向的视频生成
机构 * School of Computer Science and Technology, University of Chinese Academy of Sciences(中国科学院大学计算机科学与技术学院) ; School of Computer Science and Technology, Beijing Institute of Technology(北京理工大学计算机科学与技术学院) ; Institute of Information Engineering, Chinese Academy of Sciences(中国科学院信息工程研究所) ; School of Cyber Security, University of Chinese Academy of Sciences(中国科学院大学网络安全学院) ; Institute of Computing Technology, Chinese Academy of Sciences(中国科学院计算技术研究所)
AI总结 本文提出一种通过VLM引导的迭代自优化方法,提升视频生成的物理一致性,实验显示在PhyIQ基准上得分提升明显。
Comments ICCV 2025 Physics-IQ Challenge Third Place Solution