Towards Memory-Efficient Autoregressive Video Generation via Instance-Specific Parametric Absorption
面向内存高效的自回归视频生成:基于实例特定参数吸收
Xiaomeng Fu, Jia Li, Yiming Hu, Yong Wang, Hayden Kwok-Hay So, Jiao Dai, Xiangxiang Chu, Jizhong Han
机构
*
Institute of Information Engineering, Chinese Academy of Sciences(中国科学院信息工程研究所)
;
The University of Hong Kong(香港大学)
;
AMAP, Alibaba Group(阿里巴巴集团高德地图)
机构
*
The Chinese University of Hong Kong(香港中文大学)
;
Joy Future Academy, JD(京东探索研究院)
;
The Hong Kong University of Science and Technology(香港科学与技术大学)
;
Tsinghua University(清华大学)
;
The University of Hong Kong(香港大学)
;
University of Science and Technology of China(中国科学技术大学)
机构
*
Department of Computer Science and Software Engineering, The University of Western Australia(计算机科学与软件工程系,西澳大学)
;
Munich Center for Machine Learning (MCML) and Technical University of Munich (TUM)(慕尼黑机器学习中心(MCML)和技术大学慕尼黑(TUM))
;
School of Information Technology, Murdoch University(信息科技学院,墨尔本大学)
;
Department of Electrical, Electronics and Computer Engineering, The University of Western Australia(电子、电子与计算机工程系,西澳大学)
机构
*
Chinese Academy of Sciences(中国科学院)
;
Microsoft Research(微软研究院)
;
Sun Yat-sen University(中山大学)
;
Zhejiang University(浙江大学)
;
Shanghai Artificial Intelligence Laboratory(上海人工智能实验室)
;
Xi’an Jiaotong University(西安交通大学)
MMPhysVideo: Physically Plausible Video Generation Through Joint RGB-Perception Modeling
MMPhysVideo: 通过联合多模态建模提升视频生成的物理合理性
Shubo Lin, Xuanyang Zhang, Wei Cheng, Weiming Hu, Gang Yu, Jin Gao
机构
*
State Key Laboratory of Multimodal Artificial Intelligence Systems, CASIA(中国科学院自动化研究所多模态人工智能系统国家重点实验室)
;
School of Artificial Intelligence, University of Chinese Academy of Sciences(中国科学院大学人工智能学院)
;
StepFun(阶跃星辰)
;
Beijing Key Laboratory of Super Intelligent Security of Multi-Modal Information(多模态信息超级智能安全北京市重点实验室)
;
School of Information Science and Technology, Shanghai Tech University(上海科技大学信息科学与技术学院)