3D4D: An Interactive, Editable, 4D World Model via 3D Video Generation
专题命中 多模态生成 :multi-modal(abstract);分类 cs.CV
Comments Accepted by AAAI 2026 Demo Track
AI 大模型
跨文本、图像、视频、音频等模态的大模型与学习方法。
专题命中 多模态生成 :multi-modal(abstract);分类 cs.CV
Comments Accepted by AAAI 2026 Demo Track
专题命中 多模态生成 :image-text(abstract);分类 cs.CV
Comments Accepted by AAAI2026
机构 * Washington University in St. Louis(华盛顿大学圣路易斯分校)
专题命中 多模态生成 :cross-modal(abstract);分类 cs.CV
专题命中 多模态生成 :multi-modal(abstract);分类 cs.AI
机构 * NEC Laboratories America(NEC美国实验室) ; Lehigh University(莱特大学)
专题命中 多模态生成 :multimodal(abstract);分类 cs.CL
Comments AACL 2025
机构 * College of Computer Science, Nankai University(南开大学计算机科学学院) ; Institute of Artificial Intelligence (TeleAI), China Telecom(中国电信人工智能研究所)
专题命中 多模态生成 :multimodal(abstract);分类 eess.AS
Comments Accepted by AAAI 2026
机构 * School of Electronic and Optical Engineering, Nanjing University of Science and Technology(电子与光学工程学院,南京理工大学)
专题命中 多模态生成 :multimodal(abstract);分类 cs.CV
专题命中 多模态生成 :cross-modal(abstract);分类 cs.CV
Comments 15 pages, 9 figures, 8 tables
机构 * Tianjin University(天津大学) ; University of Trento(特伦托大学)
专题命中 多模态生成 :MLLM(abstract);分类 cs.CV
Comments Accepted by ACMMM2025, Our project webpage: https://tjulcx.github.io/FreeInsert/
机构 * Harbin Institute of Technology(哈尔滨工业大学) ; University of Science and Technology of China(中国科学技术大学) ; Westlake University(西湖大学)
专题命中 多模态生成 :cross-modal(abstract);分类 cs.CV
Comments Webpage: https://petershen-csworld.github.io/FreeBlend
机构 * Department of Computer Science, Utah Valley University(计算机科学系,犹他谷大学) ; School of Education, Utah Valley University(教育学院,犹他谷大学)
专题命中 多模态生成 :multimodal(abstract);分类 cs.AI
机构 * Nanjing University of Posts and Telecommunications(南京邮电大学) ; Peng Cheng Laboratory(鹏城实验室)
专题命中 多模态生成 :image-text(abstract);分类 cs.CV
Comments Accepted by ACM MM 2025
机构 * NLPR, MAIS, Institute of Automation(NLPR、MAIS、自动化研究所) ; School of Artificial Intelligence(人工智能学院) ; University of Chinese Academy of Sciences(中国科学院大学) ; Alibaba Group(阿里巴巴集团) ; Hupan Lab(华普实验室) ; College of Intelligence and Computing(智能与计算学院) ; National University of Singapore(新加坡国立大学) ; Cancer Science Institute of Singapore(新加坡癌症科学研究所)
专题命中 多模态生成 :multi-modal(abstract);分类 cs.AI
机构 * VERSES AI Research Lab(VERSES AI研究实验室) ; University of Tübingen(图宾根大学) ; ELLIS Institute, Tübingen(图宾根ELLIS研究所) ; MPI for Intelligent Systems, Tübingen(图宾根智能系统研究所)
专题命中 多模态生成 :multimodal(abstract);分类 cs.AI
Comments 13 pages, 3 figures, under review for World Modeling Workshop 2026
专题命中 多模态生成 :multimodal(abstract);分类 cs.AI
机构 * The Chinese University of Hong Kong(香港中文大学) ; Tsinghua University(清华大学) ; Snap Research ; Shenzhen University(深圳大学) ; HKUST(香港科技大学) ; University of Central Florida(佛罗里达大学) ; UC Davis(加州大学戴维斯分校) ; Sun Yat-Sen University(孙中山大学)
专题命中 多模态生成 :MLLM(abstract);分类 cs.CV
机构 * School of Informatics, Xiamen University(厦门大学信息学院)
专题命中 多模态生成 :multimodal(abstract);分类 cs.CV
Comments Accepted by ACM MM 2025. Project Page: https://jiajinglin.github.io/Phys4DGen
机构 * Shenzhen MSU-BIT University(深圳MSU-BIT大学) ; Beijing Institude of Technology(北京理工大学)
专题命中 多模态生成 :multimodal(abstract);分类 cs.CV
机构 * Tencent Hunyuan Team(腾讯文言团队)
专题命中 多模态生成 :multi-modal(abstract);分类 cs.AI
机构 * College of Computer Science, Sichuan University(四川大学计算机科学学院) ; School of Computer Science and Technology, Xidian University(西安电子科技大学计算机科学与技术学院) ; Engineering Research Center of Machine Learning and Industry Intelligence, Ministry of Education, Chengdu, China(教育部机器学习与工业智能工程研究中心)
专题命中 多模态生成 :multi-modal(abstract);分类 cs.CV
Journal ref Proceedings of the 33rd ACM International Conference on Multimedia (2025) 3330-3339
专题命中 多模态生成 :multimodal(abstract);分类 cs.CV
Comments Accepted in NeurIPS 2025
机构 * Technical University of Munich(慕尼黑技术大学) ; ETH Zurich(苏黎世联邦理工学院)
专题命中 多模态生成 :multi-modal(abstract);分类 cs.CV
专题命中 多模态生成 :image-text(abstract);分类 cs.AI
Comments Finding unfinished issue in this work , still refining
专题命中 多模态生成 :multimodal(abstract);分类 cs.AI
机构 * The Chinese University of Hong Kong(香港中文大学) ; Tsinghua University(清华大学) ; Huawei Noah’s Ark Lab(华为诺亚实验室) ; Shanghai Artificial Intelligence Laboratory(上海人工智能实验室)
专题命中 多模态生成 :multimodal(abstract);分类 cs.AI
Comments Neurips 2025, 29 pages, 19 figures
机构 * organization= School of Chemical \& Biomolecular Engineering, Georgia Institute of Technology , addressline= 311 Ferst Drive NW , city= Atlanta , postcode= 30332 , state= GA , country= USA ; organization= Department of Chemical Engineering, Carnegie Mellon University , addressline= 5000 Forbes Street , city= Pittsburgh , postcode= 15213 , state= PA , country= USA
专题命中 多模态生成 :multimodal(abstract);分类 cs.AI
Comments Submitted to "Current Opinion in Chemical Engineering" for peer review
机构 * Tulane University(路易斯安那州立大学)
专题命中 多模态生成 :multimodal(abstract);分类 cs.CV
Comments NeurIPS 2025
机构 * Horizon Robotics ; University of Hong Kong(香港大学) ; University of the Chinese Academy of Sciences(中国科学院大学) ; Nanjing University(南京大学) ; Beijing Jiaotong University(北京交通大学)
专题命中 多模态生成 :multi-modal(abstract);分类 cs.CV
Comments IROS 2025
机构 * Shenzhen University(深圳大学)
专题命中 多模态生成 :image-text(abstract);分类 cs.CV
Comments This work was initially drafted in November 2022
机构 * Cuhk(香港中文大学)
专题命中 多模态生成 :multimodal(abstract);分类 cs.AI