RISE-T2V: Rephrasing and Injecting Semantics with LLM for Expansive Text-to-Video Generation
专题命中 可控生成 :diffusion(abstract);分类 cs.CV
Comments 17 pages, 16 figures
视觉与机器人
图像生成、文生图、图像编辑、扩散模型和可控生成。
专题命中 可控生成 :diffusion(abstract);分类 cs.CV
Comments 17 pages, 16 figures
机构 * State Key Laboratory for Novel Software Technology, Nanjing University(南京大学新型软件技术国家重点实验室) ; National University of Singapore(新加坡国立大学)
专题命中 可控生成 :diffusion(abstract);分类 cs.CV
机构 * School of Computer Science, Shanghai Jiao Tong University(上海交通大学计算机科学学院) ; Guangdong Laboratory of Artificial Intelligence and Digital Economy (SZ)(广东省人工智能与数字经济实验室(深圳))
专题命中 可控生成 :diffusion(abstract);分类 cs.CV
Comments Accepted by NeurIPS 2025
机构 * Nanjing University of Posts and Telecommunications(南京邮电大学) ; Peng Cheng Laboratory(鹏城实验室)
专题命中 可控生成 :image generation(abstract);分类 cs.CV
Comments Accepted by ACM MM 2025
专题命中 图像生成评测 :diffusion(abstract);分类 cs.CV
机构 * School of Computer Science, Shanghai Jiao Tong University(上海交通大学计算机科学学院) ; Zhiyuan College, Shanghai Jiao Tong University(上海交通大学紫阳学院) ; China Mobile Research Institute(中国移动研究院) ; Westlake University(西湖大学) ; Huawei Consumer Business Group(华为消费者业务集团)
专题命中 效率与蒸馏 :diffusion(title,abstract);分类 cs.CV
Comments Accepted to NeurIPS 2025. Code is available at: https://github.com/zhengchen1999/DOVE
机构 * Korea Advanced Institute of Science and Technology (KAIST)(韩国科学技术院)
专题命中 效率与蒸馏 :diffusion(title,abstract)
专题命中 效率与蒸馏 :diffusion(abstract)