DreamOmni2: Multimodal Instruction-based Editing and Generation
机构 * CUHK(香港中文大学) ; HKUST(香港科技大学) ; HKU(香港大学) ; ByteDance Inc(字节跳动公司)
专题命中 多模态生成 :multimodal(title,abstract);分类 cs.CV
AI 大模型
跨文本、图像、视频、音频等模态的大模型与学习方法。
机构 * CUHK(香港中文大学) ; HKUST(香港科技大学) ; HKU(香港大学) ; ByteDance Inc(字节跳动公司)
专题命中 多模态生成 :multimodal(title,abstract);分类 cs.CV
机构 * Shanghai AI Laboratory(上海人工智能实验室) ; Shanghai Innovation Institute(上海创新研究院) ; Nanjing University(南京大学) ; The University of Sydney(悉尼大学) ; Shanghai Jiao Tong University(上海交通大学) ; Tsinghua University(清华大学) ; The Chinese University of Hong Kong(香港中文大学)
专题命中 多模态生成 :multi-modal(title,abstract);分类 cs.CV
Comments 33 pages, 13 figures, 10 tables
专题命中 多模态生成 :cross-modal(abstract)
Comments 13 main pages, 5 figures, 2 tables