Discrete Diffusion Models with MLLMs for Unified Medical Multimodal Generation
专题命中 多模态生成 :multimodal(title,abstract);MLLM(abstract);cross-modal(abstract);分类 cs.CV、cs.AI
Comments 16 pages,6 figures
AI 大模型
跨文本、图像、视频、音频等模态的大模型与学习方法。
专题命中 多模态生成 :multimodal(title,abstract);MLLM(abstract);cross-modal(abstract);分类 cs.CV、cs.AI
Comments 16 pages,6 figures
机构 * eHealth Center, Faculty of Computer Science, Multimedia and Telecommunications, Universitat Oberta de Catalunya(eHealth中心,计算机科学、多媒体与电信学院,开放大学)
专题命中 多模态生成 :multimodal(title,abstract);分类 cs.CV、cs.MM
专题命中 多模态生成 :multimodal(abstract);MLLM(abstract);分类 cs.CV
Comments 19 pages, 8 figures
机构 * Meta Superintelligence Labs(Meta超智能实验室) ; Meta FAIR ; Cornell University(康奈尔大学) ; Stony Brook University(石溪大学)
专题命中 多模态生成 :multimodal(abstract);分类 cs.CV、cs.CL、cs.AI
机构 * Lehigh University(莱维大学)
专题命中 多模态生成 :multimodal(abstract);分类 cs.CV、cs.AI
专题命中 多模态生成 :multi-modal(abstract);分类 cs.CL、cs.AI
Comments Need to be revised
机构 * Meituan Inc(美团公司) ; Shanghai Jiao Tong University(上海交通大学)
专题命中 多模态生成 :multimodal(abstract);分类 cs.CV
Comments technical report, project url:https://onecat-ai.github.io/