AnyGPT: Unified Multimodal LLM with Discrete Sequence Modeling
机构 * Fudan University(复旦大学) ; Multimodal Art Projection Research Community(多模态艺术投影研究社区) ; Shanghai AI Laboratory(上海人工智能实验室)
专题命中 音频语音多模态 :multimodal(title,abstract);any-to-any(abstract);分类 cs.CV、cs.CL、cs.AI
Comments 28 pages, 16 figures, under review, work in progress