Lego-Edit: A General Image Editing Framework with Model-Level Bricks and MLLM Builder
机构 * Xiaomi Corporation Beijing, China(小米公司北京)
专题命中 多模态生成 :MLLM(title,abstract);multi-modal(abstract);分类 cs.CV
AI 大模型
跨文本、图像、视频、音频等模态的大模型与学习方法。
机构 * Xiaomi Corporation Beijing, China(小米公司北京)
专题命中 多模态生成 :MLLM(title,abstract);multi-modal(abstract);分类 cs.CV
机构 * LiAuto Inc(LiAuto公司) ; Institute of Computing Technology, Chinese Academy of Sciences(中国科学院计算技术研究所) ; School of Vehicle and Mobility, Tsinghua University(清华大学车辆与移动研究所)
专题命中 多模态生成 :multi-modal(title,abstract);分类 cs.CV
Comments Under review
机构 * Southern University of Science and Technology(南方科技大学) ; Pengcheng Laboratory(鹏城实验室)
专题命中 多模态生成 :multimodal(abstract);分类 cs.CV、cs.AI
机构 * University of Amsterdam(阿姆斯特丹大学)
专题命中 多模态生成 :multi-modal(abstract);分类 cs.CV、cs.AI
Comments The paper is accepted by the Conference on Information and Knowledge Management (CIKM), 2025
机构 * School of Computer Science, Sichuan University(四川大学计算机学院) ; School of Cyber Science and Engineering, Sichuan University(四川大学网络科学与工程学院) ; West China Biomedical Big Data Center, Sichuan University West China Hospital(四川大学华西生物医学大数据中心)
专题命中 多模态生成 :multimodal(abstract);分类 cs.AI