ScreenCoder: Advancing Visual-to-Code Generation for Front-End Automation via Modular Multimodal Agents
机构 * CUHK(香港中文大学) ; MMLab(多模态实验室) ; ARISE Lab(ARISE实验室)
专题命中 多模态生成 :multimodal(title,abstract);MLLM(abstract);分类 cs.CV
Comments ScreenCoder-v2
AI 大模型
跨文本、图像、视频、音频等模态的大模型与学习方法。
机构 * CUHK(香港中文大学) ; MMLab(多模态实验室) ; ARISE Lab(ARISE实验室)
专题命中 多模态生成 :multimodal(title,abstract);MLLM(abstract);分类 cs.CV
Comments ScreenCoder-v2
机构 * National Taiwan University(国立台湾大学) ; NVIDIA Research(NVIDIA研究)
专题命中 多模态生成 :multimodal(title,abstract);MLLM(abstract);分类 cs.AI
机构 * School of Engineering The University of Edinburgh Edinburgh, UK(工程学院 苏格兰爱丁堡大学)
专题命中 多模态生成 :multimodal(title,abstract);MLLM(abstract)
机构 * University of Illinois Urbana-Champaign(伊利诺伊大学厄巴纳-香槟分校) ; Meta AI ; Rutgers University(罗格斯大学)
专题命中 多模态生成 :multimodal(title,abstract);cross-modal(abstract)
Comments ICML 2025
机构 * AstraZeneca Computational Pathology GmbH(阿斯利康计算病理学 GmbH)
专题命中 多模态生成 :multimodal(title,abstract);分类 cs.CV、cs.AI
专题命中 多模态生成 :multi-modal(abstract);MLLM(abstract);分类 cs.CV、cs.AI
机构 * Gaoling School of Artificial Intelligence, Renmin University of China(中国人民大学耿丽人工智能学院) ; BAAI(北京人工智能研究院)
专题命中 多模态生成 :multimodal(abstract);分类 cs.CL、cs.AI
Comments Working in progress
专题命中 多模态生成 :multimodal(abstract);分类 cs.AI
Comments 9 pages, 5 figures
机构 * Shanghai Artificial Intelligence Laboratory(上海人工智能实验室) ; Zhejiang University(浙江大学) ; Fudan University(复旦大学) ; University of British Columbia(不列颠哥伦比亚大学) ; Tongji University(同济大学) ; The Chinese University of Hong Kong(香港中文大学) ; Shanghai Jiaotong University(上海交通大学) ; Stony Brook University(石溪大学) ; Lingang Laboratory(临港实验室) ; Tsinghua University(清华大学)
专题命中 多模态生成 :multimodal(abstract)