Multimodal LLMs Can Reason about Aesthetics in Zero-Shot
机构 * The Hong Kong Polytechnic University(香港理工大学)
专题命中 多模态生成 :multimodal(title,abstract);分类 cs.CV、cs.CL、cs.AI
Comments ACM MM 2025 Camera Ready
AI 大模型
跨文本、图像、视频、音频等模态的大模型与学习方法。
机构 * The Hong Kong Polytechnic University(香港理工大学)
专题命中 多模态生成 :multimodal(title,abstract);分类 cs.CV、cs.CL、cs.AI
Comments ACM MM 2025 Camera Ready
机构 * ETH Zürich(苏黎世联邦理工学院) ; DisneyResearch | Studios(迪士尼研究室)
专题命中 多模态生成 :multimodal(title,abstract);分类 cs.CV
Comments Added more evaluation since the first version. Accepted to SMI 2025. Computers & Graphics
专题命中 多模态生成 :multimodal(title,abstract);分类 cs.CV
Comments Project Page: https://personavlog-paper.github.io/
专题命中 多模态生成 :multi-modal(title);multimodal(abstract)
专题命中 多模态生成 :multimodal(title)
Comments MICCAI'25
机构 * School of Computer Science and Engineering, South China University of Technology(华南理工大学计算机科学与工程学院) ; South China University of Technology(华南理工大学) ; Guangdong Engineering Center for Large Model and GenAI Technology(广东省大模型与生成式人工智能技术工程中心) ; Guangdong Provincial Key Lab of Computational Intelligence and Cyberspace Information(广东省计算智能与网络信息重点实验室) ; Centre for Smart Health, Hong Kong Polytechnic University(香港理工大学智能健康研究中心) ; CAS-Hong Kong Joint Laboratory for Multimodal Medical Molecular Imaging(中国科学院-香港联合多模态医学分子成像联合实验室) ; School of Electronic and Information Engineering, South China University of Technology(华南理工大学电子与信息学院) ; Pazhou Lab, Guangzhou, Guangdong, China(广州琶洲实验室) ; School of Computing and Information Systems, Singapore Management University(新加坡国立大学计算机与信息系统学院)
专题命中 多模态生成 :multimodal(abstract);分类 cs.AI、cs.MM
机构 * ONE Lab, HUST(华中科技大学 ONE 实验室) ; ONE Lab, HUST University of Maryland(华中科技大学 与 马里兰大学 ONE 实验室) ; University of Washington(华盛顿大学) ; Zhejiang University(浙江大学)
专题命中 多模态生成 :multi-modal(abstract);分类 cs.CV、cs.CL
Comments Technical Report
机构 * Technical University of Munich(慕尼黑技术大学) ; Munich Center for Machine Learning (MCML)(慕尼黑机器学习中心) ; ETH Zurich(苏黎世联邦理工学院) ; MBZUAI(穆桑比克大学人工智能研究所)
专题命中 多模态生成 :multimodal(abstract);分类 cs.CV
Comments 13 pages
机构 * Beijing Jiaotong University(北京交通大学) ; Ant Group(蚂蚁集团) ; Qinghai University(青海大学) ; Tsinghua University(清华大学)
专题命中 多模态生成 :multimodal(abstract);分类 cs.CV
Comments 14 pages, 11 figures
机构 * Consultant, high-throughput microscopy and hardware-accelerated algorithms(咨询顾问,高通量显微镜和硬件加速算法)
专题命中 多模态生成 :multi-modal(abstract);分类 cs.CL
Comments Minor changes: resolve HTML rendering issues of sideways tables; Code listing in dark mode. Cite three more journal articles
机构 * Stanford University(斯坦福大学) ; Adobe Research(Adobe研究)
专题命中 多模态生成 :multi-modal(abstract);分类 cs.CV
Comments ICML 2025. Code: https://github.com/Lakonik/GMFlow
机构 * Department of Computer Science, North Carolina State University(计算机科学系,北卡罗来纳州立大学) ; Department of Computer Science and Engineering, University of Louisville(计算机科学与工程系,路易斯维尔大学) ; Department of Electrical and Computer Engineering and Frost Institute for Data Science and Computing, University of Miami(电气与计算机工程系及弗罗斯特数据科学与计算研究所,迈阿密大学)
专题命中 多模态生成 :multi-modal(abstract)
Comments Accepted by IEEE MobiWac 2025
专题命中 多模态生成 :multi-modal(abstract)
机构 * University of Washington(华盛顿大学) ; UC San Diego(圣地亚哥大学) ; Nvidia(英伟达)
专题命中 多模态生成 :multi-modal(abstract)
专题命中 多模态生成 :multi-modal(abstract)
Comments 69 pages, 5 tables, 5 figures
专题命中 多模态生成 :multimodal(abstract)
机构 * The Hong Kong University of Science and Technology(Guangzhou)(香港科技大学(广州)) ; The Hong Kong University of Science and Technology(香港科技大学) ; Shanghai AI Laboratory(上海人工智能实验室)
专题命中 多模态生成 :multimodal(abstract)
Comments Accepted at EMNLP 2025, Findings
专题命中 多模态生成 :multimodal(abstract)
Comments Accepted in IEEE COINS 2025
Journal ref 2025 IEEE International Conference on Omni-layer Intelligent Systems (COINS)