X-MULTI: VLM-based Imaging Factor Disentanglement for Factor-Aware Image Synthesis
X-MULTI:基于VLM的成像因子解耦用于因子感知图像合成
Sonali Godavarthy, Matthias Neuwirth-Trapp, Tim-Felix Faasch, Maarten Bieshaar, Michael Moeller, Kristof Van Laerhoven, Danda Pani Paudel
机构
*
University of Siegen(锡根大学)
;
ETH Zürich(苏黎世联邦理工学院)
;
Bosch Research(博世研究中心)
;
INSAIT(INSAIT(保加利亚的智能与数据科学研究所))
;
Sofia University “St. Kliment Ohridski”(索非亚大学“圣·克利门特·奥赫里德斯基”)
CommentsPaper accepted to Workshop on Human-Centered Multimodal Intelligence in the Wild (HCMIW) in European Conference on Computer Vision (ECCV) 2026; 18 pages, 3 figures, 7 tables. Project webpage at https://apicis.github.io/aff-sheet
Human-Centric Intelligence in the Era of Foundation Models: A Survey
基础模型时代的以人为中心的智能:一项综述
Yang Chen, Tianqi Wang, Xiaorui Jiang, Yilei Man, Yihua Shao, Mengyuan Liu, Zhi Chen, Xiaofeng Cao, Qibin Zhao, Chi Harold Liu, Albert Y. Zomaya, Nicu Sebe, Jingren Zhou, Dacheng Tao, Song Guo, Jingcai Guo
机构
*
The Hong Kong Polytechnic University(香港理工大学)
;
Peking University(北京大学)
;
University of Southern Queensland(南昆士兰大学)
;
Tongji University(同济大学)
;
RIKEN Center for Advanced Intelligence Project(理化学研究所先进智能项目中心)
;
Beijing Institute of Technology(北京理工大学)
;
The University of Sydney(悉尼大学)
;
University of Trento(特伦托大学)
;
Alibaba Group(阿里巴巴集团)
;
Nanyang Technological University(南洋理工大学)
;
Hong Kong University of Science and Technology(香港科技大学)
机构
*
Sogang University(西江大学)
;
KAIST(韩国科学技术院)
;
NYU Shanghai(上海纽约大学)
;
JIUTIAN Research, China Mobile(中国移动九天研究院)
;
The State Key Laboratory of Multimedia Information Processing, Peking University(北京大学多媒体信息处理国家重点实验室)
;
China Mobile (Hong Kong) Innovation Research Institute(中国移动(香港)创新研究院)
;
Central Conservatory of Music(中央音乐学院)