Déjà Vu: Multilingual LLM Evaluation through the Lens of Machine Translation Evaluation
专题命中 其他VLM :MLLM(abstract);分类 cs.AI
AI 大模型
视觉语言模型、视觉推理、视觉问答、图文理解和视觉 grounding。
专题命中 其他VLM :MLLM(abstract);分类 cs.AI
机构 * School of Computer Science ; Engineering University of New South Wales Sydney, Australia ; Stanford University California, USA ; Computer Science University College London London, UK ; School of Eng. Maths. \& Tech University of Bristol Bristol, UK ; Inst. Logic Language \& Computation University of Amsterdam Amsterdam, NL
专题命中 其他VLM :vision-language model(abstract);分类 cs.AI
Comments Accepted to: 2025 IEEE International Conference on Quantum Artificial Intelligence (QAI), Naples, Italy, Nov 2-5, 2025. This is the authors' accepted manuscript (AAM). An IEEE copyright notice appears on page 1. The final published version will appear in IEEE Xplore; DOI to be added when available
机构 * MERaLiON Team(MERaLiON团队) ; Institute for Infocomm Research (I 2 R), A*STAR, Singapore(信息通信研究所(I2R),A*STAR,新加坡)
专题命中 其他VLM :multimodal large language model(abstract);分类 cs.AI
机构 * College of Computing and Data Science, Nanyang Technological University(计算与数据科学学院,南洋理工大学) ; DAMO Academy, Alibaba Group(阿里达摩院)
专题命中 其他VLM :vision language model(abstract);分类 cs.CV
机构 * College of Engineering & Advanced Computing(工程与高级计算学院)
专题命中 其他VLM :multimodal large language model(abstract);分类 cs.CV
机构 * University of Science, VNU-HCM(越南胡志明市科学大学)
专题命中 其他VLM :vision-language model(abstract);分类 cs.CV
Comments ACM Multimedia 2025
机构 * Shanghai Jiao Tong University(上海交通大学) ; Zhejiang University(浙江大学) ; Westlake University(西湖大学)
专题命中 其他VLM :multimodal large language model(abstract);分类 cs.AI
Comments Our platform is publicly accessible at https://www.tbox.cn/about/model-ranking
机构 * Tianjin University(天津大学) ; Zhejiang University(浙江大学) ; Zhejiang University of Technology(浙江工业大学) ; Hainan University(海南大学)
专题命中 其他VLM :vision-language model(abstract);分类 cs.CV
专题命中 其他VLM :multimodal large language model(abstract);分类 cs.CV
Comments Project Page: https://personavlog-paper.github.io/
专题命中 其他VLM :multimodal large language model(abstract);分类 cs.LG
机构 * The Hong Kong University of Science and Technology(Guangzhou)(香港科技大学(广州)) ; The Hong Kong University of Science and Technology(香港科技大学) ; Shanghai AI Laboratory(上海人工智能实验室)
专题命中 其他VLM :multimodal large language model(abstract);分类 cs.LG
Comments Accepted at EMNLP 2025, Findings
机构 * Alibaba Group(阿里巴巴集团) ; Beijing Electronic Science and Technology Institute(北京电子科技研究所) ; Nanjing University(南京大学) ; Renmin University of China(中国人民大学) ; Northeastern University(东北大学) ; BraneMatrix AI ; Nanyang Technological University(南洋理工大学)
专题命中 其他VLM :vision language model(abstract);分类 cs.AI
Comments Prior-Guided Bimodal Interactive Black-Box Jailbreak Attack for Toxicity Maximization
机构 * State Key Laboratory of Multimodal Artificial Intelligence Systems, Institute of Automation, Chinese Academy of Sciences(多模态人工智能系统国家重点实验室,自动化研究所,中国科学院) ; School of Artificial Intelligence, University of Chinese Academy of Sciences(人工智能学院,中国科学院大学) ; Department of Physics and Technology, UiT The Arctic University of Norway(物理与技术系,UiT 北极大学)
专题命中 其他VLM :vision-language model(abstract);分类 cs.CV
Comments Accepted by EMNLP 2025 Findings
机构 * Argonne National Laboratory(阿贡国家实验室) ; Cerebras(Cerebras公司) ; Pacific Northwest National Laboratory(太平洋西北国家实验室)
专题命中 其他VLM :vision language model(abstract);分类 cs.LG
Comments Preprint
机构 * Department of Artificial Intelligence, Yonsei University(人工智能系,延世大学) ; Graduate School of Information, Yonsei University(信息研究生院,延世大学)
专题命中 其他VLM :multimodal large language model(abstract);分类 cs.CV
Comments Accepted by ICCV 2025
机构 * School of Computer Science, Wuhan University(武汉大学计算机学院) ; Amap, Alibaba Group(阿里巴巴集团阿里的)
专题命中 其他VLM :vision-language model(abstract);分类 cs.CV
机构 * University of Udine(乌迪大学) ; University of Naples Federico II(那不勒斯费德里科二世大学)
专题命中 其他VLM :vision-language model(abstract);分类 cs.CV
Comments Accepted for publication at the 23rd International Conference on Image Analysis and Processing (ICIAP 2025)
机构 * University of Maryland, College Park(马里兰大学学院市分校) ; DEVCOM Army Research Laboratory(陆军研究实验室)
专题命中 其他VLM :multimodal large language model(abstract);分类 cs.CV
Comments ICCV 2025
机构 * Portland State University(波特兰州立大学)
专题命中 其他VLM :vision-language model(abstract);分类 cs.CV
Comments ICCV 2025 AI4VA workshop (oral), Code: https://github.com/abhijay9/wpclip
机构 * Henan Polytechnic University(河南理工大学)
专题命中 其他VLM :MLLM(abstract);分类 cs.CV
机构 * Shanghai Artificial Intelligence Laboratory(上海人工智能实验室) ; ZheJiang University(浙江大学) ; The Chinese University of Hong Kong(香港中文大学) ; Center for Earth System Modeling and Prediction of China Meteorological Administration(中国气象局地球系统模拟与预测中心)
专题命中 其他VLM :MLLM(abstract);分类 cs.AI
机构 * Virginia Tech(弗吉尼亚理工大学)
专题命中 其他VLM :vision language model(abstract);分类 cs.CV
Comments 5 pages, 2 figures, 5 tables, accepted in CIKM 2025
机构 * Mediterranean Agronomic Institute of Montpellier - CIHEAM-IAMM(地中海农业研究院-CIHEAM-IAMM) ; Inria(法国国家信息与自动化技术研究院) ; INRAE(法国国家农业研究咨询中心) ; Cirad(国际热带农业研究中心) ; UMR TETIS(TETIS联合研究单位) ; Univ. of Montpellier(蒙彼利埃大学)
专题命中 其他VLM :vision-language model(abstract);分类 cs.CV
Comments Accepted at WACV'25
专题命中 其他VLM :MLLM(abstract);分类 cs.CV
专题命中 其他VLM :multimodal large language model(abstract);分类 cs.CV
机构 * Carnegie Mellon University(卡内基梅隆大学)
专题命中 其他VLM :multimodal large language model(abstract);分类 cs.CV
机构 * Mercari, Inc.(Mercari公司)
专题命中 其他VLM :vision-language model(abstract);分类 cs.AI
Comments 6 pages, KDD 2025 Workshop on Two-sided Marketplace Optimization: Search, Pricing, Matching & Growth (TSMO)
机构 * Beijing University of Post and Telecommunication(北京邮电大学)
专题命中 其他VLM :vision-language model(abstract);分类 cs.CV
Comments 8 pages, 2 supplement pages, 3 figures, ECAI2025
机构 * Dali University(大理大学) ; Bandırma Onyedi Eylül University(班迪尔马第十七个九月大学)
专题命中 其他VLM :vision-language model(abstract);分类 cs.LG
机构 * New Laboratory of Pattern Recognition, Institute of Automation, Chinese Academy of Sciences(模式识别新实验室,自动化研究所,中国科学院)
专题命中 其他VLM :visual language model(abstract);分类 cs.CV
Comments accepted by ICPR2024