Grounded Visual Factualization: Factual Anchor-Based Finetuning for Enhancing MLLM Factual Consistency
机构 * University of Padua(帕多瓦大学)
专题命中 多模态训练与对齐 :MLLM(title,abstract);multimodal(abstract);分类 cs.CV、cs.CL
AI 大模型
跨文本、图像、视频、音频等模态的大模型与学习方法。
机构 * University of Padua(帕多瓦大学)
专题命中 多模态训练与对齐 :MLLM(title,abstract);multimodal(abstract);分类 cs.CV、cs.CL
机构 * Department of Computing, The Hong Kong Polytechnic University(计算系,香港理工大学) ; Research Institute of Multiple Agents and Embodied Intelligence, Pengcheng Laboratory(多智能体与具身智能研究院,鹏城实验室)
专题命中 多模态训练与对齐 :multimodal(title,abstract);分类 cs.CV、cs.CL、cs.AI
Comments Accepted to IEEE TMM
机构 * Tissue Image Analytics Centre, Department of Computer Science, University of Warwick, UK(沃里克大学计算机科学系组织图像分析中心)
专题命中 多模态训练与对齐 :multimodal(title,abstract);分类 cs.CV
Comments 5 pages, 1 figure, 4 tables
专题命中 多模态训练与对齐 :multimodal(title,abstract);分类 cs.MM
专题命中 多模态训练与对齐 :multimodal(title,abstract)
专题命中 多模态训练与对齐 :multimodal(title,abstract)
Comments Accepted by AAAI 2026
专题命中 多模态训练与对齐 :multimodal(abstract);cross-modal(abstract);分类 cs.CV、cs.CL
机构 * School of Computer Science, Northwestern Polytechnical University(西北工业大学计算机学院)
专题命中 多模态训练与对齐 :multimodal(abstract);cross-modal(abstract);分类 cs.CV
专题命中 多模态训练与对齐 :cross-modal(abstract);分类 cs.CV
Comments 21pages,12 figures,published to AAAI 2026
机构 * HKUST(香港科技大学) ; NTU(国立台湾大学) ; SYSU(南方科技大学) ; NUS(国立新加坡大学) ; Alibaba Group(阿里巴巴集团)
专题命中 多模态训练与对齐 :multimodal(abstract);分类 cs.CV
Comments Project Page: https://livioni.github.io/OmniVGGT-official/
机构 * School of Optical-Electrical and Computer Engineering, University of Shanghai for Science and Technology(光学电子与计算机工程学院,上海科学技术大学) ; School of Automotive Studies, Tongji University(汽车学院,同济大学) ; Momoni AI ; College of Science and Engineering, James Cook University(科学与工程学院,詹姆斯库克大学) ; School of Engineering, Swinburne University of Technology(工程学院,斯威本技术大学) ; School of Electrical and Electronic Engineering, Nanyang Technological University(电气与电子工程学院,南洋理工大学)
专题命中 多模态训练与对齐 :cross-modal(abstract);分类 cs.CV
Comments 8 pages, 5 figures
专题命中 多模态训练与对齐 :multimodal(abstract)
Comments This is the extended version with technical appendices. The version of record appears in AAAI-26. Please cite the AAAI version