LFTR: Learning-Free Token Reduction for Multimodal Large Language Models
专题命中 多模态训练与对齐 :multimodal(title,abstract);MLLM(abstract);分类 cs.CV、cs.CL、cs.AI
AI 大模型
跨文本、图像、视频、音频等模态的大模型与学习方法。
专题命中 多模态训练与对齐 :multimodal(title,abstract);MLLM(abstract);分类 cs.CV、cs.CL、cs.AI
机构 * University of Bristol(布里斯托大学) ; Brown University(布朗大学) ; South China University of Technology(华南理工大学)
专题命中 多模态训练与对齐 :multimodal(title,abstract);cross-modal(abstract);分类 cs.CV、cs.CL
Comments Accepted by conference EMNLP2025
机构 * Northwestern University(西北大学)
专题命中 多模态训练与对齐 :multimodal(title,abstract);cross-modal(abstract)
Comments Accepted to Neurips 2025 (Spotlight)
机构 * FnGuide Inc.(FnGuide公司) ; Safe Generative AI Lab, MODULABS(MODULABS安全生成AI实验室) ; A.I.MATICS Inc.(A.I.MATICS公司) ; Ewha Womans University(成均馆大学)
专题命中 多模态训练与对齐 :multi-modal(title,abstract);分类 cs.AI
Comments Accept to ACLW 2025 (WOAH); fix typo
Journal ref ACL Workshop 2025
机构 * CISPA Helmholtz Center for Information Security(CISPA赫尔姆霍茨信息安全中心) ; Nokia Bell Labs(诺基亚贝尔实验室)
专题命中 多模态训练与对齐 :multimodal(title,abstract)
机构 * Meta Superintelligence Labs(Meta 超智能实验室) ; University of Oxford(牛津大学)
专题命中 多模态训练与对齐 :multimodal(abstract);MLLM(abstract);分类 cs.CV、cs.AI、cs.MM
Comments Project page: https://junlinhan.github.io/projects/lsbs/
专题命中 多模态训练与对齐 :multimodal(title)
机构 * School of Earth and Space Sciences, Peking University(地球与空间科学学院,北京大学) ; College of Urban and Environmental Sciences, Peking University(城市与环境科学学院,北京大学) ; State Key Laboratory of Multimodal Artificial Intelligence Systems Institute of Automation, CAS(多模态人工智能系统国家重点实验室,中国科学院自动化研究所) ; Intelligent Maintenance and Operations Systems Lab(智能维护与运营系统实验室)
专题命中 多模态训练与对齐 :multimodal(abstract);cross-modal(abstract);分类 cs.CV
Comments NeurIPS 2025
机构 * Zhejiang University(浙江大学) ; Huawei Noah’s Ark Lab(华为诺亚实验室)
专题命中 多模态训练与对齐 :multimodal(abstract);MLLM(abstract)
机构 * IDLab - Faculty of Applied Engineering, University of Antwerp - imec(IDLab - 应用工程学院,安特卫普大学 - imec) ; Cosys-Lab - Faculty of Applied Engineering, University of Antwerp(Cosys-Lab - 应用工程学院,安特卫普大学) ; Flanders Make Strategic Research Centre(弗拉芒制造战略研究中心)
专题命中 多模态训练与对齐 :multi-modal(abstract);分类 cs.CV、cs.AI
专题命中 多模态训练与对齐 :multimodal(abstract);分类 cs.CV
Comments 6 pages, 3 figures
机构 * the BRain and Artificial INtelligence Lab (BRAIN LAB), School of Automation, Northwestern Polytechnical University(人工智能实验室(BRAIN LAB)、自动化学院、西北工业大学) ; the Shenzhen Research Institute of Northwestern Polytechnical University(西北工业大学深圳研究院) ; the National Key Laboratory for Multimedia Information Processing, School of Computer Science, Peking University(多媒体信息处理国家重点实验室、计算机科学学院、北京大学) ; the National Key Laboratory of Microwave Imaging Technology, Chinese Academy of Sciences(微波成像技术国家重点实验室、中国科学院) ; the Aerospace Information Research Institute, Chinese Academy of Sciences(航空信息研究所、中国科学院)
专题命中 多模态训练与对齐 :multimodal(abstract);分类 cs.CV
Comments The manuscript is 15 pages long, includes 13 figures and 5 tables