Accelerating Multimodal Large Language Models via Dynamic Visual-Token Exit and the Empirical Findings
机构 * Key Laboratory of Multimedia Trusted Perception and Efficient Computing, Ministry of Education of China, Xiamen University(中国教育部多媒体可信感知与高效计算重点实验室,厦门大学) ; Institute of Artificial Intelligence, Xiamen University(厦门大学人工智能研究院)
专题命中 多模态训练与对齐 :multimodal(title,abstract);分类 cs.CV、cs.CL、cs.MM