B-VLLM: A Vision Large Language Model with Balanced Spatio-Temporal Tokens
机构 * School of Computer Science, The University of Sydney(悉尼大学计算机科学学院) ; Department of Engineering Science, University of Oxford(牛津大学工程科学系) ; International School of Information Science and Engineering, Dalian University of Technology(大连理工大学信息科学与工程国际学院) ; Advanced Micro Devices(先进微器件公司) ; School of Science, Edith Cowan University(埃迪斯科文大学科学学院)
专题命中 长上下文与记忆 :large language model(title,abstract);language model(title,abstract);分类 cs.AI
Comments Accepted by ICCV2025 (Poster)