机构
*
National University of Singapore(新加坡国立大学)
;
PuzzleLogic Pte Ltd(拼图逻辑私人有限公司)
;
Harbin Institute of Technology, Shenzhen(哈尔滨工业大学深圳校区)
;
Peking Union Medical College Hospital(北京协和医院)
机构
*
Department of Robotics
;
Mechatronics Engineering, University of Dhaka, Bangladesh
;
Center for Computational \& Data Sciences, Independent University, Bangladesh
专题命中
视觉推理
:VLM(abstract,abstract_cn);vision language model(abstract)
VGI-Bench: Probing Visual Intelligence in Video Generation Models
VGI-BENCH:探究视频生成模型的视觉智能
Xuan He, Cong Wei, Yuhao Cheng, Linrui Ma, Yuxuan Zhang, Zuojun Li, Yuhao Wen, Jize Jiang, Zeyi Liu, Yuren Hao, Songcheng Cai, Keming Wu, Penghui Du, Kai Zou, Rui Yang, Chenkai Sun, Ke Yang, Ping Nie, Kelsey R Allen, Chenglong Wang, Michel Galley, Jianfeng Gao, ChengXiang Zhai
机构
*
University of Illinois Urbana Champaign(伊利诺伊大学厄巴纳-香槟分校)
;
Tsinghua University(清华大学)
;
University of Waterloo(滑铁卢大学)
;
Massachusetts Institute of Technology(麻省理工学院)
;
University of British Columbia(不列颠哥伦比亚大学)
;
Vector Institute(矢量研究所)
;
Microsoft Research(微软研究院)
;
Etude AI
Beyond RGB: Benchmarking and Enhancing MLLMs for Hyperspectral Image Understanding via Training-Free Reasoning Framework
HM-Bench:多模态大语言模型在高光谱遥感中的综合基准
Xinyu Zhang, Zurong Mai, Qingmei Li, Xiaoya Fan, Zjin Liao, Haoyuan Liang, Yibin Wen, Yuhang Chen, Chan Tsz Ho, Bi Tianyuan, Ruifeng Su, Zihao Qiang, Juepeng Zheng, Jianxi Huang, Yutong Lu, Haohuan Fu
机构
*
Sun Yat-sen University(中山大学)
;
Tsinghua Shenzhen International Graduate School(清华大学深圳国际研究生院)
;
China Agricultural University(中国农业大学)
;
Southwest Jiaotong University(西南交通大学)
;
Southwest University(西南大学)
;
National Supercomputing Center in Shenzhen(国家超级计算深圳中心)
专题命中
视觉推理
:multimodal large language model(abstract);分类 cs.CV、cs.AI
A Modular Multitask Reasoning Framework Integrating Spatio-temporal Models and LLMs
整合时空模型与大语言模型(LLMs)的模块化多任务推理框架
Kethmi Hirushini Hettige, Jiahao Ji, Cheng Long, Shili Xiang, Gao Cong, Jingyuan Wang
机构
*
College of Computing and Data Science, Nanyang Technological University, Singapore(南洋理工大学计算机与数据科学学院)
;
Institute for Infocomm Research, A*STAR, Singapore(资讯通信研究院)
;
School of Computer Science and Engineering, Beihang University, China(北京航空航天大学计算机科学与工程学院)
HaReCAP: Habitual-action Grounding for Recursive Large Language Model Agents
HaReCAP:面向递归大语言模型智能体的习惯性动作 grounding 方法
Shen Liu, Zhenguo Xu, Shaopu Wang, Yike Gao, Chunlei Wang
机构
*
North China Institute of Computer System Engineering(华北计算机系统工程研究所)
;
University of Science and Technology of China(中国科学技术大学)
;
China Information Security Research Institute Co., Ltd.(中国信息安全研究院有限公司)
ExtrinSplat: Decoupling Geometry and Semantics for Open-Vocabulary Understanding in 3D Gaussian Splatting
ExtrinSplat:解耦几何与语义以实现3D高斯散射中的开放词汇理解
Jiayu Ding, Xinpeng Liu, Zhiyi Pan, Shiqiang Long, Ge Li
机构
*
Guangdong Provincial Key Laboratory of Ultra High Definition Immersive Media Technology, Shenzhen Graduate School, Peking University(广东省超高清沉浸式媒体技术重点实验室,北京大学深圳研究生院)
;
School of Computer Science and Technology, Tianjin University(天津大学计算机科学与技术学院)
;
Guangdong Bohua UHD Innovation Center Co., Ltd.(广东博华超高清创新中心有限公司)
CommentsTo appear in Proceedings of the 1st International Workshop on Specification-Driven Development Life Cycle (SpecOps 2026), co-located with SPLASH 2026
Qixiang Yin, Huanjin Yao, Yuchen Cai, Jianghao Chen, Ziyi Wang, Min Yang, Fei Su, Zhicheng Zhao
机构
*
Beijing University of Posts and Telecommunications(北京邮电大学)
;
ByteDance(字节跳动)
;
USTC(中国科学技术大学)
;
Beijing Key Laboratory of Network System and Network Culture(北京网络系统与网络文化重点实验室)
;
Key Laboratory of Interactive Technology and Experience System, Ministry of Culture and Tourism(文化和旅游部互动技术与体验系统重点实验室)
;
Zhongguancun Academy(中关村科学城)
CommentsWithdrawn due to serious concerns regarding the authenticity and accuracy of the listed authorship. The identity of one or more listed authors cannot presently be verified, and the author list may not represent distinct contributors. The manuscript is withdrawn pending institutional review
SPADE: Self-Play in Adaptive Synthetic Executable Environments
SPADE:自适应合成可执行环境中的自博弈
Bo Liu, Simon Yu, Yiding Jiang, Ao Qu, Andrew Zhao, Zichen Liu, Junsu Kim, Zijian Zhou, Seungone Kim, Tongzheng Ren, Mickel Liu, Hanfei Yu, Zhaorun Chen, Weiyan Shi, Paul Pu Liang, Luke Zettlemoyer, Yejin Choi, Natasha Jaques
机构
*
University of Washington(华盛顿大学)
;
Stanford University(斯坦福大学)
;
Northeastern University(东北大学)
;
Carnegie Mellon University(卡内基梅隆大学)
;
Massachusetts Institute of Technology(麻省理工学院)
;
National University of Singapore(新加坡国立大学)
;
Seoul National University(首尔大学)
;
Stevens Institute of Technology(史蒂文斯理工学院)
;
University of Chicago(芝加哥大学)