CommentsThis article is withdrawn because the experimental results and analysis require substantial revision. The current version should not be cited as a reliable representation of the work
机构
*
Mohamed bin Zayed University of Artificial Intelligence(穆罕默德·本·扎耶德人工智能大学)
;
University of Illinois at Urbana–Champaign(伊利诺伊大学厄巴纳-香槟分校)
;
University of Maryland, College Park(马里兰大学学院公园分校)
;
University of California, San Diego(加州大学圣地亚哥分校)
RosettaSearch: Multi-Objective Inference-Time Search for Protein Sequence Design
RosettaSearch:蛋白质序列设计的多目标推理时间搜索
Meghana Kshirsagar, Allen Nie, Ching-An Cheng, Fanglei Xue, Rahul Dodhia, Juan Lavista Ferres, Kevin K. Yang, Frank DiMaio
机构
*
AI for Good, Microsoft Redmond(微软红mond AI for Good)
;
Google DeepMind(谷歌DeepMind)
;
Google Research(谷歌研究)
;
Institute for Protein Design University of Washington(蛋白质设计研究所华盛顿大学)
;
Microsoft Research New England(微软新英格兰研究)
;
Department of Biochemistry Institute for Protein Design University of Washington(生物化学系蛋白质设计研究所华盛顿大学)
EmoTrans: A Benchmark for Understanding, Reasoning, and Predicting Emotion Transitions in Multimodal LLMs
EmoTrans:一个多模态大语言模型中情感过渡理解、推理和预测的基准测试
He Hu, Tengjin Weng, Zebang Cheng, Yu Wang, Jiachen Luo, Björn Schuller, Zheng Lian, Laizhong Cui
机构
*
Shenzhen University(深圳大学)
;
Guangdong Laboratory of Artificial Intelligence(广东人工智能与数字经济实验室)
;
Queen Mary University of London(女王学院伦敦大学)
;
Technical University of Munich(慕尼黑技术大学)
;
Imperial College London(伦敦帝国学院)
;
Tongji University(同济大学)
专题命中
视觉推理
:multimodal large language model(abstract);分类 cs.CV、cs.AI
Zhimu Zhou, Yanpeng Zhao, Qiuyu Liao, Bo Zhao, Xiaojian Ma
机构
*
Shanghai Jiao Tong University(上海交通大学)
;
Renmin University of China(中国人民大学)
;
State Key Laboratory of General Artificial Intelligence, BIGAI(通用人工智能国家重点实验室,BIGAI)
CommentsAccepted by ACL 2026 Main. 17 pages, 7 figures, 8 tables. TL;DR: We propose MM-Mem, a cognition-inspired, dual-trace hierarchical memory framework for long-horizon video understanding grounded in Fuzzy-Trace Theory. It features adaptive memory compression via the Information Bottleneck and employs an entropy-driven top-down retrieval to access fine-grained details only when necessary
Chuanyu Qin, Chenxu Yang, Qingyi Si, Naibin Gu, Dingyu Yao, Zheng Lin, Peng Fu, Nan Duan, Jiaqi Wang
机构
*
Institute of Information Engineering, Chinese Academy of Sciences, Beijing, China(中国科学院信息工程研究所)
;
School of Cyber Security, University of Chinese Academy of Sciences, Beijing, China(中国科学院大学网络安全学院)
Seeing the Intangible: Survey of Image Classification into High-Level and Abstract Categories
看见无形:图像分类到高级和抽象类别的调查
Delfina Sol Martinez Pandiani, Valentina Presutti
机构
*
University of Bologna(博洛尼亚大学)
;
University of Bologna Department of Computer Science(博洛尼亚大学计算机科学系)
;
University of Bologna Department of Modern Languages, Literatures(博洛尼亚大学现代语言文学系)
;
Centrum Wiskunde en Informatica(数学与信息学研究中心)
;
Institute for Clarity in Documentation(文档清晰研究所)
;
Inria Paris-Rocquencourt(巴黎-罗克奎恩特研究所)
;
Rajiv Gandhi University(拉贾·甘地大学)
;
Tsinghua University(清华大学)
;
Palmer Research Laboratories(帕尔默研究实验室)
Free Lunch for Unified Multimodal Models: Enhancing Generation via Reflective Rectification with Inherent Understanding
统一多模态模型的免费午餐:通过内在理解的反思校正增强生成
Yibo Jiang, Tao Wu, Rui Jiang, Yehao Lu, Chaoxiang Cai, Zequn Qin, Xi Li
机构
*
School of Software Technology, Zhejiang University(浙江大学软件技术学院)
;
College of Computer Science and Technology, Zhejiang University(浙江大学计算机科学与技术学院)
;
College of Computer Science(计算机科学学院)
A Study of Failure Modes in Two-Stage Human-Object Interaction Detection
两阶段人-物交互检测中故障模式的研究
Lemeng Wang, Qinqian Lei, Vidhi Bakshi, Daniel Yi, Yifan Liu, Jiacheng Hou, Asher Seng Hao, Zheda Mai, Wei-Lun Chao, Robby T. Tan, Bo Wang
机构
*
The Ohio State University(俄亥俄州立大学)
;
National University of Singapore(新加坡国立大学)
;
Boston University(波士顿大学)
;
Independent Researcher(独立研究者)
;
University of Mississippi(密西西比大学)
CamReasoner: Reinforcing Camera Movement Understanding via Structured Spatial Reasoning
CamReasoner:通过结构化空间推理强化相机运动理解
Hang Wu, Yujun Cai, Zehao Li, Haonan Ge, Bowen Sun, Junsong Yuan, Yiwei Wang
机构
*
University of California, Merced(加州大学梅尔德分校)
;
University of Queensland(昆士兰大学)
;
Ant Group(蚂蚁集团)
;
University of Chinese Academy of Sciences(中国科学院大学)
;
University at Buffalo, State University of New York(纽约州立大学布法罗分校)
From UAV Imagery to Agronomic Reasoning: A Multimodal LLM Benchmark for Plant Phenotyping
从无人机图像到农学推理:一种多模态大语言模型基准用于植物表型分析
Yu Wu, Guangzeng Han, Ibra Niang Niang, Francia Ravelombola, Maiara Oliveira, Jason Davis, Dong Chen, Feng Lin, Xiaolei Huang
机构
*
University of Memphis(孟菲斯大学)
;
University of Missouri(密苏里大学)
;
University of Arkansas Division of Agriculture(阿肯色大学农业分部)
;
Mississippi State University(密西西比州立大学)