机构
*
Harbin Institute of Technology (Shenzhen)(哈尔滨工业大学(深圳))
;
Microsoft Research(微软研究院)
;
National University of Singapore(新加坡国立大学)
;
Technical University of Dresden(德累斯顿技术大学)
Thinking in Frames: How Visual Context and Test-Time Scaling Empower Video Reasoning
基于框架的思考:视觉上下文和测试时扩展如何增强视频推理
Chengzu Li, Zanyi Wang, Jiaang Li, Yi Xu, Han Zhou, Huanyu Zhang, Ruichuan An, Dengyang Jiang, Zhaochong An, Ivan Vulić, Serge Belongie, Anna Korhonen
机构
*
University of Cambridge(剑桥大学)
;
Pioneer Center for AI, University of Copenhagen(人工智能先锋中心,哥本哈根大学)
;
Hong Kong University of Science(香港科学大学)
;
University of California San Diego(加州大学圣地亚哥分校)
;
Institute of Automation, Chinese Academy of Sciences(中国科学院自动化研究所)
;
Peking University(北京大学)
机构
*
National University of Singapore(新加坡国立大学)
;
Stanford University(斯坦福大学)
;
Peking University(北京大学)
;
UniMelb(墨尔本大学)
;
University of Auckland(奥克兰大学)
;
MBZUAI(穆斯林人工智能研究所)
;
University of California, Santa Barbara(加州大学圣芭芭拉分校)
专题命中
视觉推理
:vision language model(abstract);分类 cs.CV
RSGround-R1: Rethinking Remote Sensing Visual Grounding through Spatial Reasoning
RSGround-R1: 重新思考通过空间推理的遥感视觉定位
Shiqi Huang, Shuting He, Bihan Wen
机构
*
School of Electrical and Electronic Engineering, Nanyang Technological University(电气电子工程学院,南洋理工大学)
;
MoE Key Laboratory of Interdisciplinary Research of Computation and Economics, Shanghai University of Finance and Economics(教育部计算与经济交叉学科重点实验室,上海财经大学)
专题命中
视觉定位与Grounding
:grounding(title,abstract);multimodal large language model(abstract);分类 cs.CV
机构
*
EPIC Lab, SJTU(SJTU实验室)
;
Shanghai Jiao Tong University(上海交通大学)
;
National University of Singapore(新加坡国立大学)
;
Duke University(杜克大学)
;
The University of Chicago(芝加哥大学)
Wikontic: Constructing Wikidata-Aligned, Ontology-Aware Knowledge Graphs with Large Language Models
Wikontic: 构建与维基数据对齐、本体感知的知识图谱
Alla Chepurova, Aydar Bulatov, Mikhail Burtsev, Yuri Kuratov
机构
*
Cognitive AI Systems Lab(认知人工智能系统实验室)
;
Moscow Independent Research Institute of Artificial Intelligence(莫斯科独立人工智能研究所)
;
London Institute for Mathematical Sciences(伦敦数学科学研究所)
机构
*
Organizational Management Department, School of Management, Xi’an Jiaotong University(管理学院组织管理部,西安交通大学)
;
West China Longquan Hospital, Sichuan University(四川大学西部临床医学院)
;
School of Electronic Science and Engineering, Xi’an Jiaotong University(西安交通大学电子科学与工程学院)
;
Systems Engineering Institute, Xi’an Jiaotong University(西安交通大学系统工程研究院)
;
Institute of Medical Artificial Intelligence, the Second Affiliated Hospital of Xi’an Jiaotong University(西安交通大学第二附属医院医学人工智能研究所)
;
School of Human Settlements and Civil Engineering, Xi’an Jiaotong University(西安交通大学人居环境与土木工程学院)
;
School of Life Science and Technology, Xi’an Jiaotong University(西安交通大学生命科学与技术学院)
KID: Knowledge-Injected Dual-Head Learning for Knowledge-Grounded Harmful Meme Detection
KID: 基于知识注入的双头学习用于知识引导的有害迷因检测
Yaocong Li, Leihan Zhang, Le Zhang, Qiang Yan
机构
*
School of Economics and Management, Beijing University of Posts and Telecommunications(经济管理学院,北京邮电大学)
;
College of Computing, Beijing Information Science and Technology University(计算机学院,北京信息科技大学)
Manuel Benavent-Lledo, David Mulero-Pérez, David Ortiz-Perez, Jose Garcia-Rodriguez
机构
*
Department of Computer Technology, University of Alicante(阿拉维大学计算机技术系)
;
ValgrAI - Valencian Graduate School and Research Network of Artificial Intelligence(瓦伦西亚人工智能研究生学校和研究网络)
;
Institute of Informatics Research, University of Alicante(阿拉维大学信息研究所)
Faster Predictive Coding Networks via Better Initialization
通过更好的初始化提升预测编码网络的速度
Luca Pinchetti, Simon Frieder, Thomas Lukasiewicz, Tommaso Salvatori
机构
*
Department of Computer Science, University of Oxford(牛津大学计算机科学系)
;
Institute of Logic and Computation, Vienna University of Technology(维也纳技术大学逻辑与计算研究所)
;
VERSES AI Research Lab(VERSES AI研究实验室)
Concept Component Analysis: A Principled Approach for Concept Extraction in LLMs
概念成分分析:一种用于大语言模型中概念提取的原则性方法
Yuhang Liu, Erdun Gao, Dong Gong, Anton van den Hengel, Javen Qinfeng Shi
机构
*
Australian Institute for Machine Learning, The University of Adelaide(澳大利亚机器学习研究所,阿德莱德大学)
;
School of Computer Science and Engineering, The University of New South Wales(计算机科学与工程学院,新南威尔士大学)
机构
*
State Key Laboratory of Multimodal Artificial Intelligence Systems, Institute of Automation, Chinese Academy of Sciences(多模态人工智能系统国家重点实验室,自动化研究所,中国科学院)
;
School of Artificial Intelligence, University of Chinese Academy of Sciences(中国科学院大学人工智能学院)
;
The Grainger College of Engineering, University of Illinois Urbana-Champaign(伊利诺伊大学厄巴纳-香槟分校格拉inger工程学院)
;
CAS Center for Excellence in Brain Science and Intelligence Technology(中国科学院脑科学与智能技术卓越创新中心)
;
Joint Laboratory of Intelligence Science and Technology, Institute of Systems Engineering, Macau University of Science and Technology(澳门科技大学系统工程学院智能科学与技术联合实验室)
Dynamic Topology Awareness: Breaking the Granularity Rigidity in Vision-Language Navigation
动态拓扑感知:突破视觉语言导航中的粒度刚性
Jiankun Peng, Jianyuan Guo, Ying Xu, Yue Liu, Jiashuang Yan, Xuanwei Ye, Houhua Li, Xiaoming Wang
机构
*
Aerospace Information Research Institute, Chinese Academy of Sciences(中国科学院航空航天信息研究所)
;
School of Electronic, Electrical and Communication Engineering, University of Chinese Academy of Sciences(中国科学院大学电子电气与通信工程学院)
;
Department of Computer Science, City University of Hong Kong(香港城市大学计算机科学系)
CommentsCore methods unchanged; title updated and full-text narrative refined for clarity and logical coherence. No changes to key findings and conclusions
When Ads Become Profiles: Uncovering the Invisible Risk of Web Advertising at Scale with LLMs
当广告成为资料:利用LLMs揭示大规模网络广告中的隐形风险
Baiyu Chen, Benjamin Tag, Hao Xue, Daniel Angus, Flora Salim
机构
*
The University of New South Wales(新南威尔士大学)
;
Hong Kong University of Science and Technology (Guangzhou)(香港科技大学(广州))
;
Queensland University of Technology(昆士兰理工大学)
专题命中
幻觉与鲁棒性
:multimodal large language model(abstract);分类 cs.AI