Improving Reasoning Capabilities in Small Models through Mixture-of-Layers Distillation with Stepwise Attention on Key Information
通过关键信息的分步注意力混合层蒸馏提升小模型的推理能力
Yao Chen, Jiawei Sheng, Wenyuan Zhang, Tingwen Liu
机构
*
Institute of Information Engineering, Chinese Academy of Sciences(中国科学院信息工程研究所)
;
School of Cyber Security, University of Chinese Academy of Sciences(中国科学院大学网络安全学院)
CommentsThis preprint is withdrawn due to significant errors in the emergent geometric isomorphism results that necessitate full rewriting, coupled with unresolved author disagreement on authorship. A corrected and revised manuscript will be released separately
机构
*
Shanghai University of Finance and Economics(上海金融学院)
;
Alibaba Group(阿里巴巴集团)
;
Southern University of Science and Technology(南方科技大学)
;
Beihang University(北航)
;
MoE Key Laboratory of Interdisciplinary Research of Computation and Economics(计算与经济交叉学科研究教育部重点实验室)
Journal refProceedings of the 49th International ACM SIGIR Conference on Research and Development in Information Retrieval (SIGIR '26), July 20--24, 2026, Melbourne, VIC, Australia
Language on Demand, Knowledge at Core: Composing LLMs with Encoder-Decoder Translation Models for Extensible Multilinguality
按需语言,知识为核心:通过编码器-解码器翻译模型组成LLM以实现可扩展的多语言性
Mengyu Bu, Yang Feng
机构
*
Key Laboratory of Intelligent Information Processing, Institute of Computing Technology, Chinese Academy of Sciences(智能信息处理重点实验室,计算技术研究所,中国科学院)
;
State Key Laboratory of AI Safety, Institute of Computing Technology, Chinese Academy of Sciences(人工智能安全国家重点实验室,计算技术研究所,中国科学院)
;
University of Chinese Academy of Sciences, Beijing, China(中国科学院大学,北京,中国)
Towards Fine-grained Temporal Perception: Post-Training Large Audio-Language Models with Audio-Side Time Prompt
迈向细粒度时间感知:通过音频侧时间提示进行后训练的大音频-语言模型
Yanfeng Shi, Pengfei Cai, Jun Liu, Qing Gu, Nan Jiang, Lirong Dai, Ian McLoughlin, Yan Song
机构
*
National Engineering Research Center of Speech and Language Information Processing, University of Science and Technology of China, Hefei, China(语音与语言信息处理国家级工程研究中心,中国科学技术大学,合肥,中国)
;
ICT Cluster, Singapore Institute of Technology, Singapore(新加坡理工学院ICT集群,新加坡)
专题命中
其他安全
:alignment(abstract);分类 cs.AI
AI总结
本文提出Audio-Side Time Prompt方法,结合强化学习改进大音频-语言模型的时间感知能力,在音频定位、事件检测等任务中取得显著提升。
EvoSpark: Endogenous Interactive Agent Societies for Unified Long-Horizon Narrative Evolution
EvoSpark:内生交互代理社会的统一长周期叙述进化
Shiyu He, Minchi Kuang, Mengxian Wang, Bin Hu, Tingxiang Gu
机构
*
School of Computer Science and Technology, Xinjiang University(新疆大学计算机科学与技术学院)
;
Department of Precision Instrument, Tsinghua University(清华大学精密仪器系)
机构
*
Department of Mathematics, IIT Delhi(印度德里理工学院数学系)
;
Department of Electrical Engineering, IIT Delhi(印度德里理工学院电气工程系)
;
Department of Computer Science and Engineering, IIIT Manipur(曼尼普尔理工学院计算机科学与工程系)
Human-Centric Topic Modeling with Goal-Prompted Contrastive Learning and Optimal Transport
以人为中心的主题建模:基于目标提示对比学习与最优传输
Rui Wang, Yi Zheng, Dongxin Wang, Haiping Huang, Yuanzhi Yao, Yuxiang Zhou, Jialin Yu, Philip Torr
机构
*
School of Computer Science, Nanjing University of Posts and Telecommunications(南京邮电大学计算机科学学院)
;
School of Computer Science and Information Engineering, Hefei University of Technology(合肥工业大学计算机科学与信息工程学院)
;
School of Electronic Engineering and Computer Science, Queen Mary University of London(伦敦玛丽女王大学电子工程与计算机科学学院)
;
Department of Engineering Science, University of Oxford(牛津大学工程科学系)
机构
*
Basic Model Technology Center, WeChat AI, Tencent Inc.(腾讯基本模型技术中心、微信AI、腾讯公司)
;
State Key Laboratory of Multimedia Information Processing, School of Computer Science, Peking University(多媒体信息处理国家重点实验室、计算机学院、北京大学)
机构
*
The University of Hong Kong(香港大学)
;
Fudan University(复旦大学)
;
LMU Munich(慕尼黑大学)
;
Tsinghua University(清华大学)
;
Technische Universität Darmstadt(达姆施塔特技术大学)
;
Technische Universität Berlin(柏林技术大学)
;
Technische Universität Dresden(德累斯顿技术大学)
;
The Chinese University of Hong Kong(香港中文大学)
;
University of Pennsylvania(宾夕法尼亚大学)
;
Nanjing University(南京大学)
;
University of Manchester(曼彻斯特大学)
;
Dartmouth College(达特茅斯学院)
;
University of California Los Angeles(加州大学洛杉矶分校)
;
University of Michigan(密歇根大学)
;
Microsoft(微软)
;
Tencent(腾讯)
Autonomous Diffractometry Enabled by Visual Reinforcement Learning
由视觉强化学习实现的自主衍射测量
J. Oppliger, M. Stifter, A. Rüegg, I. Biało, L. Martinelli, P. G. Freeman, D. Prabhakaran, J. Zhao, Q. Wang, J. Chang
机构
*
Jeremiah Horrocks Institute for Mathematics, Physics and Astronomy, University of Central Lancashire(中央兰开夏大学杰里迈亚·霍罗克斯数学、物理与天文研究所)
;
Department of Physics, Clarendon Laboratory, University of Oxford(牛津大学克拉伦登实验室物理系)
;
State Key Laboratory of Surface Physics and Department of Physics, Fudan University(复旦大学表面物理国家重点实验室和物理系)
;
Department of Physics, The Chinese University of Hong Kong(香港中文大学物理系)
;
State Key Laboratory of Quantum Information Technologies and Materials, The Chinese University of Hong Kong(香港中文大学量子信息与材料国家重点实验室)
机构
*
Vermont Artificial Intelligence Lab, Department of Computer Science, University of Vermont(佛蒙特大学计算机科学系佛蒙特人工智能实验室)
;
Intelligent Machines Lab, Department of Artificial Intelligence, Information Technology University(信息技术大学人工智能系智能机器实验室)
;
Institute of Artificial Intelligence, University of Central Florida(中佛罗里达大学人工智能研究所)
机构
*
School of Computer Science and Engineering, Macau University of Science and Technology, China(澳门科技大学计算机科学与工程学院)
;
SKLPlanets, Macau University of Science and Technology, China(澳门科技大学月球与行星科学国家重点实验室)
;
School of Economics, Anhui University, China(安徽大学经济学院)
;
School of Energy and Power Engineering, Huazhong University of Science and Technology, China(华中科技大学能源与动力工程学院)