机构
*
School of Informatics, Xiamen University(厦门大学信息学院)
;
Institute of Artificial Intelligence, Xiamen University(厦门大学人工智能研究院)
;
The Hong Kong Polytechnic University(香港理工大学)
Reasoning Primitives in Hybrid and Non-Hybrid LLMs: Do Architectural Differences Yield Advantages in State-Tracking and Recall?
混合与非混合大语言模型中的推理原语:架构差异在状态追踪和召回中是否带来优势?
Shivam Rawat, Lucie Flek, Florian Mai, Nicholas Kluge Corrêa
机构
*
Lamarr Institute for Machine Learning and Artificial Intelligence(拉玛尔机器学习与人工智能研究所)
;
Rheinische Friedrich-Wilhelms-Universität Bonn(波恩莱茵河弗里德里希-威廉大学)
Catching The Correct Answer Trap: Characterising AI Tutor Blind Spots When Analysing Student Reasoning
捕捉正确答案陷阱:分析学生推理时AI导师盲点的特征化
Moiz Imran, Sahan Bulathwela
机构
*
Department of Computer Science, University College London, UK(英国伦敦大学学院计算机科学系)
;
Centre for Artificial Intelligence, University College London, UK(英国伦敦大学学院人工智能中心)
Metacognition as Reward: Reinforcing LLM Reasoning via Knowledge and Regulation Signals
元认知作为奖励:通过知识和调节信号强化LLM推理
Sirui Chen, Lei Xu, Yuying Zhao, Yutian Chen, Yu Wang, Beier Zhu, Hanwang Zhang, Shengjie Zhao, Chaochao Lu
机构
*
Tongji University(同济大学)
;
Shanghai AI Laboratory(上海人工智能实验室)
;
Nanyang Technological University(南洋理工大学)
;
University of Science and Technology of China(中国科学技术大学)
;
EPFL(苏黎世联邦理工学院)
;
Wuhan University(武汉大学)
Double-Calibration: Towards Reliable LLMs via Calibrating Knowledge and Reasoning Confidence
双重校准:通过校准知识和推理置信度实现可靠的LLM
Yuyin Lu, Ziran Liang, Yanghui Rao, Wenqi Fan, Fu Lee Wang, Qing Li
机构
*
School of Computer Science and Engineering, Sun Yat-sen University, Guangzhou, China(中山大学计算机科学与工程学院,广州,中国)
;
Department of Computing, The Hong Kong Polytechnic University, Hong Kong SAR(香港理工大学计算机系,香港特别行政区)
;
School of Science and Technology, Hong Kong Metropolitan University, Hong Kong SAR(香港 Metropolitan 大学科技学院,香港特别行政区)
Fine-Tuning Small Reasoning Models for Quantum Field Theory
对量子场论进行小规模推理模型的微调
Nathaniel S. Woodward, Zhiqi Gao, Yurii Kvasiuk, Kendrick M. Smith, Frederic Sala, Moritz Münchmeyer
机构
*
Department of Physics, University of Wisconsin-Madison(威斯康星大学麦迪逊分校物理系)
;
Department of Computer Science, University of Wisconsin-Madison(威斯康星大学麦迪逊分校计算机科学系)
;
Perimeter Institute for Theoretical Physics(理论物理研究所)
Reasoning-targeted Jailbreak Attacks on Large Reasoning Models via Semantic Triggers and Psychological Framing
通过语义触发和心理框架针对大推理模型的推理定向劫持攻击
Zehao Wang, Lanjun Wang
机构
*
College of Intelligence and Computing(智能与计算学院)
;
School of New Media and Communication(新媒体与传播学院)
;
Shanghai Key Laboratory of Data Science(上海数据科学 key laboratory)
机构
*
The Key Laboratory of Knowledge Engineering with Big Data (the Ministry of Education of China), Hefei University of Technology, China(合肥工业大学大数据知识工程教育部重点实验室)
;
School of Computer Science and Information Engineering, Hefei University of Technology, China(合肥工业大学计算机与信息学院)
;
Nanjing University of Science and Technology, China(南京理工大学)
;
China Unicom Digital Technology Co., Ltd., Beijing, China(联通数字科技有限公司)
;
China Unicom Internet of Things Co., Ltd., Nanjing, China(联通物联网有限责任公司)
;
Shandong Inspur Science Research Institute, Jinan, China(山东浪潮科学研究院)
;
Griffith University, Australia(格里菲斯大学)
机构
*
Institute for Clarity in Documentation(清晰文档研究所)
;
Inria Paris-Rocquencourt(巴黎- Rocquencourt 国家信息与自动化研究所)
;
Rajiv Gandhi University(拉吉夫·甘地大学)
;
Tsinghua University(清华大学)
;
Palmer Research Laboratories(帕勒实验室)
;
University of Chicago(芝加哥大学)
;
Stanford University(斯坦福大学)
;
Purdue University(普渡大学)
;
University of Southern California(南加州大学)