arXivDaily arXiv每日学术速递 周一至周五更新

AI 大模型

大模型推理能力

大模型数学、逻辑、规划、多步推理和测试时计算能力。

2025-12-10 至 2025-12-10 共收录 9 信号源:cs.CL, cs.AI, cs.LG

1. 其他推理 9 篇

2512.00218 2025-12-10 cs.AI cs.CR 88%

Reasoning Under Pressure: How do Training Incentives Influence Chain-of-Thought Monitorability?

压力下的推理:训练激励如何影响推理链的可监控性?

Matt MacDermott, Qiyao Wei, Rada Djoneva, Francis Rhys Ward

专题命中 其他推理 :reasoning(title,abstract);chain-of-thought(title);CoT(abstract);分类 cs.AI

AI总结 本文研究了训练激励对推理链可监控性的影响,发现对抗性优化降低监控性能,而直接优化可监控性未显著提升。

详情

展开后加载摘要…

URL PDF HTML 收藏
2512.08270 2025-12-10 cs.AI cs.CL q-fin.GN 81%

Reasoning Models Ace the CFA Exams

推理模型在CFA考试中表现优异

Jaisal Patel, Yunzhe Chen, Kaiwen He, Keyi Wang, David Li, Kairong Xiao, Xiao-Yang Liu

机构 * Rensselaer Polytechnic Institute(罗格斯理工学院) University of North Carolina at Chapel Hill(北卡罗来纳大学教堂山分校) SecureFinAI Lab(安全金融人工智能实验室) Columbia University(哥伦比亚大学) Business School(商学院)

专题命中 其他推理 :reasoning(title,abstract);分类 cs.CL、cs.AI

AI总结 本文评估了先进推理模型在模拟CFA考试中的表现,发现Gemini 3.0 Pro在多个考试级别均取得优异成绩。

详情

展开后加载摘要…

URL PDF HTML 收藏
2512.08820 2025-12-10 cs.CV cs.AI 79%

Training-Free Dual Hyperbolic Adapters for Better Cross-Modal Reasoning

无需训练的双双曲适配器用于更高效的跨模态推理

Yi Zhang, Chun-Wun Cheng, Junyi He, Ke Yu, Yushun Tang, Carola-Bibiane Schönlieb, Zhihai He, Angelica I. Aviles-Rivero

机构 * College of Computer Science and Software Engineering, Shenzhen University(深圳大学计算机科学与软件工程学院) Department of Electrical and Electronic Engineering, Southern University of Science and Technology(南方科技大学电子与电气工程系) Department of Applied Mathematics and Theoretical Physics, University of Cambridge(剑桥大学应用数学与理论物理系) Yau Mathematical Sciences Center, Tsinghua University(清华大学应用数学中心)

专题命中 其他推理 :reasoning(title,abstract);分类 cs.AI

AI总结 本文提出无需训练的双双曲适配器方法,通过双曲空间嵌入提升跨模态推理性能,实现更高效的领域泛化和少样本识别。

Comments Accepted in IEEE Transactions on Multimedia (TMM)

详情

展开后加载摘要…

URL PDF HTML 收藏
2511.10384 2025-12-10 cs.SI cs.AI cs.CL cs.CY 62%

Simulating Misinformation Propagation in Social Networks using Large Language Models

利用大语言模型模拟社交媒体上的虚假信息传播

Raj Gaurav Maurya, Vaibhav Shukla, Raj Abhijit Dandekar, Rajat Dandekar, Sreedath Panat

机构 * Vizuara AI Labs(Vizuara AI实验室)

专题命中 其他推理 :reasoning(abstract);分类 cs.CL、cs.AI

AI总结 本文利用大语言模型模拟社交媒体虚假信息传播,通过构建人格代理网络研究虚假信息演变机制,揭示身份和意识形态驱动的人格加速虚假信息扩散,专家驱动人格则保持事实稳定。

Comments Accepted to CIKM 2025 Workshop LASS

详情

展开后加载摘要…

URL PDF HTML 收藏
2512.08524 2025-12-10 cs.CV cs.CL 57%

Beyond Real Weights: Hypercomplex Representations for Stable Quantization

超越真实权重:用于稳定量化 的超复数表示

Jawad Ibn Ahad, Maisha Rahman, Amrijit Biswas, Muhammad Rafsan Kabir, Robin Krambroeckers, Sifat Momen, Nabeel Mohammed, Shafin Rahman

机构 * Artificial Intelligence Department, RobotBulls Labs(机器人bulls实验室人工智能部门) Machine Intelligence Lab (MILab), North South University(北南大学机器智能实验室)

专题命中 其他推理 :reasoning(abstract);分类 cs.CL

AI总结 本文提出了一种基于超复数乘法的渐进式重新参数化策略,用于压缩多模态语言模型,实现参数和计算量的显著减少,同时保持模型性能。

Comments Accepted in Winter Conference on Applications of Computer Vision (WACV) 2026

详情

展开后加载摘要…

URL PDF HTML 收藏
2512.08123 2025-12-10 cs.CL 57%

Universal Adversarial Suffixes Using Calibrated Gumbel-Softmax Relaxation

利用校准的Gumbel-Softmax松弛的通用对抗后缀

Sampriti Soor, Suklav Ghosh, Arijit Sur

机构 * Center for Intelligent Cyber Physical Systems(智能网络物理系统中心) Indian Institute of Technology Guwahati(印度古瓦哈提理工学院) Department of Computer Science and Engineering(计算机科学与工程系)

专题命中 其他推理 :reasoning(abstract);分类 cs.CL

AI总结 本文提出了一种利用校准的Gumbel-Softmax松弛学习通用对抗后缀的方法,能有效降低多种任务和模型的准确性。

Comments 10 pages

详情

展开后加载摘要…

URL PDF HTML 收藏
2512.07917 2025-12-10 cs.SE cs.AI physics.flu-dyn 57%

CFD-copilot: leveraging domain-adapted large language model and model context protocol to enhance simulation automation

CFD-copilot: 利用领域适应的大语言模型和模型上下文协议增强仿真自动化

Zhehao Dong, Shanghai Du, Zhen Lu, Yue Yang

机构 * State Key Laboratory for Turbulence and Complex Systems(湍流与复杂系统国家重点实验室) School of Mechanics and Engineering Science(力学与工程科学学院) Peking University(北京大学) HEDPS-CAPT

专题命中 其他推理 :reasoning(abstract);分类 cs.AI

AI总结 CFD-copilot通过领域适应的大语言模型和模型上下文协议提升CFD仿真的自动化水平。

详情

展开后加载摘要…

URL PDF HTML 收藏
2511.14318 2025-12-10 physics.ed-ph cs.CY 50%

Report on the Scoping Workshop on AI in Science Education Research 2025

2025年人工智能在科学教育研究中的范围研讨会报告

Marcus Kubsch, Marit Kastaun, Peter Wulff, Nicole Graulich, Moriah Ariely, Alexander Bergmann-Gering, Sebastian Gombert, Bor Gregorcic, Hendrik Härtig, Benedikt Heuckmann, Andrea Horbach, Christina Krist, Gerd Kortemeyer, Ben Münch, Samuel Pazicni, Joshua M. Rosenberg, Sascha Schanze, Gena Sbeglia, Vidar Skogvoll, Christophe Speroni, Christoph Thyssen, Lars-Jochen Thoms, Brandon J. Yik, Xiaoming Zhai

专题命中 其他推理 :reasoning(abstract)

AI总结 2025年人工智能在科学教育研究中的范围研讨会报告,探讨AI在教育研究中的应用、挑战及负责任的整合方法。

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.06335 2025-12-10 cs.CV 50%

Harnessing Object Grounding for Time-Sensitive Video Understanding

利用物体接地提升时间敏感视频理解

Tz-Ying Wu, Sharath Nittur Sridhar, Subarna Tripathi

机构 * Intel(英特尔公司)

专题命中 其他推理 :reasoning(abstract)

AI总结 本文提出 GO-Tokenizer 以提升视频大型语言模型的时间敏感视频理解能力,通过实时编码紧凑的物体信息,提高模型性能并减少噪声影响。

Comments Accepted to WACV 2026

详情

展开后加载摘要…

URL PDF HTML 收藏