arXivDaily arXiv每日学术速递 周一至周五更新

AI 大模型

大模型推理能力

大模型数学、逻辑、规划、多步推理和测试时计算能力。

2025-12-05 至 2025-12-05 共收录 8 信号源:cs.CL, cs.AI, cs.LG

1. 其他推理 8 篇

2511.23002 2025-12-05 cs.CV 75%

JarvisEvo: Towards a Self-Evolving Photo Editing Agent with Synergistic Editor-Evaluator Optimization

JarvisEvo: 向自进化式照片编辑代理迈进:协同编辑器-评估器优化

Yunlong Lin, Linqing Wang, Kunjie Lin, Zixu Lin, Kaixiong Gong, Wenbo Li, Bin Lin, Zhenxi Li, Shiyi Zhang, Yuyang Peng, Wenxun Dai, Xinghao Ding, Chunyu Wang, Qinglin Lu

机构 * Tencent Hunyuan(腾讯文元)

专题命中 其他推理 :reasoning(abstract);chain-of-thought(abstract);CoT(abstract)

AI总结 JarvisEvo通过协同编辑器-评估器优化,实现自进化式照片编辑代理,提升编辑质量和内容保真度。

Comments 31 pages, 18 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2511.08003 2025-12-05 cs.CV cs.AI 74%

Sharp Eyes and Memory for VideoLLMs: Information-Aware Visual Token Pruning for Efficient and Reliable VideoLLM Reasoning

锐眼与记忆:视频大语言模型的信息感知视觉标记修剪方法

Jialong Qin, Xin Zou, Di Lu, Yibo Yan, Xuming Hu

专题命中 其他推理 :reasoning(title);分类 cs.AI

AI总结 SharpV通过自适应视觉标记修剪和键值缓存修剪,提升视频大语言模型的效率与可靠性,是首个无需暴露注意力分数的两阶段修剪框架。

Comments The 40th Annual AAAI Conference on Artificial Intelligence (AAAI-26) Poster

详情

展开后加载摘要…

URL PDF HTML 收藏
2512.04518 2025-12-05 cs.CL cs.AI 62%

UW-BioNLP at ChemoTimelines 2025: Thinking, Fine-Tuning, and Dictionary-Enhanced LLM Systems for Chemotherapy Timeline Extraction

UW-BioNLP在ChemoTimelines 2025中的表现:基于思考、微调和词典增强的LLM系统用于化疗时间线提取

Tianmai M. Zhang, Zhaoyi Sun, Sihang Zeng, Chenxi Li, Neil F. Abernethy, Barbara D. Lam, Fei Xia, Meliha Yetisgen

机构 * University of Washington(华盛顿大学)

专题命中 其他推理 :chain-of-thought(abstract);分类 cs.CL、cs.AI

AI总结 UW-BioNLP通过思考、微调和词典增强LLM方法,在ChemoTimelines 2025中实现了最佳性能,提升了化疗时间线提取的准确性。

Comments To be published in Proceedings of the 7th Clinical Natural Language Processing Workshop

详情

展开后加载摘要…

URL PDF HTML 收藏
2511.17127 2025-12-05 cs.CL cs.AI cs.DC 62%

Training Foundation Models on a Full-Stack AMD Platform: Compute, Networking, and System Design

在全栈AMD平台上训练基础模型:计算、网络和系统设计

Quentin Anthony, Yury Tokpanov, Skyler Szot, Srivatsan Rajagopal, Praneeth Medepalli, Anna Golubeva, Vasu Shyam, Robert Washbourne, Rishi Iyer, Ansh Chaurasia, Tomas Figliolia, Xiao Yang, Abhinav Sarje, Drew Thorstensen, Amartey Pearson, Zack Grossbart, Jason van Patten, Emad Barsoum, Zhenyu Gu, Yao Fu, Beren Millidge

机构 * IBM

专题命中 其他推理 :reasoning(abstract);分类 cs.CL、cs.AI

AI总结 本文介绍了在AMD全栈平台上训练大规模混合专家模型的方法,展示了ZAYA1模型在预训练中的性能表现。

详情

展开后加载摘要…

URL PDF HTML 收藏
2512.04822 2025-12-05 cs.AI 57%

Enabling Ethical AI: A case study in using Ontological Context for Justified Agentic AI Decisions

实现道德AI:利用本体上下文进行合理化智能体决策的案例研究

Liam McGee, James Harvey, Lucy Cull, Andreas Vermeulen, Bart-Floris Visscher, Malvika Sharan

专题命中 其他推理 :reasoning(abstract);分类 cs.AI

AI总结 本文提出一种协作人机AI方法,通过本体上下文构建可检查语义层,提升智能体决策的合理性和透明度。

Comments 24 pages including references, with 6 images and 2 tables. Appendices, supporting data and additional reference provided from page 25 to 117

详情

展开后加载摘要…

URL PDF HTML 收藏
2512.04532 2025-12-05 cs.CV cs.AI 57%

PhyVLLM: Physics-Guided Video Language Model with Motion-Appearance Disentanglement

PhyVLLM:具有运动-外观解耦的物理引导视频语言模型

Yu-Wei Zhan, Xin Wang, Hong Chen, Tongtong Feng, Wei Feng, Ren Wang, Guangyao Li, Qing Li, Wenwu Zhu

机构 * Department of Computer Science and Technology, Tsinghua University(计算机科学与技术系,清华大学) Department of Electronic Engineering, Tsinghua University(电子工程系,清华大学)

专题命中 其他推理 :reasoning(abstract);分类 cs.AI

AI总结 PhyVLLM通过引入物理运动建模和运动-外观解耦,提升视频语言模型在物理推理和视频理解任务中的性能。

详情

展开后加载摘要…

URL PDF HTML 收藏
2512.04513 2025-12-05 cs.AI 57%

BiTAgent: A Task-Aware Modular Framework for Bidirectional Coupling between Multimodal Large Language Models and World Models

BiTAgent: 一种面向任务的模块化框架,用于多模态大语言模型与世界模型之间的双向耦合

Yu-Wei Zhan, Xin Wang, Pengzhe Mao, Tongtong Feng, Ren Wang, Wenwu Zhu

机构 * Tsinghua University(清华大学) Shandong University(山东大学)

专题命中 其他推理 :reasoning(abstract);分类 cs.AI

AI总结 BiTAgent通过双向耦合多模态大语言模型与世界模型,实现任务感知的动态联合学习,提升多任务和跨环境的泛化能力。

详情

展开后加载摘要…

URL PDF HTML 收藏
2512.04425 2025-12-05 cs.CV cs.AI 57%

Explainable Parkinsons Disease Gait Recognition Using Multimodal RGB-D Fusion and Large Language Models

可解释的帕金森病步态识别:基于多模态RGB-D融合与大语言模型

Manar Alnaasan, Md Selim Sarowar, Sungho Kim

专题命中 其他推理 :reasoning(abstract);分类 cs.AI

AI总结 本文提出基于多模态RGB-D融合与大语言模型的帕金森病步态识别框架,通过增强时空表示和语言解释提高识别准确性和临床可解释性。

详情

展开后加载摘要…

URL PDF HTML 收藏