arXivDaily arXiv每日学术速递 周一至周五更新

AI 大模型

语言大模型 / LLM

大语言模型、预训练、指令微调、后训练和语言模型应用。

2025-11-07 至 2025-11-07 共收录 159 信号源:cs.CL, cs.AI, cs.LG

1. 长上下文与记忆 4 篇

2510.22968 2025-11-07 cs.CL cs.AI cs.CY 73%

Measuring Teaching with LLMs

Michael Hardy

机构 * Stanford University(斯坦福大学)

专题命中 长上下文与记忆 :large language model(abstract);language model(abstract);分类 cs.CL、cs.AI

详情

展开后加载摘要…

URL PDF HTML 收藏
2506.05410 2025-11-07 cs.CL 70%

Homogeneous Keys, Heterogeneous Values: Exploiting Local KV Cache Asymmetry for Long-Context LLMs

Wanyun Cui, Mingwei Xu

机构 * MoE Key Laboratory of Interdisciplinary Research of Computation and Economics(交叉计算与经济学 interdisciplinary 研究联合实验室) Shanghai University of Finance and Economics(上海财经大学)

专题命中 长上下文与记忆 :large language model(abstract);language model(abstract);分类 cs.CL

Comments 14 pages,7 figures;Accepted by NeurIPS 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2511.04595 2025-11-07 cs.CV 50%

UniSplat: Unified Spatio-Temporal Fusion via 3D Latent Scaffolds for Dynamic Driving Scene Reconstruction

Chen Shi, Shaoshuai Shi, Xiaoyang Lyu, Chunyang Liu, Kehua Sheng, Bo Zhang, Li Jiang

机构 * The Chinese University of Hong Kong, Shenzhen(香港中文大学(深圳)) Voyager Research, Didi Chuxing The University of Hong Kong(香港大学)

专题命中 长上下文与记忆 :foundation model(abstract)

详情

展开后加载摘要…

URL PDF HTML 收藏
2511.04593 2025-11-07 cs.NE q-bio.NC 50%

Neural Computation Without Slots: Steps Towards Biologically Plausible Memory and Attention in Natural and Artificial Intelligence

Shaunak Bhandarkar, James L. McClelland

专题命中 长上下文与记忆 :language model(abstract)

Comments 19 main text pages, 7 main text figures; 33 supplementary pages, 13 supplementary figures

详情

展开后加载摘要…

URL PDF HTML 收藏

2. 推理与问题求解 26 篇

2510.12829 2025-11-07 cs.CL cs.AI cs.LG cs.LO 89%

Mathematics with large language models as provers and verifiers

Hieu Le Duc, Leo Liberti

机构 * CNRS LIX Ecole Polytechnique, Institut Polytechnique de Paris(高等理工学院巴黎理工研究所)

专题命中 推理与问题求解 :large language model(title,abstract);language model(title,abstract);分类 cs.CL、cs.AI、cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
2511.04573 2025-11-07 cs.LG 88%

ARETE: an R package for Automated REtrieval from TExt with large language models

Vasco V. Branco, Jandó Benedek, Lidia Pivovarova, Luís Correia, Pedro Cardoso

专题命中 推理与问题求解 :large language model(title,abstract);language model(title,abstract);分类 cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
2511.04072 2025-11-07 cs.CL 88%

Plan of Knowledge: Retrieval-Augmented Large Language Models for Temporal Knowledge Graph Question Answering

Xinying Qian, Ying Zhang, Yu Zhao, Baohang Zhou, Xuhui Sui, Xiaojie Yuan

专题命中 推理与问题求解 :large language model(title,abstract);language model(title,abstract);分类 cs.CL

Comments Submitted to the IEEE for possible publication

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.11190 2025-11-07 cs.CV 88%

FlexAC: Towards Flexible Control of Associative Reasoning in Multimodal Large Language Models

Shengming Yuan, Xinyu Lyu, Shuailong Wang, Beitao Chen, Jingkuan Song, Lianli Gao

机构 * University of Electronic Science and Technology of China(电子科学与技术大学) Southwestern University of Finance and Economics(西南财经大学) Tongji University(同济大学)

专题命中 推理与问题求解 :large language model(title,abstract);language model(title,abstract)

Comments 19 pages, 11 figures. Accepted by the 39th Conference on Neural Information Processing Systems (NeurIPS 2025)

详情

展开后加载摘要…

URL PDF HTML 收藏
2511.04491 2025-11-07 cs.CL cs.AI cs.DB cs.IR cs.LG 87%

RUST-BENCH: Benchmarking LLM Reasoning on Unstructured Text within Structured Tables

Nikhil Abhyankar, Purvi Chaurasia, Sanchit Kabra, Ananya Srivastava, Vivek Gupta, Chandan K. Reddy

机构 * Virginia Tech(弗吉尼亚理工大学) IGDTUW New Delhi(新德里IGDTUW) Arizona State University(亚利桑那州立大学)

专题命中 推理与问题求解 :LLM(title);large language model(abstract);language model(abstract);prompting(abstract)

详情

展开后加载摘要…

URL PDF HTML 收藏
2511.03995 2025-11-07 cs.CR cs.AI 85%

Hybrid Fuzzing with LLM-Guided Input Mutation and Semantic Feedback

Shiyin Lin

机构 * Independent Researcher(独立研究者)

专题命中 推理与问题求解 :LLM(title,abstract);large language model(abstract);language model(abstract);分类 cs.AI

详情

展开后加载摘要…

URL PDF HTML 收藏
2408.13406 2025-11-07 cs.AI cs.CE cs.MA 85%

Collaboration Dynamics and Reliability Challenges of Multi-Agent LLM Systems in Finite Element Analysis

Chuan Tian, Yilei Zhang

专题命中 推理与问题求解 :LLM(title,abstract);large language model(abstract);language model(abstract);分类 cs.AI

详情

展开后加载摘要…

URL PDF HTML 收藏
2504.12795 2025-11-07 cs.CV 85%

EarthGPT-X: A Spatial MLLM for Multi-level Multi-Source Remote Sensing Imagery Understanding with Visual Prompting

Wei Zhang, Miaoxin Cai, Yaqian Ning, Tong Zhang, Yin Zhuang, Shijian Lu, He Chen, Jun Li, Xuerui Mao

机构 * School of Interdisciplinary Science, Beijing Institute of Technology(交叉科学学院,北京理工大学) College of Computing and Data Science, Nanyang Technological University(计算与数据科学学院,南洋理工大学) National Key Laboratory of Science and Technology on Space-Born Intelligent Information Processing, Beijing Institute of Technology(空间智能信息处理国家重点实验室,北京理工大学) School of Optics and Photonics, Beijing Institute of Technology(光学与 photonics 学院,北京理工大学) State Key Laboratory of Explosion Science and Safety Protection, Beijing(爆炸科学与安全防护国家重点实验室,北京)

专题命中 推理与问题求解 :prompting(title,abstract);large language model(abstract);language model(abstract)

详情

展开后加载摘要…

URL PDF HTML 收藏
2511.03908 2025-11-07 cs.CL 79%

Context informs pragmatic interpretation in vision-language models

Alvin Wei Ming Tan, Ben Prystawski, Veronica Boyce, Michael C. Frank

机构 * Department of Psychology Stanford University(心理学系 斯坦福大学)

专题命中 推理与问题求解 :language model(title,abstract);分类 cs.CL

Comments Accepted at CogInterp Workshop, NeurIPS 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2505.18574 2025-11-07 cs.PL cs.AI cs.AR cs.LG 79%

Autocomp: A Powerful and Portable Code Optimizer for Tensor Accelerators

Charles Hong, Sahil Bhatia, Alvin Cheung, Yakun Sophia Shao

专题命中 推理与问题求解 :LLM(abstract);large language model(abstract);language model(abstract);分类 cs.AI、cs.LG

Comments 10 pages + appendices

详情

展开后加载摘要…

URL PDF HTML 收藏
2511.04646 2025-11-07 cs.AI cs.CL cs.LG cs.MA 78%

DR. WELL: Dynamic Reasoning and Learning with Symbolic World Model for Embodied LLM-Based Multi-Agent Collaboration

Narjes Nourzad, Hanqing Yang, Shiyu Chen, Carlee Joe-Wong

机构 * University of Southern California(南加州大学) Carnegie Mellon University(卡内基梅隆大学)

专题命中 推理与问题求解 :LLM(title);分类 cs.CL、cs.AI、cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
2511.04023 2025-11-07 cs.SE cs.CR 78%

LLM-Driven Adaptive Source-Sink Identification and False Positive Mitigation for Static Analysis

Shiyin Lin

专题命中 推理与问题求解 :LLM(title,abstract)

详情

展开后加载摘要…

URL PDF HTML 收藏
2511.04654 2025-11-07 cs.CL 77%

Logit-Entropy Adaptive Stopping Heuristic for Efficient Chain-of-Thought Reasoning

Mohammad Atif Quamar, Mohammad Areeb

机构 * Purdue University(普渡大学)

专题命中 推理与问题求解 :large language model(abstract);language model(abstract);prompting(abstract);分类 cs.CL

Comments Presented at the 1st Workshop on Efficient Reasoning (NeurIPS 2025)

详情

展开后加载摘要…

URL PDF HTML 收藏
2511.04464 2025-11-07 cs.AI 77%

Beyond Shortest Path: Agentic Vehicular Routing with Semantic Context

Carnot Braun, Rafael O. Jarczewski, Gabriel U. Talasso, Leandro A. Villas, Allan M. de Souza

机构 * Institute of Computing (IC)(计算学院) Hub de Inteligência Artificial e Arquiteturas Cognitivas (H.IAAC)(人工智能与认知架构中心) Universidade Estadual de Campinas (UNICAMP)(坎皮纳斯州立大学)

专题命中 推理与问题求解 :LLM(abstract);large language model(abstract);language model(abstract);分类 cs.AI

详情

展开后加载摘要…

URL PDF HTML 收藏
2511.04093 2025-11-07 cs.AI 77%

KGFR: A Foundation Retriever for Generalized Knowledge Graph Question Answering

Yuanning Cui, Zequn Sun, Wei Hu, Zhangjie Fu

机构 * School of Computer Science, Nanjing University of Information Science and Technology(南京信息工程大学计算机科学学院) State Key Laboratory for Novel Software Technology, Nanjing University(南京大学软件新技术国家重点实验室) National Institute of Healthcare Data Science, Nanjing University(南京大学健康数据科学国家研究院) Engineering Research Center of Digital Forensics, Ministry of Education, Nanjing University of Information Science and Technology(南京信息工程大学教育部长江数字取证工程研究中心)

专题命中 推理与问题求解 :LLM(abstract);large language model(abstract);language model(abstract);分类 cs.AI

详情

展开后加载摘要…

URL PDF HTML 收藏
2511.03739 2025-11-07 cs.CL 77%

TextualVerifier: Verify TextGrad Step-by-Step

Eugenius Mario Situmorang, Adila Alfa Krisnadhi, Ari Wibisono

专题命中 推理与问题求解 :LLM(abstract);large language model(abstract);language model(abstract);分类 cs.CL

详情

展开后加载摘要…

URL PDF HTML 收藏
2511.04662 2025-11-07 cs.AI cs.CL 73%

VeriCoT: Neuro-symbolic Chain-of-Thought Validation via Logical Consistency Checks

Yu Feng, Nathaniel Weir, Kaj Bostrom, Sam Bayless, Darion Cassel, Sapana Chaudhary, Benjamin Kiesl-Reiter, Huzefa Rangwala

机构 * University of Pennsylvania(宾夕法尼亚大学) Amazon Web Services(亚马逊网络服务)

专题命中 推理与问题求解 :SFT(abstract);preference optimization(abstract);分类 cs.CL、cs.AI

详情

展开后加载摘要…

URL PDF HTML 收藏
2511.03845 2025-11-07 cs.AI cs.LG 73%

To See or To Read: User Behavior Reasoning in Multimodal LLMs

Tianning Dong, Luyi Ma, Varun Vasudevan, Jason Cho, Sushant Kumar, Kannan Achan

机构 * Personalization Team, Walmart Global Tech(Walmart全球科技个性化团队)

专题命中 推理与问题求解 :large language model(abstract);language model(abstract);分类 cs.AI、cs.LG

Comments Accepted by the 39th Conference on Neural Information Processing Systems (NeurIPS 2025) Workshop: Efficient Reasoning

详情

展开后加载摘要…

URL PDF HTML 收藏
2511.04432 2025-11-07 cs.CL 70%

If I Could Turn Back Time: Temporal Reframing as a Historical Reasoning Task for LLMs

Lars Bungum, Charles Yijia Huang, Abeer Kashar

机构 * NTNU Norway(挪威诺汉大学) University of Waterloo(滑铁卢大学)

专题命中 推理与问题求解 :LLM(abstract);prompting(abstract);分类 cs.CL

Comments 8 pages, 1 figure, 3 tables, submitted to aconference

详情

展开后加载摘要…

URL PDF HTML 收藏
2505.23653 2025-11-07 cs.LG 70%

How do Transformers Learn Implicit Reasoning?

Jiaran Ye, Zijun Yao, Zhidian Huang, Liangming Pan, Jinxin Liu, Yushi Bai, Amy Xin, Weichuan Liu, Xiaoyin Che, Lei Hou, Juanzi Li

机构 * DCST, BNRist(中国清华大学人工智能研究院) KIRC, Institute for Artificial Intelligence, Tsinghua University, China

专题命中 推理与问题求解 :large language model(abstract);language model(abstract);分类 cs.LG

Comments Accepted as Spotlight at NeurIPS 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2511.01833 2025-11-07 cs.CV 67%

TIR-Bench: A Comprehensive Benchmark for Agentic Thinking-with-Images Reasoning

Ming Li, Jike Zhong, Shitian Zhao, Haoquan Zhang, Shaoheng Lin, Yuxiang Lai, Chen Wei, Konstantinos Psounis, Kaipeng Zhang

机构 * Shanghai AI Laboratory(上海人工智能实验室) University of Southern California(南加州大学) Emory University(埃默里大学) Chinese University of Hong Kong(香港中文大学) Rice University(德克萨斯大学奥斯汀分校)

专题命中 推理与问题求解 :large language model(abstract);language model(abstract)

Comments Preprint

详情

展开后加载摘要…

URL PDF HTML 收藏
2511.03958 2025-11-07 cs.MA cs.CL cs.HC 57%

Multi-Agent Collaborative Framework For Math Problem Generation

Kia Karbasi, Kevin Hong, Mohammad Amin Samadi, Gregory Pottie

专题命中 推理与问题求解 :language model(abstract);分类 cs.CL

Comments Published in the Proceedings of the 18th International Conference on Educational Data Mining, 6 pages, 5 figures

Journal ref Kia Karbasi, Kevin Hong, Mohammad Amin Samadi, & Gregory Pottie. (2025). Multi-Agent Collaborative Framework For Math Problem Generation. Proceedings of the 18th International Conference on Educational Data Mining, 613--618

详情

展开后加载摘要…

URL PDF HTML 收藏
2511.03768 2025-11-07 cs.LG cs.CV 57%

What's in Common? Multimodal Models Hallucinate When Reasoning Across Scenes

Candace Ross, Florian Bordes, Adina Williams, Polina Kirichenko, Mark Ibrahim

机构 * FAIR at Meta(Meta 的 FAIR)

专题命中 推理与问题求解 :language model(abstract);分类 cs.LG

Comments 10 pages, 6 figures. Accepted to NeurIPS Datasets & Benchmarks 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.21814 2025-11-07 cs.CV cs.AI 57%

Gestura: A LVLM-Powered System Bridging Motion and Semantics for Real-Time Free-Form Gesture Understanding

Zhuoming Li, Aitong Liu, Mengxi Jia, Yubi Lu, Tengxiang Zhang, Changzhi Sun, Dell Zhang, Xuelong Li

机构 * Institute of Artificial Intelligence (TeleAI) of China Telecom(中国电信人工智能研究院(TeleAI))

专题命中 推理与问题求解 :language model(abstract);分类 cs.AI

Comments IMWUT2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.18751 2025-11-07 cs.AI cs.CV 57%

Seg the HAB: Language-Guided Geospatial Algae Bloom Reasoning and Segmentation

Patterson Hsieh, Jerry Yeh, Mao-Chi He, Wen-Han Hsieh, Elvis Hsieh

机构 * UC San Diego(加州大学圣地亚哥分校) UC Berkeley(加州大学伯克利分校)

专题命中 推理与问题求解 :language model(abstract);分类 cs.AI

详情

展开后加载摘要…

URL PDF HTML 收藏
2506.13757 2025-11-07 cs.CV 50%

AutoVLA: A Vision-Language-Action Model for End-to-End Autonomous Driving with Adaptive Reasoning and Reinforcement Fine-Tuning

Zewei Zhou, Tianhui Cai, Seth Z. Zhao, Yun Zhang, Zhiyu Huang, Bolei Zhou, Jiaqi Ma

机构 * University of California, Los Angeles(加州大学洛杉矶分校)

专题命中 推理与问题求解 :language model(abstract)

Comments NeurIPS 2025; Website link:https://autovla.github.io/

详情

展开后加载摘要…

URL PDF HTML 收藏