arXivDaily arXiv每日学术速递 周一至周五更新

AI 大模型

语言大模型 / LLM

大语言模型、预训练、指令微调、后训练和语言模型应用。

共收录 12228 信号源:cs.CL, cs.AI, cs.LG

1. 其他LLM 12228 篇

1502.00804 2015-03-09 cs.IR 78%

A Polya Urn Document Language Model for Improved Information Retrieval

Ronan Cummins, Jiaul Hoque Paik, Yuanhua Lv

专题命中 其他LLM :language model(title,abstract)

Comments 37 page journal submission (accepted for publication in TOIS)

详情

展开后加载摘要…

URL PDF HTML 收藏
1401.3896 2014-01-17 cs.IR 78%

The Opposite of Smoothing: A Language Model Approach to Ranking Query-Specific Document Clusters

Oren Kurland, Eyal Krikon

专题命中 其他LLM :language model(title,abstract)

Journal ref Journal Of Artificial Intelligence Research, Volume 41, pages 367-395, 2011

详情

展开后加载摘要…

URL PDF HTML 收藏
1401.1732 2014-01-09 cs.IR 78%

Looking at Vector Space and Language Models for IR using Density Matrices

Alessandro Sordoni, Jian-Yun Nie

专题命中 其他LLM :language model(title,abstract)

Comments In Proceedings of Quantum Interaction 2013

详情

展开后加载摘要…

URL PDF HTML 收藏
1307.5839 2013-07-24 cs.NE math.OC 78%

A New Approach for Finding the Global Optimal Point Using Subdividing Labeling Method (SLM)

Masoumeh Vali

专题命中 其他LLM :SLM(title,abstract)

Comments arXiv admin note: text overlap with arXiv:1307.5667, arXiv:1307.5840

详情

展开后加载摘要…

URL PDF HTML 收藏
1208.5979 2013-02-11 hep-th 78%

Beyond LLM in M-theory

Eoin Ó Colgáin

专题命中 其他LLM :LLM(title,abstract)

Comments 1+30 pages, footnote added

Journal ref JHEP 1212 (2012) 023

详情

展开后加载摘要…

URL PDF HTML 收藏
1210.4247 2012-10-17 cs.IT math.IT 78%

Deterministic Selection of Phase Sequences in Low Complexity SLM Scheme

Jun-Young Woo, Hyun-Seung Joo, Kee-Hoon Kim, Jong-Seon No, Dong-Joon Shin

专题命中 其他LLM :SLM(title,abstract)

详情

展开后加载摘要…

URL PDF HTML 收藏
1208.6412 2012-09-03 cs.IT math.IT 78%

Adaptive Generation Method of OFDM Signals in SLM Schemes for Low-complexity

Kee-Hoon Kim, Hyun-Seung Joo, Jong-Seon No, Dong-Joon Shin

专题命中 其他LLM :SLM(title,abstract)

详情

展开后加载摘要…

URL PDF HTML 收藏
1005.4752 2012-05-02 cs.IR cs.DB 78%

A database approach to information retrieval: The remarkable relationship between language models and region models

Djoerd Hiemstra, Vojkan Mihajlovic

专题命中 其他LLM :language model(title,abstract)

Comments Published as CTIT Technical Report 05-35

详情

展开后加载摘要…

URL PDF HTML 收藏
1010.5982 2011-04-18 hep-th 78%

On the generality of the LLM geometries in M-theory

Eoin Ó Colgáin, Jun-Bao Wu, Hossein Yavartanoo

专题命中 其他LLM :LLM(title,abstract)

Comments 15 pages, v2. minor improvements

Journal ref JHEP 1104:002,2011

详情

展开后加载摘要…

URL PDF HTML 收藏
1010.3101 2011-01-27 hep-th 78%

The electrostatic view on M-theory LLM geometries

Aristomenis Donos, Joan Simon

专题命中 其他LLM :LLM(title,abstract)

Comments 35 pages

Journal ref JHEP 1101:067,2011

详情

展开后加载摘要…

URL PDF HTML 收藏
hep-th/0508177 2009-12-01 hep-th 78%

Dynamics of Giant-Gravitons in the LLM geometry and the Fractional Quantum Hall Effect

Jian Dai, Xiao-Jun Wang, Yong-Shi Wu

专题命中 其他LLM :LLM(title,abstract)

Comments 32 pages, 1 figure; v.2: references added, the relation between the level shift and filling fraction elaborated

Journal ref Nucl.Phys. B731 (2005) 285-308

详情

展开后加载摘要…

URL PDF HTML 收藏
2507.02169 2026-08-17 cs.LG 版本更新 77%

Responsiveness Verification: Will Predictions Change? How Much? How Often?

响应性验证:预测会改变吗?改变多少?改变频率如何?

Harry Cheon, Meredith Stewart, Bogdan Kulynych, Tsui-Wei Weng, Berk Ustun

专题命中 其他LLM :LLM(summary_cn,abstract_cn);分类 cs.LG

AI总结 本研究提出测量响应性的算法与交互模型框架,搭配统计保证,可用于检测累犯预测的排除情况、估计内容审核博弈成本、测试LLM基准鲁棒性,提升模型安全性与可靠性。

详情

展开后加载摘要…

URL PDF HTML 收藏
2608.04607 2026-08-06 math.OC cs.LG 新提交 77%

On MUON optimization: From non-convergence to an error analysis with Polar Express and the Newton-Schulz polynomial from implementations

MUON优化:从非收敛性到结合Polar Express与Newton-Schulz多项式实现的误差分析

Thang Do, Steffen Dereich, Arnulf Jentzen

专题命中 其他LLM :LLM(abstract_cn);large language model(abstract);language model(abstract);分类 cs.LG

AI总结 本文针对2024年提出的MUON优化器,提出含任意NS步骤的广义变体,证明其在部分随机优化问题中无法收敛,建立误差分析并在多个实例中验证相关结论。

Comments 82 pages

详情

展开后加载摘要…

URL PDF HTML 收藏
2608.00549 2026-08-04 cs.IR cs.AI cs.HC 新提交 77%

A Context-Aware Cultural Heritage Guide Powered by LLMs

由大语言模型(LLMs)驱动的上下文感知文化遗产指南

Liliana Ardissono, Fabio Ferrero, Angelo Geninatti Cossatin, Claudio Mattutino, Noemi Mauro

专题命中 其他LLM :LLM(abstract);large language model(abstract);language model(abstract);分类 cs.AI

AI总结 本研究扩展了文化遗产网页应用Triangolazioni,构建与LLMs无关的松耦合架构,实现LLMs驱动的上下文感知文化遗产信息搜索与展示,丰富了文化遗产内容。

Journal ref In Proceedings of the 34th ACM Conference on User Modeling, Adaptation and Personalization (UMAP '26). 2026

详情

展开后加载摘要…

URL PDF HTML 收藏
2607.13683 2026-08-03 cs.CL 版本更新 77%

HarnessBank: Semantic Gene-Bank Search with Gated Verification for Agent-Harness Self-Evolution

通过门控语义质量多样性实现自我进化智能体的驾驭

Xiaotian Luo, Dizhan Xue, Fengxingyu Wang, Chuanrui Hu, Yafeng Deng

专题命中 其他LLM :large language model(abstract);language model(abstract);foundation model(abstract);分类 cs.CL

AI总结 研究大语言模型智能体驾驭因素提升性能的问题,提出自我进化智能体驾驭框架,将提出与评估更改分开,通过门控存档避免过度拟合,在多领域测试中取得较好泛化效果,证明诊断与评估循环的有效性。

Comments 9 pages, 4 figures, 3 tables

详情

展开后加载摘要…

URL PDF HTML 收藏
2505.21544 2026-08-03 cs.CV cs.CL 77%

Vision Meets Language: A RAG-Augmented YOLOv8 Framework for Coffee Disease Diagnosis and Farmer Assistance

Semanto Mondal

机构 * University of Naples Federico II(那不勒斯费德里科二世大学)

专题命中 其他LLM :LLM(abstract);large language model(abstract);language model(abstract);分类 cs.CL

Comments There are 14 pages, 8 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2607.21517 2026-07-31 cs.IT cs.AI cs.DM math.CO math.IT 版本更新 77%

Improved lower bounds for the Shannon capacity of odd cycles

奇数圈香农容量的改进下界

Nathaniel Itty, Christopher D. Rosin, Chase Carstensen, Daniel Reichman

机构 * Worcester Polytechnic Institute(沃斯特理工学院)

专题命中 其他LLM :LLM(abstract);large language model(abstract);language model(abstract);分类 cs.AI

AI总结 研究通过与大语言模型迭代交互,在奇数圈\(C_7^{10}\)、\(C_{11}^{6}\)、\(C_{13}^{6}\)中构造独立集,改进了这些图香农容量的已知下界,还改进了几个奇数圈单个强幂独立数的已知下界。

Comments v2: added improvement on lower bound for the Shannon capacity of C15

详情

展开后加载摘要…

URL PDF HTML 收藏
2607.19843 2026-07-23 cs.SE cs.AI 新提交 77%

Beyond Fail-to-Pass: Iterative Hardening of Co-Generated Bug Reproduction Tests and Fixes

超越未通过测试:协同生成的错误重现测试与修复的迭代强化

Yuhao Tan, Zhibang Yang, Fangkai Yang, Yuan Yao, Yu Kang, Lu Wang, Pu Zhao, Xin Zhang, Xiaoxing Ma, Qingwei Lin, Saravan Rajmohan, Dongmei Zhang

机构 * Nanjing University(南京大学) Peking University(北京大学) Microsoft(微软公司)

专题命中 其他LLM :LLM(abstract);large language model(abstract);language model(abstract);分类 cs.AI

AI总结 研究自动化程序修复中错误重现测试问题,指出仅用未通过到通过标准不足。提出CoHarden框架,先生成测试再迭代强化测试与修复,实验证明该框架在解决率等方面优于现有基线。

Comments 29 pages, 5 figures, preprint

详情

展开后加载摘要…

URL PDF HTML 收藏
2607.07846 2026-07-10 cs.AI cs.CY 新提交 77%

VectorizationLLM: Smart Vectorization Based AI Assistant

向量化语言模型:基于智能向量化的人工智能助手

Ryan Duke

机构 * New York Institute of Technology(纽约理工学院)

专题命中 其他LLM :LLM(abstract);large language model(abstract);language model(abstract);分类 cs.AI

AI总结 研究旨在设计基于谷歌开放权重语言模型的VectorizationLLM,用于辅助学生学习MATLAB相关知识。采用RAG知识库和系统提示架构,不直接给答案,通过课堂笔记示例详细解释概念,为课程应用提供有指导意义的智能助手。

Comments 44 pages, 6 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2607.07021 2026-07-09 cs.AI cs.HC 新提交 77%

Learning social norms enhances compatibility in dynamic human-AI coordination

学习社会规范可增强动态人机协作中的兼容性

Yi Yang, Siyuan Liu, Xin Gao, Huamu Sun, Chao Liu, Qing Zhou, Bingbing Nie

机构 * School of Vehicle and Mobility, Tsinghua University, Beijing, China(清华大学车辆与移动系统学院) State Key Laboratory of Intelligent Green Vehicle and Mobility, Tsinghua University, Beijing 100084, China(清华大学智能绿色车辆与移动系统国家重点实验室) State Key Laboratory of Cognitive Neuroscience and Learning & IDG/McGovern Institute for Brain Research, Beijing Normal University, Beijing, China(北京师范大学认知神经科学与学习国家重点实验室) Beijing Key Laboratory of Safe AI and Superalignment, Beijing, China(北京安全人工智能与超对齐关键实验室) Beijing Institute of AI Safety and Governance, Beijing, China(北京人工智能安全与治理研究院)

专题命中 其他LLM :LLM(abstract);large language model(abstract);language model(abstract);分类 cs.AI

AI总结 研究以行人与车辆互动为代表,搭建实验平台,识别出人类社会规范的三项原则,将其融入人工智能显著改善人机协作,在闭环互动任务中,融入社会规范的大语言模型总分比基线策略高出近四倍,比人际互动高出43%。

Comments 44 pages, 5 figures, supplementary information included

详情

展开后加载摘要…

URL PDF HTML 收藏
2412.14995 2026-07-07 cs.NE cs.AI 77%

HSEvo: Elevating Automatic Heuristic Design with Diversity-Driven Harmony Search and Genetic Algorithm Using LLMs

HSEvo: 通过多样性驱动的和谐搜索和遗传算法提升自动启发式设计

Pham Vu Tuan Dat, Long Doan, Huynh Thi Thanh Binh

机构 * Hanoi University of Science and Technology(河内科学技术大学) George Mason University(乔治·梅森大学)

专题命中 其他LLM :LLM(abstract);large language model(abstract);language model(abstract);分类 cs.AI

AI总结 本文提出HSEvo框架,通过结合和谐搜索算法平衡多样性与收敛性,解决自动启发式设计中探索与利用的平衡问题,提升搜索效率与稳定性。

Comments 18 pages, 12 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2603.04390 2026-07-03 cs.AI cs.SE 版本更新 77%

A Dual-Helix Governance Approach Towards Reliable Agentic Artificial Intelligence for WebGIS Development

面向可靠WebGIS开发的代理人工智能双螺旋治理方法

Boyuan Guan, Wencong Cui, Levente Juhasz

机构 * Geographic Information Systems Center, Florida International University(佛罗里达国际大学地理信息系统中心) Geospatial Analytics, Technology and Open Research Lab, University of Florida(佛罗里达大学地理空间分析、技术与开放研究实验室)

专题命中 其他LLM :LLM(abstract,abstract_cn);prompting(abstract);分类 cs.AI

AI总结 针对代理AI在WebGIS开发中的不可靠问题,提出双螺旋治理框架,通过知识-行为-技能三轨架构和持久知识图谱外化事实与协议,实验证明能降低代码复杂度、减少输出方差并防止制图错误。

Comments Paper submitted to and under review in Transactions in GIS

详情

展开后加载摘要…

URL PDF HTML 收藏
2606.28926 2026-06-30 cs.IT cs.LG math.IT 77%

A Theoretical Interpretation of In-Context Learning via Probabilistic Modeling

通过概率建模对上下文学习的理论解释

Zhenyu Liu, Huaze Tang, Shao-Lun Huang

机构 * Tsinghua Shenzhen International Graduate School, Tsinghua University, Shenzhen, China(清华大学深圳国际研究生院,清华大学,深圳,中国)

专题命中 其他LLM :LLM(abstract_cn);large language model(abstract);language model(abstract);分类 cs.LG

AI总结 本文提出上下文学习的概率模型,推导其在一般参数分布和指数族下的性能,并解释示例数量、模型参数敏感性和示例-查询相似性对性能的影响。

详情

展开后加载摘要…

URL PDF HTML 收藏
2505.16164 2026-06-30 cs.CL 77%

Can LLMs Simulate Human Behavioral Variability? A Case Study in the Phonemic Fluency Task

LLMs能否模拟人类行为变异?一项在音素流畅任务中的案例研究

Mengyang Qiu, Zoe Brisebois, Siena Sun

专题命中 其他LLM :LLM(abstract);large language model(abstract);language model(abstract);分类 cs.CL

AI总结 研究探讨LLMs在音素流畅任务中是否能模拟个体差异,发现尽管部分模型能模拟平均值和词汇偏好,但无法再现人类行为变异范围,且新模型和思考模式反而降低了多样性。

Journal ref Proceedings of the 15th Workshop on Cognitive Modeling and Computational Linguistics (2026) 250-263

详情

展开后加载摘要…

URL PDF HTML 收藏
2602.18364 2026-06-26 cs.IT cs.LG math.IT quant-ph stat.ML 版本更新 77%

Quantum Maximum Likelihood Prediction via Hilbert Space Embeddings

通过希尔伯特空间嵌入的量子最大似然预测

Sreejith Sreekumar, Nir Weinberger

机构 * L2S, CNRS, CentraleSupélec, University of Paris-Saclay, France(L2S、CNRS、CentraleSupélec、巴黎-萨克雷大学、法国)

专题命中 其他LLM :LLM(abstract_cn);large language model(abstract);language model(abstract);分类 cs.LG

AI总结 研究量子最大似然预测任务,通过将经验概率分布嵌入量子态并最小化量子相对熵,提出统一框架,给出非渐近性能保证。

Comments 38+4 pages, 1 figure

详情

展开后加载摘要…

URL PDF HTML 收藏
2507.09839 2026-06-24 cs.LG 版本更新 77%

Stabilizing Black-Box Prompt Optimization with Textual Regularization and Signal Aggregation

通过文本正则化与信号聚合稳定黑盒提示优化

MohammadReza Davari, Utkarsh Garg, Weixin Cai, Eugene Belilovsky

专题命中 其他LLM :LLM(abstract_cn);large language model(abstract);language model(abstract);分类 cs.LG

AI总结 提出TRAS框架,利用成功预测的文本正则化与蒙特卡洛信号聚合,解决黑盒提示优化中的不稳定性和语义漂移,提升准确率与收敛速度。

详情

展开后加载摘要…

URL PDF HTML 收藏
2606.21619 2026-06-23 cs.SE cs.LG cs.PL 新提交 77%

The Alignment Problem in Constrained Code Generation

约束代码生成中的对齐问题

Matteo Biagiola, Jahrim Gabriele Cesario, Luca Di Grazia, George Zakhour, Guido Salvaneschi

机构 * University of St. Gallen(圣加尔登大学) Università della Svizzera italiana (USI)(瑞士联邦理工学院)

专题命中 其他LLM :LLM(abstract_cn);large language model(abstract);language model(abstract);分类 cs.LG

AI总结 研究约束解码中约束器、语言模型与目标语言之间的对齐问题,发现约束器的不完整性会扭曲模型分布,导致功能正确性下降高达97%,并提出设计约束器的定量见解。

详情

展开后加载摘要…

URL PDF HTML 收藏
2606.20946 2026-06-23 cs.CL cs.CV 新提交 77%

Scaling Diverse Language Generation for 3D Visual Grounding

面向3D视觉定位的多样化语言生成扩展

Austin T. Wang, Dongchen Yang, Angel X. Chang

机构 * Simon Fraser University(西蒙菲莎大学)

专题命中 其他LLM :LLM(summary_cn,abstract_cn);分类 cs.CL

AI总结 提出ViGiL3D++方法,通过场景图约束采样与LLM语言生成结合,生成多样化视觉定位查询,提升3DVG模型泛化能力并揭示VLM局限性。

Comments 39 pages, 14 figures, 16 tables. Project Page: https://3dlg-hcvc.github.io/vigil3dpp

详情

展开后加载摘要…

URL PDF HTML 收藏
2606.19864 2026-06-19 cs.CL 新提交 77%

The Almost Intelligent Revolution: Options for Scaling Up Deliberation and Empowering People with AI

近乎智能的革命:扩大审议规模并利用AI赋能人类的选项

Serge Sharoff

机构 * Centre on Participatory and Deliberative Democracy(参与性和协商性民主研究中心)

专题命中 其他LLM :LLM(abstract_cn);large language model(abstract);language model(abstract);分类 cs.CL

AI总结 探讨大型语言模型如何通过系统功能语言学视角扩大民主审议规模,增强包容性并赋权边缘群体,同时警惕过度承诺与低估风险。

Comments Published in /Handbook of Democracy in the Era of Artificial Intelligence/ edited by Evangelos Pournaras, Srijoni Majumdar, Carina Ines Hausladen, and Dirk Helbing. 2026

详情

展开后加载摘要…

URL PDF HTML 收藏
2606.07833 2026-06-09 cs.CR cs.AI 新提交 77%

Beyond Pass/Fail: Using Process Mining to Understand How LLMs Resist (and Fail) Red Team Attacks

超越通过/失败:使用过程挖掘理解LLM如何抵抗(和失败)红队攻击

Zvi Topol

机构 * MuyVentive LLC

专题命中 其他LLM :LLM(title_cn,abstract_cn);分类 cs.AI

AI总结 提出将过程挖掘应用于红队攻击轨迹,通过分析事件日志提取直接跟随图和状态转移矩阵,揭示GPT-OSS和Llama 3.3在防御结构上的差异,发现传统攻击成功率指标无法捕捉的模型防御模式。

详情

展开后加载摘要…

URL PDF HTML 收藏