A Polya Urn Document Language Model for Improved Information Retrieval
专题命中 其他LLM :language model(title,abstract)
Comments 37 page journal submission (accepted for publication in TOIS)
AI 大模型
大语言模型、预训练、指令微调、后训练和语言模型应用。
专题命中 其他LLM :language model(title,abstract)
Comments 37 page journal submission (accepted for publication in TOIS)
专题命中 其他LLM :language model(title,abstract)
Journal ref Journal Of Artificial Intelligence Research, Volume 41, pages 367-395, 2011
专题命中 其他LLM :language model(title,abstract)
Comments In Proceedings of Quantum Interaction 2013
专题命中 其他LLM :SLM(title,abstract)
Comments arXiv admin note: text overlap with arXiv:1307.5667, arXiv:1307.5840
专题命中 其他LLM :LLM(title,abstract)
Comments 1+30 pages, footnote added
Journal ref JHEP 1212 (2012) 023
专题命中 其他LLM :SLM(title,abstract)
专题命中 其他LLM :SLM(title,abstract)
专题命中 其他LLM :language model(title,abstract)
Comments Published as CTIT Technical Report 05-35
专题命中 其他LLM :LLM(title,abstract)
Comments 15 pages, v2. minor improvements
Journal ref JHEP 1104:002,2011
专题命中 其他LLM :LLM(title,abstract)
Comments 35 pages
Journal ref JHEP 1101:067,2011
专题命中 其他LLM :LLM(title,abstract)
Comments 32 pages, 1 figure; v.2: references added, the relation between the level shift and filling fraction elaborated
Journal ref Nucl.Phys. B731 (2005) 285-308
响应性验证:预测会改变吗?改变多少?改变频率如何?
专题命中 其他LLM :LLM(summary_cn,abstract_cn);分类 cs.LG
AI总结 本研究提出测量响应性的算法与交互模型框架,搭配统计保证,可用于检测累犯预测的排除情况、估计内容审核博弈成本、测试LLM基准鲁棒性,提升模型安全性与可靠性。
MUON优化:从非收敛性到结合Polar Express与Newton-Schulz多项式实现的误差分析
专题命中 其他LLM :LLM(abstract_cn);large language model(abstract);language model(abstract);分类 cs.LG
AI总结 本文针对2024年提出的MUON优化器,提出含任意NS步骤的广义变体,证明其在部分随机优化问题中无法收敛,建立误差分析并在多个实例中验证相关结论。
Comments 82 pages
由大语言模型(LLMs)驱动的上下文感知文化遗产指南
专题命中 其他LLM :LLM(abstract);large language model(abstract);language model(abstract);分类 cs.AI
AI总结 本研究扩展了文化遗产网页应用Triangolazioni,构建与LLMs无关的松耦合架构,实现LLMs驱动的上下文感知文化遗产信息搜索与展示,丰富了文化遗产内容。
Journal ref In Proceedings of the 34th ACM Conference on User Modeling, Adaptation and Personalization (UMAP '26). 2026
通过门控语义质量多样性实现自我进化智能体的驾驭
专题命中 其他LLM :large language model(abstract);language model(abstract);foundation model(abstract);分类 cs.CL
AI总结 研究大语言模型智能体驾驭因素提升性能的问题,提出自我进化智能体驾驭框架,将提出与评估更改分开,通过门控存档避免过度拟合,在多领域测试中取得较好泛化效果,证明诊断与评估循环的有效性。
Comments 9 pages, 4 figures, 3 tables
机构 * University of Naples Federico II(那不勒斯费德里科二世大学)
专题命中 其他LLM :LLM(abstract);large language model(abstract);language model(abstract);分类 cs.CL
Comments There are 14 pages, 8 figures
奇数圈香农容量的改进下界
机构 * Worcester Polytechnic Institute(沃斯特理工学院)
专题命中 其他LLM :LLM(abstract);large language model(abstract);language model(abstract);分类 cs.AI
AI总结 研究通过与大语言模型迭代交互,在奇数圈\(C_7^{10}\)、\(C_{11}^{6}\)、\(C_{13}^{6}\)中构造独立集,改进了这些图香农容量的已知下界,还改进了几个奇数圈单个强幂独立数的已知下界。
Comments v2: added improvement on lower bound for the Shannon capacity of C15
超越未通过测试:协同生成的错误重现测试与修复的迭代强化
机构 * Nanjing University(南京大学) ; Peking University(北京大学) ; Microsoft(微软公司)
专题命中 其他LLM :LLM(abstract);large language model(abstract);language model(abstract);分类 cs.AI
AI总结 研究自动化程序修复中错误重现测试问题,指出仅用未通过到通过标准不足。提出CoHarden框架,先生成测试再迭代强化测试与修复,实验证明该框架在解决率等方面优于现有基线。
Comments 29 pages, 5 figures, preprint
向量化语言模型:基于智能向量化的人工智能助手
机构 * New York Institute of Technology(纽约理工学院)
专题命中 其他LLM :LLM(abstract);large language model(abstract);language model(abstract);分类 cs.AI
AI总结 研究旨在设计基于谷歌开放权重语言模型的VectorizationLLM,用于辅助学生学习MATLAB相关知识。采用RAG知识库和系统提示架构,不直接给答案,通过课堂笔记示例详细解释概念,为课程应用提供有指导意义的智能助手。
Comments 44 pages, 6 figures
学习社会规范可增强动态人机协作中的兼容性
机构 * School of Vehicle and Mobility, Tsinghua University, Beijing, China(清华大学车辆与移动系统学院) ; State Key Laboratory of Intelligent Green Vehicle and Mobility, Tsinghua University, Beijing 100084, China(清华大学智能绿色车辆与移动系统国家重点实验室) ; State Key Laboratory of Cognitive Neuroscience and Learning & IDG/McGovern Institute for Brain Research, Beijing Normal University, Beijing, China(北京师范大学认知神经科学与学习国家重点实验室) ; Beijing Key Laboratory of Safe AI and Superalignment, Beijing, China(北京安全人工智能与超对齐关键实验室) ; Beijing Institute of AI Safety and Governance, Beijing, China(北京人工智能安全与治理研究院)
专题命中 其他LLM :LLM(abstract);large language model(abstract);language model(abstract);分类 cs.AI
AI总结 研究以行人与车辆互动为代表,搭建实验平台,识别出人类社会规范的三项原则,将其融入人工智能显著改善人机协作,在闭环互动任务中,融入社会规范的大语言模型总分比基线策略高出近四倍,比人际互动高出43%。
Comments 44 pages, 5 figures, supplementary information included
HSEvo: 通过多样性驱动的和谐搜索和遗传算法提升自动启发式设计
机构 * Hanoi University of Science and Technology(河内科学技术大学) ; George Mason University(乔治·梅森大学)
专题命中 其他LLM :LLM(abstract);large language model(abstract);language model(abstract);分类 cs.AI
AI总结 本文提出HSEvo框架,通过结合和谐搜索算法平衡多样性与收敛性,解决自动启发式设计中探索与利用的平衡问题,提升搜索效率与稳定性。
Comments 18 pages, 12 figures
面向可靠WebGIS开发的代理人工智能双螺旋治理方法
机构 * Geographic Information Systems Center, Florida International University(佛罗里达国际大学地理信息系统中心) ; Geospatial Analytics, Technology and Open Research Lab, University of Florida(佛罗里达大学地理空间分析、技术与开放研究实验室)
专题命中 其他LLM :LLM(abstract,abstract_cn);prompting(abstract);分类 cs.AI
AI总结 针对代理AI在WebGIS开发中的不可靠问题,提出双螺旋治理框架,通过知识-行为-技能三轨架构和持久知识图谱外化事实与协议,实验证明能降低代码复杂度、减少输出方差并防止制图错误。
Comments Paper submitted to and under review in Transactions in GIS
通过概率建模对上下文学习的理论解释
机构 * Tsinghua Shenzhen International Graduate School, Tsinghua University, Shenzhen, China(清华大学深圳国际研究生院,清华大学,深圳,中国)
专题命中 其他LLM :LLM(abstract_cn);large language model(abstract);language model(abstract);分类 cs.LG
AI总结 本文提出上下文学习的概率模型,推导其在一般参数分布和指数族下的性能,并解释示例数量、模型参数敏感性和示例-查询相似性对性能的影响。
LLMs能否模拟人类行为变异?一项在音素流畅任务中的案例研究
专题命中 其他LLM :LLM(abstract);large language model(abstract);language model(abstract);分类 cs.CL
AI总结 研究探讨LLMs在音素流畅任务中是否能模拟个体差异,发现尽管部分模型能模拟平均值和词汇偏好,但无法再现人类行为变异范围,且新模型和思考模式反而降低了多样性。
Journal ref Proceedings of the 15th Workshop on Cognitive Modeling and Computational Linguistics (2026) 250-263
通过希尔伯特空间嵌入的量子最大似然预测
机构 * L2S, CNRS, CentraleSupélec, University of Paris-Saclay, France(L2S、CNRS、CentraleSupélec、巴黎-萨克雷大学、法国)
专题命中 其他LLM :LLM(abstract_cn);large language model(abstract);language model(abstract);分类 cs.LG
AI总结 研究量子最大似然预测任务,通过将经验概率分布嵌入量子态并最小化量子相对熵,提出统一框架,给出非渐近性能保证。
Comments 38+4 pages, 1 figure
通过文本正则化与信号聚合稳定黑盒提示优化
专题命中 其他LLM :LLM(abstract_cn);large language model(abstract);language model(abstract);分类 cs.LG
AI总结 提出TRAS框架,利用成功预测的文本正则化与蒙特卡洛信号聚合,解决黑盒提示优化中的不稳定性和语义漂移,提升准确率与收敛速度。
约束代码生成中的对齐问题
机构 * University of St. Gallen(圣加尔登大学) ; Università della Svizzera italiana (USI)(瑞士联邦理工学院)
专题命中 其他LLM :LLM(abstract_cn);large language model(abstract);language model(abstract);分类 cs.LG
AI总结 研究约束解码中约束器、语言模型与目标语言之间的对齐问题,发现约束器的不完整性会扭曲模型分布,导致功能正确性下降高达97%,并提出设计约束器的定量见解。
面向3D视觉定位的多样化语言生成扩展
机构 * Simon Fraser University(西蒙菲莎大学)
专题命中 其他LLM :LLM(summary_cn,abstract_cn);分类 cs.CL
AI总结 提出ViGiL3D++方法,通过场景图约束采样与LLM语言生成结合,生成多样化视觉定位查询,提升3DVG模型泛化能力并揭示VLM局限性。
Comments 39 pages, 14 figures, 16 tables. Project Page: https://3dlg-hcvc.github.io/vigil3dpp
近乎智能的革命:扩大审议规模并利用AI赋能人类的选项
机构 * Centre on Participatory and Deliberative Democracy(参与性和协商性民主研究中心)
专题命中 其他LLM :LLM(abstract_cn);large language model(abstract);language model(abstract);分类 cs.CL
AI总结 探讨大型语言模型如何通过系统功能语言学视角扩大民主审议规模,增强包容性并赋权边缘群体,同时警惕过度承诺与低估风险。
Comments Published in /Handbook of Democracy in the Era of Artificial Intelligence/ edited by Evangelos Pournaras, Srijoni Majumdar, Carina Ines Hausladen, and Dirk Helbing. 2026
超越通过/失败:使用过程挖掘理解LLM如何抵抗(和失败)红队攻击
机构 * MuyVentive LLC
专题命中 其他LLM :LLM(title_cn,abstract_cn);分类 cs.AI
AI总结 提出将过程挖掘应用于红队攻击轨迹,通过分析事件日志提取直接跟随图和状态转移矩阵,揭示GPT-OSS和Llama 3.3在防御结构上的差异,发现传统攻击成功率指标无法捕捉的模型防御模式。