arXivDaily arXiv每日学术速递 周一至周五更新

AI 大模型

语言大模型 / LLM

大语言模型、预训练、指令微调、后训练和语言模型应用。

2026-01-06 至 2026-01-06 共收录 10 信号源:cs.CL, cs.AI, cs.LG

1. 其他LLM 10 篇

2512.24478 2026-01-06 cs.LG cs.AI stat.ME 90%

HOLOGRAPH: Active Causal Discovery via Sheaf-Theoretic Alignment of Large Language Model Priors

HOLOGRAPH:通过sheaf理论对大型语言模型先验进行对齐以实现主动因果发现

Hyunjun Kim

机构 * Korea Advanced Institute of Science \'Ecole Polytechnique F\'ed\'erale de Lausanne (EPFL), Lausanne, Switzerland

专题命中 其他LLM :large language model(title,abstract);language model(title,abstract);LLM(abstract);分类 cs.AI、cs.LG

AI总结 HOLOGRAPH通过sheaf理论对齐大型语言模型先验,实现主动因果发现,提供严谨的数学基础并实现竞争性性能。

详情

展开后加载摘要…

URL PDF HTML 收藏
2601.00810 2026-01-06 q-fin.PM cs.AI cs.LG econ.GN q-fin.EC q-fin.ST 90%

Can Large Language Models Improve Venture Capital Exit Timing After IPO?

大语言模型能否在IPO后改善风险投资退出时机?

Mohammadhossien Rashidi

专题命中 其他LLM :large language model(title,abstract);language model(title,abstract);LLM(abstract);分类 cs.AI、cs.LG

AI总结 本研究利用大语言模型分析IPO后财务数据,预测风险投资退出时机,并评估AI指导对退出决策的经济影响。

详情

展开后加载摘要…

URL PDF HTML 收藏
2601.01954 2026-01-06 cs.SE 89%

Reporting LLM Prompting in Automated Software Engineering: A Guideline Based on Current Practices and Expectations

报告LLM提示在自动化软件工程中的使用:基于当前实践和期望的指南

Alexander Korn, Lea Zaruchas, Chetan Arora, Andreas Metzger, Sven Smolka, Fanyu Wang, Andreas Vogelsang

专题命中 其他LLM :LLM(title,abstract);prompting(title);large language model(abstract);language model(abstract)

AI总结 本文提出了一项基于当前实践和期望的指南,旨在提高LLM在自动化软件工程中的透明度、可重复性和方法学严谨性。

Comments To be published at The 3rd ACM International Conference on AI Foundation Models and Software Engineering FORGE 2026

详情

展开后加载摘要…

URL PDF HTML 收藏
2601.00828 2026-01-06 cs.AI 84%

Decomposing LLM Self-Correction: The Accuracy-Correction Paradox and Error Depth Hypothesis

分解大语言模型的自我纠正:准确性-纠正悖论与错误深度假说

Yin Li

机构 * University of Birmingham(伯明翰大学)

专题命中 其他LLM :LLM(title,comments);large language model(abstract);language model(abstract);分类 cs.AI

AI总结 本研究揭示大语言模型在自我纠正中的准确性-纠正悖论,提出错误深度假说,发现更强模型犯更深入的错误,且错误检测与纠正成功率无直接关联。

Comments 9 pages, 2 figures, 3 tables. Code available at https://github.com/Kevin0304-li/llm-self-correction

详情

展开后加载摘要…

URL PDF HTML 收藏
2505.14435 2026-01-06 cs.CY cs.AI 83%

Choosing a Model, Shaping a Future: Comparing LLM Perspectives on Sustainability and its Relationship with AI

选择模型,塑造未来:比较LLM对可持续性和其与AI关系的视角

Annika Bush, Meltem Aksoy, Markus Pauly, Greta Ontrup

机构 * Research Center Trustworthy Data Science and Security, University Alliance Ruhr(可信数据科学与安全研究中心,鲁尔大学联盟) Department of Computer Science, Technical University Dortmund(计算机科学系,图林根技术大学) Chair of Mathematical Statistics and Applications in Industry, Technical University Dortmund(工业数学统计与应用教授职位,图林根技术大学) Department of Computer Science, University of Duisburg-Essen(计算机科学系,杜伊斯堡-埃森大学)

专题命中 其他LLM :LLM(title);large language model(abstract);language model(abstract);分类 cs.AI

AI总结 本研究比较了五个先进LLM对可持续性与AI关系的视角,发现模型间存在显著差异,强调模型选择对可持续性战略的影响。

Comments Accepted for EMNLP Conference

Journal ref Conference on Empirical Methods in Natural Language Processing (EMNLP 2025)

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.12421 2026-01-06 cs.SE cs.AI 77%

Understanding Prompt Management in GitHub Repositories: A Call for Best Practices

理解GitHub仓库中的提示管理:对最佳实践的呼吁

Hao Li, Hicham Masri, Filipe R. Cogo, Abdul Ali Bangash, Bram Adams, Ahmed E. Hassan

机构 * Queen’s University(女王大学) Huawei Technologies(华为技术) L L ahore University of Management Sciences(拉瓦尔管理科学大学)

专题命中 其他LLM :large language model(abstract);language model(abstract);foundation model(abstract);分类 cs.AI

AI总结 本研究通过分析GitHub仓库中的开源提示,揭示了提示管理中的关键挑战,并提出了提升提示软件可用性和可维护性的最佳实践建议。

详情

展开后加载摘要…

URL PDF HTML 收藏
2503.19004 2026-01-06 physics.ed-ph cs.CY quant-ph 67%

The Quantum Technology Job Market: Data Driven Analysis of 3641 Job Posts

量子技术就业市场:对3641份职位的驱动数据分析

Simon Goorney, Eleni Karydi, Borja Munoz, Otto Santesson, Zeki Can Seskir, Ana Alina Tudoran, Jacob Sherson

专题命中 其他LLM :large language model(abstract);language model(abstract)

AI总结 本研究通过分析3641份量子技术职位公告,揭示了该领域在北美地区的需求趋势及劳动力结构变化。

详情

展开后加载摘要…

URL PDF HTML 收藏
2601.00931 2026-01-06 cond-mat.supr-con 67%

AI-Guided Computational Design of a Room-Temperature, Ambient- Pressure Superconductor Candidate: Grokene

AI引导的室温常压超导体候选物Grokene的计算设计

DEARDAO DeSci Collaborative Team, Yanhuai Ding

专题命中 其他LLM :large language model(abstract);language model(abstract)

AI总结 利用AI和多体理论设计出Grokene,预测其在室温常压下具有超导性,但需通过实验验证并优化结构以提升临界温度。

详情

展开后加载摘要…

URL PDF HTML 收藏
2601.00965 2026-01-06 cs.LG cs.AI 62%

Adapting Feature Attenuation to NLP

适应特征衰减到NLP

Tianshuo Yang, Ryan Rabinowitz, Terrance E. Boult, Jugal Kalita

机构 * University of Michigan(密歇根大学) University of Colorado Colorado Springs(科罗拉多州立大学)

专题命中 其他LLM :language model(abstract);分类 cs.AI、cs.LG

AI总结 本文将特征衰减假设从计算机视觉移植到NLP,评估了COSTARR等方法在文本开放集识别中的表现,发现其在不重新训练的情况下效果有限,但指出了需要更大模型和定制策略的改进方向。

详情

展开后加载摘要…

URL PDF HTML 收藏
2601.02231 2026-01-06 eess.AS 50%

On the Role of Spatial Features in Foundation-Model-Based Speaker Diarization

在基于基础模型的语音辨识中空间特征的作用

Marc Deegen, Tobias Gburrek, Tobias Cord-Landwehr, Thilo von Neumann, Jiangyu Han, Lukáš Burget, Reinhold Haeb-Umbach

专题命中 其他LLM :foundation model(abstract)

AI总结 本文研究了在基于基础模型的语音辨识中引入空间特征的影响,发现虽然空间信息能提升性能,但其效果不如预期,因现有模型已能有效捕捉所需信息。

Comments Accepted at HSCMA 2026

详情

展开后加载摘要…

URL PDF HTML 收藏