超过一半的近期天文学论文借助语言模型辅助撰写
More than half of recent astronomy papers are written with language-model assistance
浏览论文内容
中文总结 AI 辅助
该研究通过分析207,111篇astro-ph论文全文中的语言模型特征词,利用分层贝叶斯模型估计出2025年超过一半(54%)的天文学论文借助语言模型辅助撰写,而仅0.81%的论文明确披露。此方法可追踪AI辅助写作的普及趋势及其随时间衰减的可检测性。
中文摘要 AI 辅助
语言模型在其辅助撰写的文本中留下了独特的词汇痕迹,我们测量了当前天文学文献中有多少带有这种痕迹。基于2015年至2026年中期207,111篇astro-ph论文的全文,我们统计每篇论文中这些词汇的出现次数,并将其计数按论文长度比例建模为辅助与非辅助写作的混合,采用分层贝叶斯模型。2020年之前的论文用于校准非辅助写作的比例,而392篇披露使用模型的论文则用于校准辅助写作的比例。我们的答案取决于如果没有人使用模型,这些词汇如今出现的频率,这一比例必须通过建模而非直接观测获得,因此我们在三种假设下将其外推至2020年之后,并报告所有三种结果。对于2025年,这给出54%(统计误差,95%置信区间为-8至+8;系统背景误差为-0至+26)的论文,第二个误差为三种假设之间的差异。当我们改变该选择、校准方式以及仅允许采纳率上升的要求时,估计值保持在36%或以上。基于astro-ph语料库构建的词汇表,仅保留在所有子领域均上升的词汇,使2025年的估计值保持在相同范围内。辅助写作也变得越来越难以察觉,因为作者会适应那些暴露痕迹的词汇,且标记过量在2023年至2026年间减少了一半以上。我们的模型允许这种衰减,因此能够将更微弱的痕迹与使用减少区分开来。因此,超过一半的近期astro-ph论文带有语言模型痕迹,而2025年论文中仅有0.81%披露了这一点,即每约66篇带痕迹的论文中只有一篇声明。
英文摘要
Language models leave a distinctive vocabulary in the prose they help write, and we measure how much of the astronomy literature now carries it. From the full text of 207,111 astro-ph papers spanning 2015 to mid-2026, we count those words in each paper and model the counts, in proportion to paper length, as a mixture of assisted and unassisted writing in a hierarchical Bayesian model. Papers from before 2020 calibrate the unassisted rate, and the 392 papers that disclose model use calibrate the assisted one. Our answer depends on how often these words would appear today if nobody used a model, a rate that must be modeled rather than observed, so we extend it past 2020 under three assumptions and report all three. For 2025 that gives $54^{+8}_{-8}\,(\mathrm{stat},\,95\%)\,^{+26}_{-0}\,(\mathrm{sys,\ background})$% of papers, the second error being the spread across the three. The estimate stays at or above 36% when we vary that choice, the calibration, and the requirement that adoption only rises. A word list built from the astro-ph corpus, keeping only words that rose across every subfield, leaves 2025 in the same range. Assisted writing is also getting harder to see, since authors adapt to the words that reveal it and the marker excess more than halves between 2023 and 2026. Our model allows for that fading, so it can separate a fainter trace from reduced use. More than half of recent astro-ph papers therefore carry a language-model trace, while only 0.81% of 2025 papers disclose it, one declaration for every $\sim$66 papers with a trace.
发表机构
- The Ohio State University(俄亥俄州立大学)
- Max-Planck-Institut für Astronomie(马克斯·普朗克天文研究所)
机构由 AI 辅助整理,请以论文原文为准。