arXivDaily arXiv每日学术速递 周一至周五更新

AI 大模型

代码大模型 / AI 编程

代码生成、软件工程智能体、程序修复、测试生成和开发者工具。

2026-01-08 至 2026-01-08 共收录 3 信号源:cs.SE, cs.CL, cs.AI, cs.LG, cs.PL

1. 代码生成 3 篇

2601.03878 2026-01-08 cs.SE 79%

Understanding Specification-Driven Code Generation with LLMs: An Empirical Study Design

通过LLM理解基于规范的代码生成:一项实证研究设计

Giovanni Rosa, David Moreno-Lumbreras, Gregorio Robles, Jesús M. González-Barahona

专题命中 代码生成 :code generation(title,abstract);分类 cs.SE

AI总结 本文通过实证研究设计,探讨人类干预在基于规范的LLM代码生成过程中对代码质量和动态的影响。

Comments This paper is a Stage 1 Registered Report. The study protocol and analysis plan were peer reviewed and accepted at SANER 2026 with a Continuity Acceptance (CA) score for Stage 2

详情

展开后加载摘要…

URL PDF HTML 收藏
2601.03780 2026-01-08 cs.SE 79%

Assessing and Improving the Representativeness of Code Generation Benchmarks Using Knowledge Units (KUs) of Programming Languages -- An Empirical Study

基于编程语言知识单元(KUs)评估和改进代码生成基准的代表性——一项实证研究

Md Ahasanuzzaman, Bram Adams, Emad Fallahzadeh, Gustavo A. Oliva, Ahmed E. Hassan

专题命中 代码生成 :code generation(title,abstract);分类 cs.SE

AI总结 本文通过分析编程语言知识单元(KUs)的覆盖情况,提出基于提示的框架改进代码生成基准的代表性,从而更准确评估LLM的代码生成能力。

详情

展开后加载摘要…

URL PDF HTML 收藏
2601.03640 2026-01-08 cs.SE cs.CR 79%

Verbatim Data Transcription Failures in LLM Code Generation: A State-Tracking Stress Test

LLM代码生成中的verbatim数据转录失败:一种状态跟踪压力测试

Mohd Ariful Haque, Kishor Datta Gupta, Mohammad Ashiqur Rahman, Roy George

专题命中 代码生成 :code generation(title,abstract);分类 cs.SE

AI总结 本文提出了一种最小化的基准测试,用于评估LLM在生成代码时对verbatim数据转录的可靠性,通过精确字符串包含和状态跟踪分析来检测长周期生成失败。

详情

展开后加载摘要…

URL PDF HTML 收藏