arXivDaily arXiv每日学术速递 周一至周五更新
arXiv周末暂无论文更新,休息一下吧,周末愉快~~

TRACTOR基准:用于评估C到Rust翻译器

TRACTOR Benchmark for Evaluating C to Rust Translators

Hamed Okhravi, Brandt Ogden, Noah Luther, Ian McQuoid, Howard Mak, Nathan Burow

arXiv 2609.25121首次发表:更新:

发表机构

Massachusetts Institute of Technology(麻省理工学院)

机构由 AI 辅助整理,请以论文原文为准。

AI 中文总结

TRACTOR基准由MIT林肯实验室开发,用于系统评估C到Rust翻译工具,包含渐进式测试组和里程碑项目,并提供评估正确性、安全性、惯用性和性能的公开基础设施。

AI 中文摘要

内存安全漏洞仍然是关键软件中持续存在的安全风险来源,其中大部分软件是用C和C++等内存不安全语言实现的。编程语言、程序分析和人工智能的最新进展为通过自动翻译到Rust等内存安全语言来实现这些遗留系统的现代化创造了新的机会。DARPA的“将所有C翻译为Rust”(TRACTOR)项目旨在开发可扩展的技术,将大型C代码库翻译成安全、高性能且可维护的Rust代码。麻省理工学院林肯实验室作为该项目的独立测试与评估组织,开发了一个标准化基准,用于系统评估C到Rust的翻译工具。本报告描述了TRACTOR基准,包括逐步增加难度的测试组和更大的里程碑项目,以及用于评估正确性、安全性、惯用性和性能的配套评估基础设施和指标。该基准及相关的评估基础设施已公开,以支持C到Rust翻译技术的更广泛开发和评估。

英文摘要

Memory-safety vulnerabilities remain a persistent source of security risk in critical software, much of which is implemented in memory-unsafe languages such as C and C++. Recent advances in programming languages, program analysis, and artificial intelligence have created new opportunities to modernize these legacy systems through automated translation to memory-safe languages such as Rust. The DARPA Translating All C to Rust (TRACTOR) program seeks to develop scalable techniques for translating large C codebases into safe, performant, and maintainable Rust. MIT Lincoln Laboratory serves as the program's independent test and evaluation organization and has developed a standardized benchmark for systematically assessing C-to-Rust translation tools. This report describes the TRACTOR benchmark, including progressively challenging test batteries and larger milestone projects, as well as the supporting evaluation infrastructure and metrics for assessing correctness, safety, idiomaticity, and performance. The benchmark and associated evaluation infrastructure are publicly available to support the broader development and evaluation of C-to-Rust translation technologies.

论文原文

arXiv 摘要页 · PDF 原文 · HTML 原文

↑