arXivDaily arXiv每日学术速递 周一至周五更新
arXiv周末暂无论文更新,休息一下吧,周末愉快~~
arXiv 2608.19838cs.AIcs.PL

规范增量驱动的数据治理:以“规范增量(spec-delta)”作为湖仓数据平台变更单元的实证研究

Specification-delta-driven data governance: an empirical study of the «spec-delta» as the unit of change in lakehouse data platforms

Pablo Ramirez Amador

AI总结:

本研究通过对照实验,验证以spec-delta为变更单元的数据治理工作流,可减少缺陷、降低评审认知负荷,提供避免过度规范的指南,而非开发工具。

AI中文摘要:

规范驱动开发(SDD)已确立“规范而非代码应是AI辅助工作的主要制品”这一理念。GitHub Spec Kit等工具以及Constitutional SDD等提案已在软件领域形式化该原则,而可执行数据契约文献则将其扩展至运行时的模式与质量执行。不过,将规范增量(OpenSpec的核心思想,即每次变更应产生可评审的需求增量)作为数据平台变更单元的处理方式仍缺乏实证探索,尽管许多数据平台变更属于契约类(新数据集、服务水平协议、指标语义、访问策略)而非纯代码变更。本研究形式化了spec-delta概念,提出了数据平台变更按其对增量规范的适用性分类的分类法,并定义了对照实验,将spec-delta驱动的工作流与无增量的传统代码拉取请求工作流进行比较。响应变量包括发现到部署的时间、到达Silver和Gold湖仓层的缺陷密度、跨工具指标差异,以及用NASA TLX测量的评审者认知负荷。论文明确预留了演示与实验室部分以在真实湖仓环境中实例化。本研究的贡献并非工具,而是可复现的证据与适用性指南,有助于避免前期过度规范的反模式。

英文摘要:

Spec Driven Development SDD has consolidated the idea that the specification rather than the code should be the primary artefact governing AI assisted work. Tools such as GitHub Spec Kit, and proposals such as Constitutional SDD, have formalised this principle in the software domain, while the executable data-contracts literature has extended it to schema and quality enforcement at run time. Nevertheless, the treatment of the specification delta OpenSpec's core idea that every change should produce a reviewable increment of requirements as the unit of change in data platforms remains empirically unexplored, even though many data-platform changes are contractual (new datasets, service-level agreements, metric semantics, access policies) rather than purely code changes. This work formalises the spec-delta concept, proposes a taxonomy of data platform changes according to their suitability for incremental specification, and defines a controlled experiment comparing a spec-delta-driven workflow against a conventional code pull-request workflow without a delta. The response variables are discovery to deployment time, the density of defects reaching the Silver and Gold lakehouse layers, cross-tool metric divergence, and reviewer cognitive load measured with NASA TLX. The paper explicitly reserves a demonstration-and-laboratory section for instantiation on a real lakehouse environment. The contribution is not a tool but reproducible evidence and an applicability guide that helps to avoid the up front over specification antipattern.

↑