arXivDaily arXiv每日学术速递 周一至周五更新
arXiv周末暂无论文更新,休息一下吧,周末愉快~~

询问馆长:用HERITRACE展示专家驱动的RDF数据管理

Ask the Curator: Demonstrating Expert-Driven RDF Data Curation with HERITRACE

Arcangelo Massari, Silvio Peroni

arXiv 2607.22348首次发表:更新:

AI 中文总结

针对文献歧义案例,利用HERITRACE开源Web应用程序,技术人员配置、馆长编辑数据并记录出处,经处理重复建议、纠正错误及验证DOI等操作,展示了专家驱动的RDF数据管理过程。

AI 中文摘要

HERITRACE是一个用于管理三元组存储中RDF数据的开源Web应用程序,它还会记录出处和更改跟踪。它与数据模型无关:技术人员通过编写SHACL形状和YAML显示规则为数据集进行配置,之后领域专家通过生成的表单编辑数据,每次更改都会成为一个出处快照,可进行检查和恢复。本文以文献歧义的记录案例追溯了从配置到管理的过程:两篇不相关文章在PubMed下具有相同DOI。展示了技术人员如何准备一个小的OpenCitations Meta子集进行管理,以及馆长如何处理重复建议来合并文章,尽管元数据冲突仍意外确认进一步合并,通过时光机纠正错误,并在对照Crossref验证后更正DOI。

英文摘要

HERITRACE is an open-source Web application for curating RDF data held in triplestores, with provenance and change tracking recorded in RDF as well. It is agnostic to the data model: a technician configures it for a collection by writing SHACL shapes and YAML display rules, after which domain experts edit the data through generated forms, and every change becomes a provenance snapshot that can be inspected and restored. This paper traces the path from configuration to curation on a documented case of bibliographic ambiguity: two unrelated articles that PubMed records under the same DOI. We show how a technician prepares a small OpenCitations Meta subset for curation and how a curator acts on duplicate suggestions to merge two articles, inadvertently confirms a further merge despite conflicting metadata, reverts the mistake through the Time Machine, and corrects the DOI after verification against Crossref.

Comments7 pages, 2 figures. Submitted to the International Semantic Web Conference (ISWC) 2026 Posters and Demos Track

论文原文

arXiv 摘要页 · PDF 原文 · HTML 原文

↑