FOXDEN:面向AI就绪型科学数据集的FAIR服务
FOXDEN: FAIR Services for AI-Ready Scientific Datasets
浏览论文内容
中文总结 AI 辅助
该研究介绍了FOXDEN这一基于FAIR原则的网络基础设施,可对科学数据集添加元数据与溯源记录并分配DOI,已在两处机构部署,将为未来智能体科学工作流提供基础。
中文摘要 AI 辅助
当科学数据集遵循FAIR指导原则,由丰富的机器可读元数据和溯源信息进行描述时,最适配AI工作流。FAIR开放科学可扩展数据交换网络(FOXDEN)是康奈尔高能同步辐射光源(CHESS)开发的一组网络基础设施构建块,用于为原始、归约和分析后的数据集添加结构化与非结构化元数据及溯源记录注释,还允许研究人员为这些记录分配数字对象标识符(DOI)以创建AI就绪型数据集。本文描述了FOXDEN的架构及其在CHESS和国家高磁场实验室的部署情况,该部署将作为未来智能体科学工作流的基础。
英文摘要
Scientific datasets are most compatible with AI workflows when they are described by rich, machine-readable metadata and provenance information, following the FAIR guiding principles. The FAIR Open-Science Extensible Data Exchange Network (FOXDEN) is a set of cyberinfrastructure building blocks developed at the Cornell High Energy Synchrotron Source (CHESS) for annotating raw, reduced, and analyzed datasets with both structured and unstructured metadata and provenance records. It also allows researchers to publish these records with Digital Object Identifiers (DOIs) to create AI-ready datasets. We describe FOXDEN's architecture and its deployment at CHESS and at the National High Magnetic Field Laboratory, where it serves as a foundation for future agentic scientific workflows.
发表机构
- Cornell High Energy Synchrotron Source, Cornell University(康奈尔高能同步辐射光源,康奈尔大学)
- National High Magnetic Field Laboratory(国家高磁场实验室)
机构由 AI 辅助整理,请以论文原文为准。