arXivDaily arXiv每日学术速递 周一至周五更新
arXiv周末暂无论文更新,休息一下吧,周末愉快~~

小型研究软件团队管理IT灾难的十二条实用建议

Twelve Quick Tips for Managing IT Disasters in Small Research Software Teams

Greg Wilson

arXiv 2608.27196首次发表:更新:

AI 中文总结

本文为非资深系统管理员的小型研究软件团队提供十二条IT灾难规划与恢复建议,明确“完成”标准,并提示可借助机构相关团队协助。

AI 中文摘要

2025年,美国政府对自身科研团体发动了一系列前所未有的攻击。一年后,GitHub的可用性首次跌破90%,同时加拿大、法国、西班牙等地的野火迫使研究人员撤离家庭和实验室。这些事件及其他情况提醒我们,研究计算系统的脆弱性,而灾难规划是预防灾难的最有效方式之一。本文是针对小型研究软件团队的灾难规划与恢复的简短指南。这些建议假设你在日常工作之外还要亲力亲为,且并非经验丰富的系统管理员。部分建议需要此类专业知识,但大多数研究机构都有研究计算组、数据馆员以及环境健康与安全办公室,其职责恰好是协助解决这类问题。本文明确了“完成”的标准,这些机构通常能提供相关支持。

英文摘要

In 2025, the US government launched an unprecedented series of attacks on its own scientific research groups. A year later GitHub dropped below 90% availability for the first time, while wildfires in Canada, France, Spain, and elsewhere forced researchers from the homes and labs. These events and others have reminded us just how fragile research computing systems can be, and that planning for disasters is one of the most effective ways to prevent them. This paper is a short guide to disaster planning and recovery for a small research software team. The tips assume you are doing everything yourself on top of your regular job, and that you aren't an experienced system administrator. Some of the tips do require that kind of expertise, but most research institutions have research computing groups, data librarians, and environmental health-and-safety offices whose entire job is to help with exactly these problems. This paper tells you what "done" looks like; they can often provide it.

Comments11 pages

论文原文

arXiv 摘要页 · PDF 原文 · HTML 原文

↑