Comments4 pages, 4 figures. Short paper accepted at the Workshop on Systems for Data-centric Agents with Human-in-the-loop (DASHSys 2026), co-located with VLDB 2026
BioSecBench-Refusal: A paired metric for performance and alignment in agentic biosecurity risk assessment
评估两用生物学环境中的校准拒绝和安全有用性
Edwin H. Wintermute, Harmon Bhasin, Christina M. Agapakis, Dianzhuo Wang, Evan Seeyave, Arjun Banerjee, Daniel Fulop, Matthew C. Watson, Adam J. Meyer, Sandrine Boissel, Jens H. Kuhn, Rishi Jain, Noah D. Taylor, Helena Shomar, Patrick M. Boyle, Kenny Workman
Commentsv4 adds 2 new appendices: app J evaluates a "generative" Bayesian alternative to BLF (which does worse than the "discriminative" approached used by BLF), and app K sketches how to do value of information computation to decide when to stop searching (although this idea has not been tried). v4 also adds a few more recent references
Journal refv3 was published in ICML AI Forecasting workshop 2026 (https://forecasting-workshop.github.io/)
Platonic Representations for Poverty Mapping: Unified Vision-Language Codes or Agent-Induced Novelty?
贫困地图绘制的柏拉图式表示:统一的视觉语言代码还是智能体诱导的新颖性?
Satiyabooshan Murugaboopathy, Connor T. Jerzak, Adel Daoud
机构
*
AI and Global Development Lab(人工智能与全球发展实验室)
;
Institute for Analytical Sociology(分析社会学研究所)
;
Institute of Computer Science(计算机科学研究所)
;
Department of Government(政府系)
机构
*
Shanghai Research Institute for Intelligent Autonomous Systems, Shanghai Institute of Intelligent Science and Technology, the State Key Laboratory of Autonomous Intelligent Unmanned Systems, and Frontiers Science Center for Intelligent Autonomous Systems of Ministry of Education, Tongji University(上海智能无人系统科学中心,上海智能科学与技术研究院,自主智能无人系统国家重点实验室,同济大学智能自主系统前沿科学中心)
;
Department of Automation, Shanghai Jiao Tong University(上海交通大学自动化系)
;
School of Aerospace Engineering and Institute for Robotics and Intelligent Machines, Georgia Institute of Technology(佐治亚理工学院航空航天工程学院和机器人与智能机器研究所)