机构
*
Nankai University(南开大学)
;
James Cook University(詹姆斯库克大学)
;
Western Sydney University(西悉尼大学)
;
Beijing University of Technology(北京工业大学)
;
Fuzhou University(福州大学)
;
Nanjing University of Science and Technology(南京理工大学)
;
CSIRO's Data 61(澳大利亚联邦科学与工业研究组织Data61)
;
The University of Adelaide(阿德莱德大学)
Blockchain Infrastructure for Intelligent Cyber--Physical--Social Systems:Post-Quantum Security, Interoperability, and Trustworthy Data Economies in the Era of Embodied AI
面向智能信息-物理-社会系统的区块链基础设施:具身AI时代的后量子安全、互操作性与可信数据经济
Song Guo, Huawei Huang, Dongping Liu, Aoyu Zhang, Luyao Zhang
机构
*
Hong Kong University of Science and Technology(香港理工大学)
;
Sun Yat-sen University(中山大学)
;
Amazon Web Services(亚马逊网络服务)
;
Duke Kunshan University(杜克昆山大学)
Causality Is Key to Understand and Balance Multiple Goals in Trustworthy ML and Foundation Models
因果关系是理解和平衡可信机器学习与基础模型中多个目标的关键
Ruta Binkyte, Ivaxi Sheth, Zhijing Jin, Mohammad Havaei, Bernhard Schölkopf, Mario Fritz
机构
*
CISPA Helmholtz Center for Information Security(CISPA海德堡信息安全中心)
;
Max Planck Institute for Intelligent Systems, Tübingen(马克斯·普朗克智能系统研究所(图宾根))
;
Google Research(谷歌研究)
;
ETH Zürich(苏黎世联邦理工学院)
;
University of Toronto(多伦多大学)
Domain-Conditioned Safety in Frontier Computer-Using Agents: A 793-Episode Browser Benchmark, a Coding-Domain Cross-Reference, and a Reproducibility Audit of Recent Red-Teaming
Toward AI That Understands Self and Others: A World-Model Theory of Cognitive Diversity and Alignment
迈向理解自我与他人的AI系统:人类认知多样性与世界模型对齐的多阶段推理框架
Toru Takahashi
机构
*
Human Informatics and Systems Lab, Doshisha University(立命馆大学人机系统实验室)
;
Linked Open Data Initiative, NPO Keio Research Institute at SFC(庆应义塾大学SFC研究所开放数据计划)
;
Stroly Inc(Stroly公司)
Comments87 pages. Revised version with a refined abstract emphasizing disagreement as a late-stage phenomenon, target admissibility, processability, and the methodological abstraction used to compare humans, AI systems, and institutional decision procedures under shared information-theoretic constraints
Repurposing Adversarial Perturbations for Continual Learning: From Defense to Active Alignment
重新利用对抗扰动进行持续学习:从防御到主动对齐
Ran Liu, Min Yu, Mingqi Liu, Jianguo Jiang, Gang Li, Rongsheng Li, Ning Li, Zhen Xu, Weiqing Huang, Ming Liu
机构
*
Institute of Information Engineering, Chinese Academy of Sciences(中国科学院信息工程研究所)
;
School of Cyber Security, University of Chinese Academy of Sciences(中国科学院大学网络安全学院)
;
Deakin University(德肯大学)
;
Harbin Engineering University(哈尔滨工程大学)
dashi: A Python library for Dataset Shift Characterization to Support Trustworthy AI Development and Deployment
dashi: 一个用于数据集偏移表征以支持可信AI开发和部署的Python库
David Fernández-Narro, Pablo Ferri, Ángel Sánchez-García, Juan M. García-Gómez, Carlos Sáez
机构
*
Biomedical Data Science Lab, Instituto Universitario de Tecnologías de la Información y Comunicaciones, Universitat Politècnica de Valéncia(生物医学数据科学实验室,信息与通信技术大学,巴塞罗那理工大学)
Modeling Community Attitude through Reaction Tone: A Human-AI Collaborative Framework for Evaluating LLM Alignment with Linguistic Behaviors in Online Communities
通过反应语气建模社区态度:评估LLM与在线社区语言行为对齐的人机协作框架
Nuan Wen, Xuezhe Ma
机构
*
Information Sciences Institute University of Southern California(南加州大学信息科学研究所)
机构
*
Shanghai Academy of Artificial Intelligence for Science, Shanghai, China.(上海人工智能科学研究院)
;
School of Biomedical Engineering, Shanghai Jiao Tong University, Shanghai, China.(上海交通大学生物医学工程学院)
;
Incubation Institute, Fudan University, Shanghai, China.(复旦大学孵化院)