arXivDaily arXiv每日学术速递 周一至周五更新

高校专区

Carnegie Mellon University(卡内基梅隆大学)

2025-12-04 至 2025-12-04 共收录 7
2512.04044 2025-12-04 cs.LG cs.AI cs.CR

MarkTune: Improving the Quality-Detectability Trade-off in Open-Weight LLM Watermarking

MarkTune: 改善开放权重语言模型水印中的质量-可检测性权衡

Yizhou Zhao, Zhiwei Steven Wu, Adam Block

机构 * University of Pennsylvania(宾夕法尼亚大学) Carnegie Mellon University(卡内基梅隆大学) Columbia University(哥伦比亚大学)

AI总结 MarkTune通过理论框架提升开放权重语言模型中质量与可检测性的平衡,优于GaussMark并保持生成质量。

详情

展开后加载摘要…

URL PDF HTML 收藏
2512.03579 2025-12-04 cs.LG math.PR math.ST stat.TH

Optimal Transportation and Alignment Between Gaussian Measures

高斯测度间的最优运输与对齐

Sanjit Dandapanthula, Aleksandr Podkopaev, Shiva Prasad Kasiviswanathan, Aaditya Ramdas, Ziv Goldfeld

机构 * Carnegie Mellon University, Department of Statistics(卡内基梅隆大学统计学系) Amazon Web Services (AWS) LogAnalytics(亚马逊网络服务(AWS)日志分析) Carnegie Mellon University, Machine Learning Department(卡内基梅隆大学机器学习系) Cornell University, Department of Electrical and Computer Engineering(康奈尔大学电子与计算机工程系)

AI总结 本文提出了一种针对高斯分布的最优运输与格罗莫夫-沃瑟斯坦对齐的闭式解,扩展到内积GW对齐,并应用于知识蒸馏和异构聚类。

详情

展开后加载摘要…

URL PDF HTML 收藏
2512.03399 2025-12-04 cs.LG

Full-Stack Alignment: Co-Aligning AI and Institutions with Thick Models of Value

全栈对齐:通过厚价值模型对齐人工智能与机构

Joe Edelman, Tan Zhi-Xuan, Ryan Lowe, Oliver Klingefjord, Vincent Wang-Mascianica, Matija Franklin, Ryan Othniel Kearns, Ellie Hain, Atrisha Sarkar, Michiel Bakker, Fazl Barez, David Duvenaud, Jakob Foerster, Iason Gabriel, Joseph Gubbels, Bryce Goodman, Andreas Haupt, Jobst Heitzig, Julian Jara-Ettinger, Atoosa Kasirzadeh, James Ravi Kirkpatrick, Andrew Koh, W. Bradley Knox, Philipp Koralus, Joel Lehman, Sydney Levine, Samuele Marro, Manon Revel, Toby Shorin, Morgan Sutherland, Michael Henry Tessler, Ivan Vendrov, James Wilken-Smith

机构 * Meaning Alignment Institute(意义对齐研究所) Massachusetts Institute of Technology(麻省理工学院) University College London(伦敦大学学院) University of Oxford(牛津大学) Western University(西方大学) University of Toronto(多伦多大学) McGill University(麦吉尔大学) Stanford University(斯坦福大学) Potsdam Institute for Climate Impact Research(波茨坦气候影响研究所) Yale University(耶鲁大学) Carnegie Mellon University(卡内基梅隆大学) UT Austin(德克萨斯大学奥斯汀分校) New York University(纽约大学) Harvard University(哈佛大学) Midjourney Core contributor(Midjourney核心贡献者)

AI总结 本文提出通过厚价值模型实现全栈对齐,以解决AI与机构目标不一致导致的不良后果,涵盖价值表示、规范推理和集体利益建模。

详情

展开后加载摘要…

URL PDF HTML 收藏
2512.03056 2025-12-04 cs.LG cs.AI

Delta Sampling: Data-Free Knowledge Transfer Across Diffusion Models

Delta Sampling: 数据无源的知识迁移跨扩散模型

Zhidong Gao, Zimeng Pan, Yuhang Yao, Chenyue Xie, Wei Wei

机构 * Shanxi University(山西大学) Google Cloud(谷歌云) Carnegie Mellon University(卡内基梅隆大学) University of Science and Technology of China(中国科学技术大学)

AI总结 Delta Sampling通过推理时利用模型预测差异实现跨不同架构基模型的知识迁移,无需原始训练数据。

详情

展开后加载摘要…

URL PDF HTML 收藏
2511.00293 2025-12-04 cs.CV

MagicView: Multi-View Consistent Identity Customization via Priors-Guided In-Context Learning

MagicView: 通过先验引导的上下文学习实现多视角一致的身份定制

Hengjia Li, Jianjin Xu, Keli Cheng, Lei Wang, Ning Bi, Boxi Wu, Fernando De la Torre, Deng Cai

机构 * Zhejiang University(浙江大学) Carnegie Mellon University(卡内基梅隆大学) Qualcomm Inc.(高通公司)

AI总结 MagicView 通过先验引导的上下文学习实现多视角一致的身份定制,有效提升多视角一致性与文本对齐能力。

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.18212 2025-12-04 cs.AI cs.LG

A Definition of AGI

AGI 的定义

Dan Hendrycks, Dawn Song, Christian Szegedy, Honglak Lee, Yarin Gal, Erik Brynjolfsson, Sharon Li, Andy Zou, Lionel Levine, Bo Han, Jie Fu, Ziwei Liu, Jinwoo Shin, Kimin Lee, Mantas Mazeika, Long Phan, George Ingebretsen, Adam Khoja, Cihang Xie, Olawale Salaudeen, Matthias Hein, Kevin Zhao, Alexander Pan, David Duvenaud, Bo Li, Steve Omohundro, Gabriel Alfour, Max Tegmark, Kevin McGrew, Gary Marcus, Jaan Tallinn, Eric Schmidt, Yoshua Bengio

机构 * Center for AI Safety(AI安全中心) University of California, Berkeley(加州大学伯克利分校) Virtue AI Morph Labs(Morph实验室) University of Michigan(密歇根大学) LG AI Research(LG人工智能研究) University of Oxford(牛津大学) Stanford University(斯坦福大学) University of Wisconsin–Madison(威斯康星大学麦迪逊分校) Gray Swan AI Carnegie Mellon University(卡内基梅隆大学) Cornell University(康奈尔大学) Hong Kong Baptist University(香港 Baptist大学) HKUST(香港科技大学) Nanyang Technological University(南洋理工大学) KAIST(韩国科学技术院) University of California, Santa Cruz(加州大学圣克鲁兹分校) Massachusetts Institute of Technology(麻省理工学院) University of Tübingen(图宾根大学) University of Washington(华盛顿大学) University of Toronto(多伦多大学) Vector Institute(向量研究所) University of Illinois Urbana-Champaign(伊利诺伊大学厄巴纳-香槟分校) Beneficial AI Research(有益AI研究) Conjecture Institute for Applied Psychometrics(应用心理测量研究所) New York University(纽约大学) CSER Université de Montréal(蒙特利尔大学) LawZero

AI总结 本文提出了一种基于卡特尔-霍恩-卡罗尔理论的可量化框架,定义AGI为与受过良好教育的成年人认知能力相匹配,并通过心理测量电池评估AI系统,揭示当前AI在基础认知机制上的不足。

详情

展开后加载摘要…

URL PDF HTML 收藏
2506.00195 2025-12-04 cs.CL cs.AI cs.HC

Let Them Down Easy! Contextual Effects of LLM Guardrails on User Perceptions and Preferences

让它们轻松一些!LLM护栏对用户感知和偏好的情境影响

Mingqian Zheng, Wenjia Hu, Patrick Zhao, Motahhare Eslami, Jena D. Hwang, Faeze Brahman, Carolyn Rose, Maarten Sap

机构 * Carnegie Mellon University(卡内基梅隆大学) Simon Fraser University(西蒙弗雷泽大学) Allen Institute for AI(人工智能研究所)

AI总结 研究探讨了LLM护栏对用户感知和偏好的影响,发现部分合规策略能显著降低负面感知,强调应通过创造性的拒绝策略而非意图检测来提升安全性和用户体验。

Comments Accepted to Findings of EMNLP 2025

详情

展开后加载摘要…

URL PDF HTML 收藏