arXivDaily arXiv每日学术速递 周一至周五更新

期刊&会议

Transactions on Machine Learning Research · 期刊 · Machine Learning

2025-12-17 至 2025-12-17 共收录 3
2503.09330 2025-12-17 cs.LG cs.AI

Group-robust Machine Unlearning

组鲁棒机器去学习

Thomas De Min, Subhankar Roy, Stéphane Lathuilière, Elisa Ricci, Massimiliano Mancini

机构 * University of Trento(特伦托大学) University of Bergamo(贝加莫大学) Inria Grenoble(格勒诺布尔研究所) Univ. Grenoble Alpes(格勒诺布尔阿尔卑斯大学) Fondazione Bruno Kessler(布鲁诺·凯斯勒基金会)

AI总结 本文提出MIU方法,通过互信息感知实现组鲁棒的机器去学习,减少主导组性能损失,提升模型公平性。

Comments Published in Transactions on Machine Learning Research (TMLR)

详情

展开后加载摘要…

URL PDF HTML 收藏
2410.07553 2025-12-17 cs.AI

COMMA: A Communicative Multimodal Multi-Agent Benchmark

COMMA:一种基于通信的多模态多智能体基准

Timothy Ossowski, Danyal Maqbool, Jixuan Chen, Zefan Cai, Tyler Bradshaw, Junjie Hu

机构 * Department of Computer Sciences University of Wisconsin-Madison(计算机科学系威斯康星大学麦迪逊分校) Department of Computer Sciences UC San Diego(计算机科学系加州大学圣地亚哥分校) Department of Radiology University of Wisconsin-Madison(放射学系威斯康星大学麦迪逊分校) Department of Computer Sciences Department of Biostatistics and Medical Informatics University of Wisconsin-Madison(计算机科学系生物统计学与医学信息学系威斯康星大学麦迪逊分校)

AI总结 COMMA基准通过语言通信评估多模态多智能体系统的协作性能,揭示现有模型在智能体协作中的不足。

Journal ref Transactions on Machine Learning Research, 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2507.14417 2025-12-17 cs.AI cs.CL

Inverse Scaling in Test-Time Compute

测试时计算的反比例关系

Aryo Pradipta Gema, Alexander Hägele, Runjin Chen, Andy Arditi, Jacob Goldman-Wetzler, Kit Fraser-Taliente, Henry Sleight, Linda Petrini, Julian Michael, Beatrice Alex, Pasquale Minervini, Yanda Chen, Joe Benton, Ethan Perez

AI总结 研究发现,延长大型推理模型的推理时间会降低准确性,揭示了测试时计算与性能之间的反比例关系,并指出需通过多样化评估识别和解决推理中的失败模式。

Comments Published in TMLR (12/2025; Featured Certification; J2C Certification), 78 pages

Journal ref Transactions on Machine Learning Research (TMLR); 12/2025; https://openreview.net/forum?id=NXgyHW1c7M

详情

展开后加载摘要…

URL PDF HTML 收藏