arXivDaily arXiv每日学术速递 周一至周五更新

高校专区

University of Southern California(南加州大学)

2025-12-15 至 2025-12-15 共收录 3
2510.22980 2025-12-15 cs.LG stat.ML

How Muon's Spectral Design Benefits Generalization: A Study on Imbalanced Data

muon的谱设计如何促进泛化:对不平衡数据的研究

Bhavya Vasudeva, Puneesh Deora, Yize Zhao, Vatsal Sharan, Christos Thrampoulidis

机构 * University of Southern California(南加州大学) University of British Columbia(不列颠哥伦比亚大学)

AI总结 muon的谱设计通过促进数据潜在成分的平衡学习,提升了模型的泛化能力,尤其在不平衡数据中表现更优。

Comments 36 pages, 32 figures, 1 table

详情

展开后加载摘要…

URL PDF HTML 收藏
2502.11028 2025-12-15 cs.CL cs.AI

Mind the Confidence Gap: Overconfidence, Calibration, and Distractor Effects in Large Language Models

注意置信差距:大型语言模型中的过度自信、校准与干扰效应

Prateek Chhikara

机构 * University of Southern California(美国南加州大学)

AI总结 本研究探讨了大型语言模型中的过度自信问题,通过引入干扰项显著改善校准,提出针对性的改进策略以提升模型可靠性。

Comments Published in Transactions on Machine Learning Research (TMLR)

详情

展开后加载摘要…

URL PDF HTML 收藏
2410.14040 2025-12-15 cs.LG cs.AI cs.RO

WARPD: World model Assisted Reactive Policy Diffusion

WARPD:世界模型辅助的反应策略扩散

Shashank Hegde, Satyajeet Das, Gautam Salhotra, Gaurav S. Sukhatme

机构 * University of Southern California(南加州大学) Google(谷歌) Intrinsic LLC(Intrinsic 公司)

AI总结 WARPD通过直接生成闭环策略,提高了机器人任务中长动作时间跨度和鲁棒性的性能,同时显著降低了推理成本。

Comments Outstanding Paper Award at the Embodied World Models for Decision Making Workshop at NeurIPS 2025

详情

展开后加载摘要…

URL PDF HTML 收藏