arXivDaily arXiv每日学术速递 周一至周五更新
arXiv周末暂无论文更新,休息一下吧,周末愉快~~

期刊&会议

International Conference on Machine Learning · 会议 · Machine Learning

2026-02-13 至 2026-02-13 共收录 3
2602.11169 2026-02-13 cs.CL cs.AI cs.LG

Disentangling Direction and Magnitude in Transformer Representations: A Double Dissociation Through L2-Matched Perturbation Analysis

解构Transformer表示中的方向与幅度:通过L2匹配扰动分析的双重解离

Mangadoddi Srikar Vardhan, Lekkala Sai Teja

AI总结 研究揭示Transformer表示中方向与幅度在语言建模和语法处理中的不同作用,通过L2匹配扰动分析发现方向扰动影响注意力路径,幅度扰动影响语法判断,且解离依赖于架构选择。

Comments 15 pages, 7 figures. will Submit to ICML 2026

详情

展开后加载摘要…

URL PDF HTML 收藏
2502.14921 2026-02-13 cs.CL cs.CR cs.LG

The Canary's Echo: Auditing Privacy Risks of LLM-Generated Synthetic Text

知更鸟的回声:审计大语言模型生成合成文本的隐私风险

Matthieu Meeus, Lukas Wutschitz, Santiago Zanella-Béguelin, Shruti Tople, Reza Shokri

机构 * Imperial College London(帝国理工学院伦敦分校) Microsoft(微软公司) National University of Singapore(新加坡国立大学)

AI总结 本文提出通过设计具有分布内前缀和高困惑度后缀的知更鸟,提高基于数据的MIAs的威力,以更准确评估LLM生成合成数据的隐私风险。

Comments 42nd International Conference on Machine Learning (ICML 2025)

Journal ref Proc. Mach. Learn. Res. 267 (2025) 43557-43580

详情

展开后加载摘要…

URL PDF HTML 收藏
2501.02087 2026-02-13 cs.LG stat.ML

Beyond CVaR: Leveraging Static Spectral Risk Measures for Enhanced Decision-Making in Distributional Reinforcement Learning

超越CVaR:利用静态谱风险度量提升分布式强化学习中的决策质量

Mehrdad Moghimi, Hyejin Ku

机构 * Department of Mathematics and Statistics, York University, Toronto, Canada(数学与统计学系,约克大学,多伦多,加拿大)

AI总结 本文提出了一种具有收敛保证的DRL算法,优化更广泛的静态谱风险度量,提升决策质量并优于现有方法。

Comments Accepted at ICML 2025

Journal ref Proceedings of the 42nd International Conference on Machine Learning, PMLR 267:44571-44593, 2025

详情

展开后加载摘要…

URL PDF HTML 收藏