arXivDaily arXiv每日学术速递 周一至周五更新
arXiv周末暂无论文更新,休息一下吧,周末愉快~~

高校专区

University of California, Los Angeles(加州大学洛杉矶分校)

2025-10-01 至 2025-10-01 共收录 5
2507.15855 2025-10-01 cs.AI

Winning Gold at IMO 2025 with a Model-Agnostic Verification-and-Refinement Pipeline

Yichen Huang, Lin F. Yang

机构 * Department of Electrical and Computer Engineering, and Department of Computer Science, UCLA(电气与计算机工程系和计算机科学系, UCLA)

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.26114 2025-10-01 cs.LG

Clip-Low Increases Entropy and Clip-High Decreases Entropy in Reinforcement Learning of Large Language Models

Jaesung R. Park, Junsu Kim, Gyeongman Kim, Jinyoung Jo, Sean Choi, Jaewoong Cho, Ernest K. Ryu

机构 * Department of Mathematics, UCLA(UCLA数学系) Department of Mathematical Sciences, Seoul National University(首尔国立大学数学科学系) KRAFTON Department of Linguistics, Stanford University(斯坦福大学语言学系) Department of Computer Science and Engineering, Santa Clara University(圣克拉拉大学计算机科学与工程系)

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.25792 2025-10-01 cs.AI cs.CV

PUREVQ-GAN: Defending Data Poisoning Attacks through Vector-Quantized Bottlenecks

Alexander Branch, Omead Pooladzandi, Radin Khosraviani, Sunay Gajanan Bhat, Jeffrey Jiang, Gregory Pottie

机构 * University of California, Los Angeles(加州大学洛杉矶分校) California Institute of Technology(加州理工学院)

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.25598 2025-10-01 cs.AI cs.LG

Hybrid Reward Normalization for Process-supervised Non-verifiable Agentic Tasks

Peiran Xu, Zhuohao Li, Xiaoying Xing, Guannan Zhang, Debiao Li, Kunyu Shi

机构 * Accio Team, Alibaba Group(阿里集团阿西莫团队) University of California, Los Angeles (UCLA)(加州大学洛杉矶分校)

详情

展开后加载摘要…

URL PDF HTML 收藏
2502.05849 2025-10-01 cs.CL

Where Fact Ends and Fairness Begins: Redefining AI Bias Evaluation through Cognitive Biases

Jen-tse Huang, Yuhang Yan, Linqi Liu, Yixin Wan, Wenxuan Wang, Kai-Wei Chang, Michael R. Lyu

机构 * The Chinese University of Hong Kong(香港中文大学) University of California, Los Angeles(加州大学洛杉矶分校)

Comments Accepted to EMNLP 2025 (Findings)

详情

展开后加载摘要…

URL PDF HTML 收藏