arXivDaily arXiv每日学术速递 周一至周五更新

高校专区

The Chinese University of Hong Kong(香港中文大学)

2025-12-16 至 2025-12-16 共收录 11
2512.13070 2025-12-16 cs.AI cs.CL

M-GRPO: Stabilizing Self-Supervised Reinforcement Learning for Large Language Models with Momentum-Anchored Policy Optimization

M-GRPO:通过动量锚定策略优化稳定大语言模型的自监督强化学习

Bizhe Bai, Hongming Wu, Peng Ye, Tao Chen

机构 * Shanghai Innovation Institute(上海创新研究院) College of Future Information Technology, Fudan(复旦大学未来信息技术学院) Shanghai AI Laboratory(上海人工智能实验室) The Chinese University of Hong Kong(香港中文大学)

AI总结 M-GRPO通过动量锚定策略优化和IQR过滤方法,稳定大语言模型的自监督强化学习训练,提升训练稳定性和性能。

Comments 7 pages, 5 figures,Accepted NeurIPS 2025 Workshop on Efficient Reasoning

详情

展开后加载摘要…

URL PDF HTML 收藏
2512.12578 2025-12-16 quant-ph cs.CC cs.LG

Scalable Quantum Error Mitigation with Neighbor-Informed Learning

可扩展的邻域感知学习量子误差缓解

Zhenyu Chen, Bin Cheng, Minbo Gao, Xiaodie Lin, Ruiqi Zhang, Zhaohui Wei, Zhengfeng Ji

机构 * Department of Computer Science and Technology, Tsinghua University(清华大学计算机科学与技术系) Centre for Quantum Technologies, National University of Singapore(新加坡国立大学量子技术中心) Institute of Software, Chinese Academy of Sciences(中国科学院软件研究所) University of Chinese Academy of Sciences(中国科学院大学) Department of Mechanical and Automation Engineering, The Chinese University of Hong Kong(香港中文大学机械与自动化工程系) College of Computer and Data Science, Fuzhou University(福州大学计算机与数据科学学院) Yau Mathematical Sciences Center, Tsinghua University(清华大学姚期 Banking 数学科学中心) Department of Mathematics, Tsinghua University(清华大学数学系) Yanqi Lake Beijing Institute of Mathematical Sciences and Applications(燕琦湖北京数学科学与应用研究院)

AI总结 本文提出邻域感知学习框架,通过灵活高效的训练方法提升量子误差缓解的性能,实现理论与实际的双重优势。

详情

展开后加载摘要…

URL PDF HTML 收藏
2512.12552 2025-12-16 cs.AI

Large Language Newsvendor: Decision Biases and Cognitive Mechanisms

大语言新闻供应商:决策偏差与认知机制

Jifei Liu, Zhi Chen, Yuanguang Zhong

机构 * School of Business Administration, South China University of Technology(华南理工大学商学院) Department of Decisions, Operations and Technology, The Chinese University of Hong Kong(香港中文大学决策、运营与技术系)

AI总结 研究分析大语言模型在新闻供应商问题中的决策偏差,发现其放大人类认知偏差,提出通过结构化提示优化AI决策可靠性。

详情

展开后加载摘要…

URL PDF HTML 收藏
2512.08648 2025-12-16 cs.CV

Repulsor: Accelerating Generative Modeling with a Contrastive Memory Bank

Repulsor:利用对比记忆库加速生成建模

Shaofeng Zhang, Xuanqi Chen, Ning Liao, Haoxiang Zhao, Xiaoxing Wang, Haoru Tan, Sitong Wu, Xiaosong Jia, Qi Fan, Junchi Yan

机构 * School of Artificial Intelligence and Data Science, University of Science and Technology of China(人工智能与数据科学学院,中国科学技术大学) Shanghai Jiao Tong University(上海交通大学) HKU(香港大学) CUHK(香港大学) Fudan University(复旦大学) Nanjing University(南京大学)

AI总结 Repulsor通过对比记忆库机制,无需外部编码器,实现高效的生成建模,显著提升收敛速度和生成质量。

Comments 19 pages, 19 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.27562 2025-12-16 stat.ML cs.LG math.ST stat.TH

Optimal Convergence Analysis of DDPM for General Distributions

DDPM在一般分布下的最优收敛性分析

Yuchen Jiao, Yuchen Zhou, Gen Li

机构 * Department of Statistics and Data Science, Chinese University of Hong Kong(香港中文大学统计与数据科学系) Department of Statistics, University of Illinois Urbana-Champaign(伊利诺伊大学厄巴纳-香槟分校统计系)

AI总结 本文研究了DDPM在一般分布下的收敛性,通过引入放松光滑性条件,证明了其在KL散度中的收敛速率,并揭示了DDPM与DDIM在维度依赖上的相似性。

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.12793 2025-12-16 cs.CV

ViCO: A Training Strategy towards Semantic Aware Dynamic High-Resolution

ViCO: 一种面向语义感知动态高分辨率的训练策略

Long Cui, Weiyun Wang, Jie Shao, Zichen Wen, Gen Luo, Linfeng Zhang, Yanting Zhang, Yu Qiao, Wenhai Wang

机构 * Shanghai Jiao Tong University(上海交通大学) Shanghai Artificial Intelligence Laboratory(上海人工智能实验室) Fudan University(复旦大学) Nanjing University(南京大学) Donghua University(东华大学) The Chinese University of Hong Kong(香港中文大学)

AI总结 ViCO通过动态调整视觉标记数量以适应图像语义复杂度,有效降低推理成本的同时保持模型性能。

详情

展开后加载摘要…

URL PDF HTML 收藏
2505.14502 2025-12-16 cs.LG

Learning to Integrate Diffusion ODEs by Averaging the Derivatives

通过平均导数学习整合扩散ODEs

Wenze Liu, Xiangyu Yue

机构 * MMLab, The Chinese University of Hong Kong(中国香港大学)

AI总结 本文提出割线损失方法,通过学习ODE整合来提升扩散模型的推断速度和稳定性,实验表明其在CIFAR-10和ImageNet-256×256上取得良好效果。

Comments Accepted to NeurIPS 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2503.22140 2025-12-16 eess.IV cs.CV eess.SP

Score-Based Turbo Message Passing for Plug-and-Play Compressive Image Recovery

基于评分的turbo信息传递用于即插即用压缩图像恢复

Chang Cai, Xiaojun Yuan, Ying-Jun Angela Zhang

机构 * Department of Information Engineering, The Chinese University of Hong Kong, Shatin, N.T., Hong Kong SAR(信息工程系,香港中文大学(深圳校区)) National Key Lab. of Wireless Commun., Uni. of Electronic Sci. and Tech. of China, Chengdu, China(无线通信国家重点实验室,电子科技大学,成都,中国)

AI总结 本文提出了一种基于评分的turbo信息传递框架,利用评分建模与经验贝叶斯最优去噪的关系,提升压缩图像恢复的性能-复杂度权衡。

Comments IEEE SPAWC 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2502.13707 2025-12-16 cs.RO

Human-Like Robot Impedance Regulation Skill Learning from Human-Human Demonstrations

类人机器人阻抗调节技能从人与人示范中学习

Chenzui Li, Xi Wu, Yiming Chen, Tao Teng, Xuefeng Zhang, Sylvain Calinon, Darwin Caldwell, Fei Chen

机构 * Department of Mechanical and Automation Engineering, T-Stone Robotics Institute, The Chinese University of Hong Kong(机械与自动化工程系,T-Stone机器人研究院,香港中文大学) College of Science and Technology, Ningbo University(科学与技术学院,宁波大学) Idiap Research Institute(Idiap研究机构)

AI总结 本文提出HIImpRSL框架,通过模仿学习和LSTM模块实现机器人在多人协作任务中类人阻抗调节与适应能力。

Comments 13 pages, 9 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2411.14507 2025-12-16 cs.LG cs.AI cs.CL

From Pruning to Grafting: Dynamic Knowledge Redistribution via Learnable Layer Fusion

从剪枝到嫁接:通过可学习层融合实现动态知识再分配

Zehua Pei, Hui-Ling Zhen, Xianzhi Yu, Sinno Jialin Pan, Mingxuan Yuan, Bei Yu

机构 * The Chinese University of Hong Kong, Hong Kong SAR(香港中文大学) Noah’s Ark Lab, Huawei, Hong Kong SAR(华为诺亚实验室,香港)

AI总结 FuseGPT通过动态知识再分配和可学习层融合,实现更高效的模型压缩与性能提升。

详情

展开后加载摘要…

URL PDF HTML 收藏
2412.02421 2025-12-16 cs.CV

TimeWalker: Personalized Neural Space for Lifelong Head Avatars

TimeWalker: 为终身头像的个性化神经空间

Dongwei Pan, Yang Li, Hongsheng Li, Kwan-Yee Lin

机构 * Shanghai AI Laboratory(上海人工智能实验室) CUHK(香港中文大学)

AI总结 TimeWalker通过个性化神经空间实现终身头像的重建与动画,结合动态神经基融合模块和2D高斯点散射技术,实现真实感的面部表情和年龄变化建模。

Comments Project Page: https://timewalker2025.github.io/timewalker.github.io/

详情

展开后加载摘要…

URL PDF HTML 收藏