arXivDaily arXiv每日学术速递 周一至周五更新

高校专区

The Chinese University of Hong Kong(香港中文大学)

2025-12-30 至 2025-12-30 共收录 6
2512.23705 2025-12-30 cs.CV

Diffusion Knows Transparency: Repurposing Video Diffusion for Transparent Object Depth and Normal Estimation

扩散知道透明:将视频扩散用于透明物体深度和法线估计

Shaocong Xu, Songlin Wei, Qizhe Wei, Zheng Geng, Hong Li, Licheng Shen, Qianpu Sun, Shu Han, Bin Ma, Bohan Li, Chongjie Ye, Yuhang Zheng, Nan Wang, Saining Zhang, Hao Zhao

机构 * Beijing Academy of Artificial Intelligence(北京人工智能研究院) University of Southern California(南加州大学) Tsinghua University(清华大学) Beihang University(北航) Wuhan University(武汉大学) Shanghai Jiao Tong University(上海交通大学) European Institute of Innovation and Technology Ningbo(创新与技术欧洲研究所宁波) FNii, The Chinese University of Hong Kong, Shenzhen(FNii,香港中文大学(深圳)) National University of Singapore(新加坡国立大学)

AI总结 本文提出DKT模型,利用视频扩散模型估计透明物体的深度和法线,实现零样本SOTA,提升现实和合成视频中的透明感知性能。

Comments Project Page: https://daniellli.github.io/projects/DKT/; Code: https://github.com/Daniellli/DKT; Dataset: https://huggingface.co/datasets/Daniellesry/TransPhy3D

详情

展开后加载摘要…

URL PDF HTML 收藏
2512.17703 2025-12-30 cond-mat.str-el cond-mat.mtrl-sci cs.LG physics.comp-ph

Revisiting the Broken Symmetry Phase of Solid Hydrogen: A Neural Network Variational Monte Carlo Study

重新审视固态氢的破缺对称相:一种神经网络变分蒙特卡洛研究

Shengdu Chai, Chen Lin, Xinyang Dong, Yuqiang Li, Wanli Ouyang, Lei Wang, X. C. Xie

机构 * Interdisciplinary Center for Theoretical Physics(理论物理交叉中心) Information Sciences (ICTPIS), Fudan University(信息科学(ICTPIS),复旦大学) Shanghai Artificial Intelligence Laboratory(上海人工智能实验室) Department of Engineering, University of Oxford(工程系,牛津大学) Beijing National Laboratory for Condensed Matter Physics(北京凝聚态物理实验室) Institute of Physics, Chinese Academy of Sciences(物理研究所,中国科学院) Department of Information Engineering, The Chinese University of Hong Kong(信息工程系,香港中文大学) International Center for Quantum Materials, School of Physics, Peking University(国际量子材料中心,物理系,北京大学) Hefei National Laboratory(合肥国家实验室)

AI总结 本文通过神经网络变分蒙特卡洛方法研究固态氢的破缺对称相,发现其基态结构候选者Cmcm空间群对称性,并验证其稳定性及与实验数据的一致性。

详情

展开后加载摘要…

URL PDF HTML 收藏
2512.22452 2025-12-30 cs.CV

SAM 3D for 3D Object Reconstruction from Remote Sensing Images

SAM 3D用于从遥感图像中进行3D物体重建

Junsheng Yao, Lichao Mou, Qingyu Li

机构 * The Chinese University of Hong Kong, Shenzhen(香港中文大学(深圳)) MedAI Technology(MedAI技术)

AI总结 本文提出SAM 3D模型用于遥感图像的3D建筑重建,通过实验验证其在生成几何和边界上的优势,并扩展至城市场景重建,探讨其在城市建模中的应用潜力。

详情

展开后加载摘要…

URL PDF HTML 收藏
2512.22315 2025-12-30 cs.CV cs.AI

VideoZoomer: Reinforcement-Learned Temporal Focusing for Long Video Reasoning

VideoZoomer: 用于长视频推理的强化学习时序聚焦

Yang Ding, Yizhen Zhang, Xin Lai, Ruihang Chu, Yujiu Yang

机构 * Tsinghua University(清华大学) The Chinese University of Hong Kong(香港中文大学)

AI总结 VideoZoomer通过强化学习实现动态时序聚焦,提升长视频推理性能和效率。

详情

展开后加载摘要…

URL PDF HTML 收藏
2512.22192 2025-12-30 cs.LG

Frequency Regularization: Unveiling the Spectral Inductive Bias of Deep Neural Networks

频域正则化:揭示深度神经网络的频域归纳偏置

Jiahao Lu

机构 * School of Artificial Intelligence(人工智能学院) The Chinese University of Hong Kong, Shenzhen(香港中文大学(深圳))

AI总结 本研究通过频域正则化揭示深度神经网络对低频结构的归纳偏置,展示L2正则化在抑制高频能量和提升鲁棒性方面的优势。

Comments 9 pages, 5 figures. Code available at https://github.com/lujiahao760/FrequencyRegularization

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.12974 2025-12-30 cs.SD eess.AS

The CCF AATC 2025 Speech Restoration Challenge: A Retrospective

2025年CCF AATC语音恢复挑战:回顾

Junan Zhang, Mengyao Zhu, Xin Xu, Hui Bu, Zhenhua Ling, Zhizheng Wu

机构 * The Chinese University of Hong Kong, Shenzhen(香港中文大学(深圳)) Audio Department, Huawei CBG(华为CBG音频部门) Beijing AISHELL Technology Co., Ltd.(北京艾斯HELL科技有限公司) University of Science and Technology of China(中国科学技术大学)

AI总结 2025年CCF AATC挑战回顾了语音恢复领域的现状,揭示了轻量模型、生成模型权衡及度量差距问题。

Comments Technical Report. Homepage: https://ccf-aatc.org.cn. Code & Data: https://github.com/viewfinder-annn/anyenhance-v1-ccf-aatc

详情

展开后加载摘要…

URL PDF HTML 收藏