arXivDaily arXiv每日学术速递 周一至周五更新

高校专区

The Hong Kong University of Science and Technology(香港科技大学)

2026-02-03 至 2026-02-03 共收录 26
2602.02159 2026-02-03 cs.CL

Focus-dLLM: Accelerating Long-Context Diffusion LLM Inference via Confidence-Guided Context Focusing

Focus-dLLM: 通过置信度引导的上下文聚焦加速长上下文扩散语言模型推理

Lingkun Long, Yushi Huang, Shihao Bai, Ruihao Gong, Jun Zhang, Ao Zhou, Jianlei Yang

机构 * Beihang University(北航) Hong Kong University of Science and Technology(香港理工大学) SenseTime Research(商汤科技研究院)

AI总结 Focus-dLLM通过置信度引导的上下文聚焦技术,实现了长上下文扩散语言模型推理的高效加速。

详情

展开后加载摘要…

URL PDF HTML 收藏
2602.02156 2026-02-03 cs.CV

LoopViT: Scaling Visual ARC with Looped Transformers

LoopViT: 通过循环变换器扩展视觉ARC

Wen-Jie Shu, Xuerui Qiu, Rui-Jie Zhu, Harold Haodong Chen, Yexin Liu, Harry Yang

机构 * HKUST(香港科技大学) CASIA(中国科学院自动化研究所) UC Santa Cruz(加州大学圣克ruz分校)

AI总结 LoopViT通过循环变换器实现视觉推理的高效扩展,采用权重绑定递归结构和动态退出机制,提升ARC-AGI基准测试性能。

Comments 8 pages, 11 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2602.01965 2026-02-03 cs.CL cs.AI

Breaking the Static Graph: Context-Aware Traversal for Robust Retrieval-Augmented Generation

打破静态图:面向鲁棒检索增强生成的上下文感知遍历

Kwun Hang Lau, Fangyuan Zhang, Boyu Ruan, Yingli Zhou, Qintian Guo, Ruiyuan Zhang, Xiaofang Zhou

机构 * Huawei Hong Kong Research Center(华为香港研究中心) The Hong Kong University of Science and Technology(香港科学与技术大学) The Chinese University of Hong Kong, Shenzhen(香港中文大学(深圳))

AI总结 CatRAG通过上下文感知遍历框架,改进检索增强生成模型,提升多跳查询的推理完整性和证据链恢复能力。

详情

展开后加载摘要…

URL PDF HTML 收藏
2602.01877 2026-02-03 cs.LG math.OC

Autocorrelated Optimize-via-Estimate: Predict-then-Optimize versus Finite-sample Optimal

自相关优化-通过估计:预测-然后优化与有限样本最优

Zichun Wang, Gar Goei Loke, Ruiting Zuo

机构 * The Hong Kong University of Science and Technology (Guangzhou)(香港科技大学(广州)) Durham University Business School(杜伦大学商学院)

AI总结 本文提出了一种自相关优化-通过估计模型,用于在有限样本情况下优化资产组合,表现出优于传统方法的性能。

详情

展开后加载摘要…

URL PDF HTML 收藏
2602.01825 2026-02-03 stat.ME cs.LG math.OC stat.ML

Learning Sequential Decisions from Multiple Sources via Group-Robust Markov Decision Processes

通过群鲁棒马尔可夫决策过程从多个来源学习顺序决策

Mingyuan Xu, Zongqi Xia, Tianxi Cai, Doudou Zhou, Nian Si

机构 * Department of Statistics and Data Science, National University of Singapore(新加坡国立大学统计与数据科学系) Department of Neurology, University of Pittsburgh(匹兹堡大学神经病学系) Department of Biostatistics, Harvard T.H. Chan School of Public Health(哈佛大学T.H. Chan公共卫生学院生物统计学系) Department of Biomedical Informatics, Harvard Medical School(哈佛医学院生物医学信息学系) Department of Industrial Engineering and Decision Analytics, Hong Kong University of Science and Technology(香港科学与技术大学工业工程与决策分析系)

AI总结 本文提出了一种群鲁棒马尔可夫决策过程框架,通过特征层面的不确定性集和悲观价值迭代算法,从多地点异质数据中学习鲁棒的顺序决策策略。

详情

展开后加载摘要…

URL PDF HTML 收藏
2512.10942 2026-02-03 cs.CV

VL-JEPA: Joint Embedding Predictive Architecture for Vision-language

VL-JEPA:面向视觉-语言的联合嵌入预测架构

Delong Chen, Mustafa Shukor, Theo Moutakanni, Willy Chung, Jade Yu, Tejaswi Kasarla, Yejin Bang, Allen Bolourchi, Yann LeCun, Pascale Fung

机构 * Meta FAIR HKUST(香港科技大学) Sorbonne Université(索邦大学) NYU(纽约大学)

AI总结 VL-JEPA通过联合嵌入预测架构,在减少参数的情况下实现更强的视觉-语言性能,支持多种任务且无需架构修改。

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.13745 2026-02-03 cs.CV

UniCalli: A Unified Diffusion Framework for Column-Level Generation and Recognition of Chinese Calligraphy

UniCalli: 一种统一的扩散框架,用于中文书法的列级生成与识别

Tianshuo Xu, Kai Wang, Zhifei Chen, Leyi Wu, Tianshui Wen, Fei Chao, Ying-Cong Chen

机构 * HKUST(GZ)(香港科技大学(广州)) China University of Geoscience Beijing(中国地质大学(北京)) Xiamen University(厦门大学) HKUST(香港科技大学)

AI总结 UniCalli提出一种统一的扩散框架,通过联合训练实现中文书法列级生成与识别,提升生成质量和识别性能,同时扩展至其他古代文字。

Comments Page: https://envision-research.github.io/UniCalli/

详情

展开后加载摘要…

URL PDF HTML 收藏
2507.12094 2026-02-03 cs.LG cs.GT

Is This Predictor More Informative than Another? A Decision-Theoretical Comparison

这个预测器比另一个更有信息量吗?一种决策理论的比较

Yiding Feng, Liuhan Qian, Wei Tang

机构 * Hong Kong University of Science and Technology(香港科学与技术大学) The Chinese University of Hong Kong(香港中文大学)

AI总结 本文提出信息量差距的概念,用于比较预测器的决策相关性,提供了一种评估预测模型在不同决策任务中表现的理论框架和实验验证。

详情

展开后加载摘要…

URL PDF HTML 收藏
2505.18561 2026-02-03 cs.CV

CoT-RVS: Zero-Shot Chain-of-Thought Reasoning Segmentation for Videos

CoT-RVS:零样本链式推理视频对象分割

Shiu-hong Kao, Yu-Wing Tai, Chi-Keung Tang

机构 * The Hong Kong University of Science and Technology(香港科学与技术大学) National University of Singapore(新加坡国立大学) Dartmouth College(达特茅斯学院)

AI总结 CoT-RVS通过零样本链式推理能力,实现了对视频对象的高效分割,无需训练即可处理复杂查询和在线视频流。

Comments Accepted to ICLR 2026. Project page: https://danielshkao.github.io/cot-rvs.html. Code: https://github.com/DanielSHKao/CoT-RVS

详情

展开后加载摘要…

URL PDF HTML 收藏
2505.15386 2026-02-03 cs.CL cs.AI

RePPL: Recalibrating Perplexity by Uncertainty in Semantic Propagation and Language Generation for Explainable QA Hallucination Detection

RePPL:通过语义传播和语言生成中的不确定性重新校准困惑度以检测可解释问答幻觉

Yiming Huang, Junyan Zhang, Zihao Wang, Biquan Bie, Yunzhong Qiu, Xuming Hu, Yi R. Fung, Xinlei He

机构 * The Hong Kong University of Science and Technology(香港科学与技术大学) Tsinghua University(清华大学)

AI总结 RePPL通过校准语义传播和语言生成中的不确定性,提升问答幻觉检测性能,生成token级不确定性评分作为解释。

详情

展开后加载摘要…

URL PDF HTML 收藏
2602.01156 2026-02-03 cs.LG cs.RO

PolicyFlow: Policy Optimization with Continuous Normalizing Flow in Reinforcement Learning

PolicyFlow: 在强化学习中使用连续归一化流进行策略优化

Shunpeng Yang, Ben Liu, Hua Chen

机构 * Hong Kong University of Science and Technology(香港科技大学) Southern University of Science and Technology(南方科技大学) Zhejiang University-University of Illinois Urbana-Champaign Institute(浙江大学-伊利诺伊大学厄巴纳-香槟分校联合研究所) LimX Dynamics

AI总结 PolicyFlow是一种基于连续归一化流的在线强化学习算法,通过减少似然计算开销实现稳定训练,并通过布朗正则化器促进多样化行为。

Comments Submitted to ICLR 2026

详情

展开后加载摘要…

URL PDF HTML 收藏
2602.00969 2026-02-03 cs.LG

On the Spectral Flattening of Quantized Embeddings

量化嵌入谱扁平化研究

Junlin Huang, Wenyi Fang, Zhenheng Tang, Yuxin Wang, Xueze Kang, Yang Zheng, Bo Li, Xiaowen Chu

机构 * The Hong Kong University of Science and Technology (Guangzhou)(香港科学与技术大学(广州)) Huawei Technologies Co., Ltd(华为技术有限公司)

AI总结 本研究揭示了量化嵌入谱扁平化现象,证明了谱保真度对稳定低比特优化的必要性。

详情

展开后加载摘要…

URL PDF HTML 收藏
2602.00808 2026-02-03 cs.RO

Physics-informed Diffusion Mamba Transformer for Real-world Driving

物理引导的扩散Mamba变换器用于现实驾驶

Hang Zhou, Qiang Zhang, Peiran Liu, Yihao Qin, Zhaoxu Yan, Yiding Ji

机构 * Robotics and Autonomous Systems Thrust, The Hong Kong University of Science and Technology(Guangzhou)(香港科学与技术大学(广州)机器人与自主系统方向) MoSense Technologies

AI总结 本文提出了一种结合Mamba和注意力机制的扩散模型,通过整合物理约束提升自动驾驶轨迹预测的准确性和可解释性。

Journal ref International Conference on Robotics & Automation (ICRA) 2026

详情

展开后加载摘要…

URL PDF HTML 收藏
2602.00739 2026-02-03 cs.CV

Diffusion-Driven Inter-Outer Surface Separation for Point Clouds with Open Boundaries

基于扩散的点云双层表面分离:用于开放边界

Zhengyan Qin, Liyuan Qiu

机构 * Hong Kong University of Science and Technology (HKUST)(香港理工大学) Hong Kong Applied Science and Technology Research Institute (ASTRI)(香港应用科技研究院)

AI总结 本文提出了一种基于扩散的算法,用于分离双层点云的内层和外层表面,特别针对具有开放边界的点云,通过提取真实内层来解决重叠表面和法线紊乱问题。

详情

展开后加载摘要…

URL PDF HTML 收藏
2602.00726 2026-02-03 cs.HC cs.AI

Augmenting Clinical Decision-Making with an Interactive and Interpretable AI Copilot: A Real-World User Study with Clinicians in Nephrology and Obstetrics

通过交互式和可解释的AI助手增强临床决策:与泌尿科和产科医生的现实世界用户研究

Yinghao Zhu, Dehao Sui, Zixiang Wang, Xuning Hu, Lei Gu, Yifan Qi, Tianchen Wu, Ling Wang, Yuan Wei, Wen Tang, Zhihan Cui, Yasha Wang, Lequan Yu, Ewen M Harrison, Junyi Gao, Liantao Ma

机构 * Peking University(北京大学) University of Hong Kong(香港大学) Hong Kong University of Science and Technology(香港科学与技术大学) Peking University Third Hospital(北京大学第三医院) Affiliated Xuzhou Municipal Hospital of Xuzhou Medical University(徐州医科大学附属徐州市人民医院) University of Edinburgh(爱丁堡大学) Health Data Research UK(英国健康数据研究)

AI总结 AICare通过交互式和可解释的AI助手提升临床决策,通过实验证明其降低认知负荷并增强医生信任

Comments Accepted by ACM CHI 2026

详情

展开后加载摘要…

URL PDF HTML 收藏
2602.00686 2026-02-03 cs.RO

Learning to Accelerate Vision-Language-Action Models through Adaptive Visual Token Caching

通过自适应视觉令牌缓存学习加速视觉-语言-动作模型

Yujie Wei, Jiahan Fan, Jiyu Guo, Ruichen Zhen, Rui Shao, Xiu Su, Zeke Xie, Shuo Yang

机构 * Harbin Institute of Technology(哈尔滨工业大学) Harbin Institute of Technology, Shenzhen(哈尔滨工业大学(深圳)) Meituan Academy of Robotics Shenzhen, Meituan(美团机器人深圳研究院) Central South University(中南大学) HKUST(GZ)(香港科技大学(广州))

AI总结 本文提出通过自适应视觉令牌缓存学习加速VLA模型,提升推理效率并提高任务成功率。

详情

展开后加载摘要…

URL PDF HTML 收藏
2602.00528 2026-02-03 cs.AI

How Far Are LLMs from Professional Poker Players? Revisiting Game-Theoretic Reasoning with Agentic Tool Use

LLMs距离专业扑克玩家还有多远?结合代理工具使用的博弈论推理再探

Minhua Lin, Enyan Dai, Hui Liu, Xianfeng Tang, Yuliang Yan, Zhenwei Dai, Jingying Zeng, Zhiwei Zhang, Fali Wang, Hongcheng Gao, Chen Luo, Xiang Zhang, Qi He, Suhang Wang

机构 * The Pennsylvania State University(宾夕法尼亚州立大学) HKUST (GZ)(香港科技大学) Amazon(亚马逊) Tsinghua University(清华大学) Microsoft(微软公司)

AI总结 本文提出ToolPoker框架,通过整合外部求解器和专业解释,提升LLMs在扑克博弈中的推理和游戏表现。

Comments Accepted by ICLR 2026

详情

展开后加载摘要…

URL PDF HTML 收藏
2601.20327 2026-02-03 cs.CL

CE-RM: A Pointwise Generative Reward Model Optimized via Two-Stage Rollout and Unified Criteria

CE-RM:一种通过两阶段回放和统一标准优化的点wise生成奖励模型

Xinyu Hu, Yancheng He, Weixun Wang, Tao Feng, Li Lin, Jiashun Liu, Wenbo Su, Bo Zheng, Xiaojun Wan

机构 * Wangxuan Institute of Computer Technology, Peking University(北京大学计算机技术研究院) Hong Kong University of Science and Technology(香港理工大学) Alibaba Group(阿里巴巴集团)

AI总结 CE-RM通过两阶段回放和统一标准优化,提升生成奖励模型在开放性自然语言生成和强化学习中的表现。

Comments Under Review

详情

展开后加载摘要…

URL PDF HTML 收藏
2511.15248 2026-02-03 cs.LG cs.AI

EntroPIC: Towards Stable Long-Term Training of LLMs via Entropy Stabilization with Proportional-Integral Control

EntroPIC: 通过比例-积分控制实现LLM长期训练的稳定性

Kai Yang, Xin Xu, Yangkun Chen, Weijie Liu, Jiafei Lyu, Zichuan Lin, Deheng Ye, Saiyong Yang

机构 * Tencent Hunyuan(腾讯文言) The Hong Kong University of Science and Technology(香港科学与技术大学)

AI总结 EntroPIC通过比例-积分控制实现LLM长期训练的熵稳定,提升探索效率和训练稳定性。

详情

展开后加载摘要…

URL PDF HTML 收藏
2507.16559 2026-02-03 cs.CV

Comparative validation of surgical phase recognition, instrument keypoint estimation, and instrument instance segmentation in endoscopy: Results of the PhaKIR 2024 challenge

内镜手术阶段识别、器械关键点估计和器械实例分割的比较验证:PhaKIR 2024挑战赛结果

Tobias Rueckert, David Rauber, Raphaela Maerkl, Leonard Klausmann, Suemeyye R. Yildiran, Max Gutbrod, Danilo Weber Nunes, Alvaro Fernandez Moreno, Imanol Luengo, Danail Stoyanov, Nicolas Toussaint, Enki Cho, Hyeon Bae Kim, Oh Sung Choo, Ka Young Kim, Seong Tae Kim, Gonçalo Arantes, Kehan Song, Jianjun Zhu, Junchen Xiong, Tingyi Lin, Shunsuke Kikuchi, Hiroki Matsuzaki, Atsushi Kouno, João Renato Ribeiro Manesco, João Paulo Papa, Tae-Min Choi, Tae Kyeong Jeong, Juyoun Park, Oluwatosin Alabi, Meng Wei, Tom Vercauteren, Runzhi Wu, Mengya Xu, An Wang, Long Bai, Hongliang Ren, Amine Yamlahi, Jakob Hennighausen, Lena Maier-Hein, Satoshi Kondo, Satoshi Kasai, Kousuke Hirasawa, Shu Yang, Yihui Wang, Hao Chen, Santiago Rodríguez, Nicolás Aparicio, Leonardo Manrique, Juan Camilo Lyons, Olivia Hosie, Nicolás Ayobi, Pablo Arbeláez, Yiping Li, Yasmina Al Khalil, Sahar Nasirihaghighi, Stefanie Speidel, Daniel Rueckert, Hubertus Feussner, Dirk Wilhelm, Christoph Palm

机构 * Regensburg Medical Image Computing (ReMIC), OTH Regensburg(雷根萨大学医学影像计算中心) Research Group MITI, TUM University Hospital, School of Medicine and Health(技术大学慕尼黑大学医院MITI研究组) Regensburg Center of Biomedical Engineering (RCBE), OTH Regensburg(雷根萨生物医学工程研究中心) Regensburg University(雷根萨大学) Regensburg Center of Health Sciences and Technology (RCHST), OTH Regensburg(雷根萨健康科学与技术研究中心) AI Centre of Excellence, Medtronic Ltd.(医学影像人工智能卓越中心) Engineering Sciences, University College London(伦敦大学学院工程科学系) Augmented Intelligence Lab, Kyung Hee University(庆熙大学增强智能实验室) University of Minho, Braga(明霍大学) Jmees Inc.(Jmees公司) School of Sciences, São Paulo State University (UNESP), Bauru(圣保罗州立大学科学学院) KIST HARILAB, Center for Humanoid Research, Artificial Intelligence and Robot Institute, Korea Institute of Science and Technology (KIST)(韩国科学技术院HARILAB中心) King's College London(伦敦国王学院) The Chinese University of Hong Kong(香港中文大学) Division of Intelligent Medical Systems, German Cancer Research Center (DKFZ)(德国癌症研究中心智能医疗系统部门) Muroran Institute of Technology, Hokkaido(北海道Muroran技术学院) Niigata University of Health and Welfare(Niigata医疗福利大学) Konica Minolta, Inc.(东宝株式会社) Department of Computer Science and Engineering, The Hong Kong University of Science and Technology(香港科技大学计算机科学与工程系) Department of Chemical and Biological Engineering, The Hong Kong University of Science and Technology(香港科技大学化学与生物工程系) HKUST Shenzhen-Hong Kong Collaborative Innovation Research Institute, Shenzhen(香港科技大学深圳-香港协同创新研究院) Center for Research and Formation in Artificial Intelligence (CinfonIA), Los Andes University, Bogota(安第斯大学人工智能研究与培养中心) Department of Biomedical Engineering, Medical Image Analysis Group, Eindhoven University of Technology(埃因霍温理工大学生物医学工程系) Institute of Information Technology (ITEC), Klagenfurt University(克雷格夫大学信息技术研究所) Center for Tactile Internet with Human-in-the-loop (CeTI), TU Dresden(德累斯顿技术大学触觉互联网中心)

AI总结 PhaKIR 2024挑战赛通过多中心数据集验证内镜手术阶段识别、关键点估计和实例分割的性能,推动RAMIS领域的时间感知和情境驱动方法发展。

Comments A challenge report pre-print accepted by the journal Medical Image Analysis (MedIA), containing 37 pages, 15 figures, and 14 tables

详情

展开后加载摘要…

URL PDF HTML 收藏
2505.24061 2026-02-03 cs.LG

Measure gradients, not activations! Enhancing neuronal activity in deep reinforcement learning

测量梯度,而非激活!增强深度强化学习中的神经元活动

Jiashun Liu, Zihao Wu, Johan Obando-Ceron, Pablo Samuel Castro, Aaron Courville, Ling Pan

机构 * Hong Kong University of Science and Technology(香港科技大学) Mila - Québec AI Institute(魁北克AI研究所) Université de Montréal(蒙特利尔大学)

AI总结 本文提出GraMa用于衡量深度强化学习中神经元的学习能力,通过梯度而非激活来增强神经元活动,提升学习性能。

详情

展开后加载摘要…

URL PDF HTML 收藏
2505.16530 2026-02-03 cs.CR cs.AI cs.CL

DuFFin: A Dual-Level Fingerprinting Framework for LLMs IP Protection

DuFFin:一种用于LLM知识产权保护的双层指纹框架

Yuliang Yan, Haochun Tang, Shuo Yan, Enyan Dai

机构 * The Hong Kong University of Science and Technology (Guangzhou)(香港科学与技术大学(广州)) Jilin University(吉林大学)

AI总结 DuFFin通过双层指纹框架在黑盒环境下实现LLM知识产权验证,准确识别模型来源并达到高验证精度。

Comments Accepted by EACL 2026, Findings

详情

展开后加载摘要…

URL PDF HTML 收藏
2502.06914 2026-02-03 q-bio.QM cs.AI cs.LG

UniZyme: A Unified Protein Cleavage Site Predictor Enhanced with Enzyme Active-Site Knowledge

UniZyme:一种结合酶活性位点知识的统一蛋白质裂解位点预测器

Chenao Li, Shuo Yan, Enyan Dai

机构 * Hong Kong University of Science and Technology (Guangzhou)(香港科技大学(广州))

AI总结 UniZyme通过结合活性位点知识的统一模型,实现了对多种酶裂解位点的高精度预测。

Comments 22 pages,9 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2501.13448 2026-02-03 cs.MA cs.AI cs.ET cs.LG

BMG-Q: Localized Bipartite Match Graph Attention Q-Learning for Ride-Pooling Order Dispatch

BMG-Q:局部双图匹配图注意力Q学习用于拼车订单调度

Yulong Hu, Siyuan Feng, Sen Li

机构 * Department of Civil and Environmental Engineering, The Hong Kong University of Science and Technology(香港科学与技术大学土木与环境工程系) Intelligent Transportation Thrust, Systems Hub, The Hong Kong University of Science and Technology (Guangzhou)(香港科学与技术大学(广州)智能交通 thrust,系统中心) Department of Aeronautical and Aviation Engineering, The Hong Kong Polytechnic University(香港理工大学航空与航空工程系)

AI总结 BMG-Q通过局部双图匹配图注意力Q学习提升拼车订单调度的决策效率与鲁棒性。

Journal ref IEEE Transactions on Intelligent Transportation Systems ( Volume: 26, Issue: 10, October 2025)

详情

展开后加载摘要…

URL PDF HTML 收藏
2407.19420 2026-02-03 cs.LG

UniGAP: A Universal and Adaptive Graph Upsampling Approach to Mitigate Over-Smoothing in Node Classification Tasks

UniGAP: 一种通用且自适应的图上采样方法以缓解节点分类任务中的过平滑问题

Xiaotang Wang, Yun Zhu, Haizhou Shi, Yongchao Liu, Yongqi Zhang

机构 * The Hong Kong University of Science and Technology (Guangzhou)(香港科学与技术大学(广州)) Shanghai Artificial Intelligence Laboratory(上海人工智能实验室) Rutgers University(罗格斯大学)

AI总结 UniGAP通过自适应图上采样缓解节点分类任务中的过平滑问题,提升模型性能并启发进一步研究。

详情

展开后加载摘要…

URL PDF HTML 收藏
2511.04120 2026-02-03 cs.CL

RIDE: Difficulty Evolving Perturbation with Item Response Theory for Mathematical Reasoning

RIDE: 通过项目反应理论进行数学推理的难度演变扰动

Xinyuan Li, Murong Xu, Wenbiao Tao, Hanlun Zhu, Yike Zhao, Jipeng Zhang, Yunshi Lan

机构 * East China Normal University(东华大学) The Hong Kong University of Science and Technology(香港科学与技术大学)

AI总结 RIDE通过项目反应理论生成更具挑战性的数学问题,评估大型语言模型的数学推理能力,实验显示其性能下降21.73%,验证了评估方法的有效性。

详情

展开后加载摘要…

URL PDF HTML 收藏