arXivDaily arXiv每日学术速递 周一至周五更新

高校专区

The Hong Kong University of Science and Technology(香港科技大学)

2026-02-09 至 2026-02-09 共收录 12
2602.06949 2026-02-09 cs.RO cs.AI cs.CV cs.LG

DreamDojo: A Generalist Robot World Model from Large-Scale Human Videos

DreamDojo:从大规模人类视频中学习通用机器人世界模型

Shenyuan Gao, William Liang, Kaiyuan Zheng, Ayaan Malik, Seonghyeon Ye, Sihyun Yu, Wei-Cheng Tseng, Yuzhu Dong, Kaichun Mo, Chen-Hsuan Lin, Qianli Ma, Seungjun Nah, Loic Magne, Jiannan Xiang, Yuqi Xie, Ruijie Zheng, Dantong Niu, You Liang Tan, K. R. Zentner, George Kurian, Suneel Indupuru, Pooya Jannaty, Jinwei Gu, Jun Zhang, Jitendra Malik, Pieter Abbeel, Ming-Yu Liu, Yuke Zhu, Joel Jang, Linxi "Jim" Fan

机构 * NVIDIA HKUST(香港科技大学) UC Berkeley(加州大学伯克利分校) Stanford(斯坦福大学) KAIST(韩国科学技术院) UofT(多伦多大学) UCSD(加州大学圣地亚哥分校) UT Austin(德克萨斯大学奥斯汀分校)

AI总结 DreamDojo通过大规模人类视频学习通用机器人世界模型,解决动作标签稀缺问题,实现高精度物理理解和实时交互。

Comments Project page: https://dreamdojo-world.github.io/

详情

展开后加载摘要…

URL PDF HTML 收藏
2602.06674 2026-02-09 cs.CV cs.HC cs.LG

CytoCrowd: A Multi-Annotator Benchmark Dataset for Cytology Image Analysis

CytoCrowd:一种用于细胞学图像分析的多标注基准数据集

Yonghao Si, Xingyuan Zeng, Zhao Chen, Libin Zheng, Caleb Chen Cao, Lei Chen, Jian Yin

机构 * Sun Yat-sen University(中山大学) Hong Kong University of Science and Technology(香港科技大学)

AI总结 CytoCrowd是一个包含446张高分辨率细胞学图像的数据集,提供四个病理学家的冲突注释和一个资深专家的金牌标准,用于评估注释聚合算法和计算机视觉任务的基准测试。

详情

展开后加载摘要…

URL PDF HTML 收藏
2602.06453 2026-02-09 cs.LG

On the Plasticity and Stability for Post-Training Large Language Models

关于后训练大语言模型的可塑性与稳定性

Wenwen Qiang, Ziyin Gu, Jiahuan Zhou, Jie Hu, Jingyao Wang, Changwen Zheng, Hui Xiong

机构 * Institute of Software Chinese Academy of Sciences, Beijing, China(中国科学院软件研究所) University of the Chinese Academy of Sciences, Beijing, China(中国科学院大学) Wangxuan Institute of Computer Technology, Peking University, Beijing, China(北京大学王轩计算机技术研究所) Thrust of Artificial Intelligence, The Hong Kong University of Science and Technology (Guangzhou), China(香港科技大学(广州)人工智能研究所) Department of Computer Science and Engineering, The Hong Kong University of Science and Technology Hong Kong SAR, China(香港科技大学(香港特别行政区)计算机科学与工程系)

AI总结 本文提出PCR框架,通过概率方法解决GRPO中可塑性与稳定性之间的几何冲突,提升训练稳定性与推理性能。

详情

展开后加载摘要…

URL PDF HTML 收藏
2602.06319 2026-02-09 cs.AI

Exposing Weaknesses of Large Reasoning Models through Graph Algorithm Problems

通过图算法问题揭示大推理模型的弱点

Qifan Zhang, Jianhao Ruan, Aochuan Chen, Kang Zeng, Nuo Chen, Jing Tang, Jia Li

机构 * The Hong Kong University of Science and Technology (Guangzhou)(香港科学与技术大学(广州))

AI总结 GrAlgoBench通过图算法问题揭示大推理模型在长上下文推理和过度思考方面的不足,为改进推理研究提供严格测试平台。

详情

展开后加载摘要…

URL PDF HTML 收藏
2602.05885 2026-02-09 cs.LG cs.AI cs.CL

Dr. Kernel: Reinforcement Learning Done Right for Triton Kernel Generations

Dr. Kernel:为Triton内核生成正确强化学习

Wei Liu, Jiawei Xu, Yingru Li, Longtao Zheng, Tianjian Li, Qian Liu, Junxian He

机构 * HKUST(香港科技大学) TikTok CUHK(SZ)(香港中文大学(深圳)) NTU(国立科技大学)

AI总结 Dr. Kernel通过强化学习方法优化内核生成,其模型在Kernelbench测试中达到与Claude-4.5-Sonnet相当的性能,并在速度提升方面超越其他模型。

详情

展开后加载摘要…

URL PDF HTML 收藏
2602.05444 2026-02-09 cs.CL

Causal Front-Door Adjustment for Robust Jailbreak Attacks on LLMs

因果前门调整用于对抗大语言模型的鲁棒性劫持攻击

Yao Zhou, Zeen Song, Wenwen Qiang, Fengge Wu, Shuyi Zhou, Changwen Zheng, Hui Xiong

机构 * Institute of Software Chinese Academy of Sciences, Beijing, China(中国科学院软件研究所) University of Chinese Academy of Sciences, Beijing, China(中国科学院大学) Institute of Information Engineering Chinese Academy of Sciences, Beijing, China(中国科学院信息工程研究所) Hong Kong University of Science and Technology, China, Hong Kong, China(香港科学与技术大学)

AI总结 本文提出CFA²攻击方法,通过因果前门准则和稀疏自编码器实现对大语言模型的鲁棒性劫持,提升攻击成功率并提供机制解释。

详情

展开后加载摘要…

URL PDF HTML 收藏
2602.00169 2026-02-09 cond-mat.mtrl-sci cs.AI

Towards Agentic Intelligence for Materials Science

迈向材料科学的代理智能

Huan Zhang, Yizhan Li, Wenhao Huang, Ziyu Hou, Yu Song, Xuye Liu, Farshid Effaty, Jinya Jiang, Sifan Wu, Qianggang Ding, Izumi Takahara, Leonard R. MacGillivray, Teruyasu Mizoguchi, Tianshu Yu, Lizi Liao, Yuyu Luo, Yu Rong, Jia Li, Ying Diao, Heng Ji, Bang Liu

机构 * DIRO & Institut Courtois, Université de Montréal(蒙特利尔大学DIRO与Courtois研究所) Mila – Quebec AI Institute(魁北克AI研究所) University of Waterloo(滑铁卢大学) Université de Sherbrooke(Sherbrooke大学) University of California, San Diego(加州大学圣地亚哥分校) The University of Tokyo, Institute of Industrial Science(东京大学工业科学研究所) The Chinese University of Hong Kong, Shenzhen(香港中文大学(深圳)) Singapore Management University(新加坡管理学院) The Hong Kong University of Science and Technology, Guangzhou(香港科学与技术大学(广州)) Alibaba DAMO Academy(阿里巴巴达摩院) University of Illinois Urbana-Champaign(伊利诺伊大学厄巴纳-香槟分校) CIFAR AI Chair(CIFAR人工智能主席)

AI总结 本文提出以流程为中心的代理系统框架,通过整合人工智能与材料科学,推动材料发现的自主化与智能化。

Comments 81 pages

详情

展开后加载摘要…

URL PDF HTML 收藏
2601.02075 2026-02-09 cs.CE cs.LG

MDAgent2: Large Language Model for Code Generation and Knowledge Q&A in Molecular Dynamics

MDAgent2:用于分子动力学中代码生成和知识问答的大型语言模型

Zhuofan Shi, Hubao A, Yufei Shao, Dongliang Huang, Hongxu An, Chunxiao Xin, Haiyang Shen, Zhenyu Wang, Yunshan Na, Gang Huang, Xiang Jing

机构 * Peking University(北京大学) National Key Laboratory of Data Space Technology and System(国家数据空间技术与系统重点实验室) The Hong Kong University of Science and Technology(香港科学与技术大学) Liaoning Technical University(辽宁技术大学) Wenjing Future Lab (Beijing) Technology Co., Ltd(文景未来实验室(北京)科技有限公司)

AI总结 MDAgent2是首个能同时进行分子动力学知识问答和代码生成的端到端框架,通过领域特定数据集和强化学习方法提升性能。

Comments 24 pages,4 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.03054 2026-02-09 cs.LG cs.AI

Calibration and Transformation-Free Weight-Only LLMs Quantization via Dynamic Grouping

无需校准和转换的权重-only LLMs 量化 via 动态分组

Xinzhe Zheng, Zhen-Qun Yang, Zishan Liu, Haoran Xie, S. Joe Qin, Arlene Chen, Fangzhen Lin

机构 * Division of Artificial Intelligence, School of Data Science, Lingnan University, Hong Kong, China(岭南大学人工智能学院) Department of Computer Science and Engineering, The Hong Kong University of Science and Technology, Hong Kong, China(香港理工大学计算机科学与工程系) Department of Computing, The Hong Kong Polytechnic University, Hong Kong, China(香港理工大学计算机系) Xiaoi Robot Inc., Shanghai, China(小蚁机器人有限公司)

AI总结 MSB提出一种无需校准和转换的低比特PTQ方法,通过动态分组优化实现多尺度量化,提升LLMs在内存和计算约束下的性能。

Comments 34 pages, 10 figures. Version 3 corrects the bit-length error and adds new experiments and analysis; the core methodology remains unchanged. Under review

详情

展开后加载摘要…

URL PDF HTML 收藏
2505.11781 2026-02-09 cs.LG

Multi-Order Wavelet Derivative Transform for Deep Time Series Forecasting

多阶小波导数变换用于深度时间序列预测

Ziyu Zhou, Jiaxi Hu, Qingsong Wen, James T. Kwok, Yuxuan Liang

机构 * The Hong Kong University of Science and Technology (Guangzhou)(香港科学与技术大学(广州)) Squirrel AI Learning

AI总结 多阶小波导数变换通过提取时间感知模式,提升深度时间序列预测的精度与效率。

Comments Preprint

详情

展开后加载摘要…

URL PDF HTML 收藏
2410.04010 2026-02-09 cs.LG cs.AI cs.CL cs.NE

Hyperbolic Fine-Tuning for Large Language Models

双曲微调用于大语言模型

Menglin Yang, Ram Samarth B B, Aosong Feng, Bo Xiong, Jihong Liu, Irwin King, Rex Ying

机构 * HKUST(GZ)(香港科技大学(广州)) HKUST(香港科技大学) Indian Institute of Science(印度科学研究院) Yale University(耶鲁大学) Stanford University(斯坦福大学) The Chinese University of Hong Kong(香港中文大学)

AI总结 HypLoRA通过在双曲空间中进行低秩适应,提升大语言模型在算术和常识推理任务中的性能。

Comments NeurIPS 2025; https://github.com/marlin-codes/HypLoRA

详情

展开后加载摘要…

URL PDF HTML 收藏
2407.14069 2026-02-09 cs.CV

Self-Supervised Video Representation Learning in a Heuristic Decoupled Perspective

基于启发式解耦视角的自监督视频表示学习

Zeen Song, Wenwen Qiang, Changwen Zheng, Hui Xiong, Gang Hua

机构 * National Key Laboratory of Space Integrated Information System, Institute of Software Chinese Academy of Sciences(空间信息集成国家重点实验室,软件研究所中国科学院) University of Chinese Academy of Sciences(中国科学院大学) Hong Kong University of Science and Technology(香港理工大学) Dolby Laboratories Inc(杜比实验室有限公司) Xi’an Jiaotong University(西安交通大学)

AI总结 本文提出BOD-VCL方法,通过解耦静态和动态语义,提升视频对比学习的性能。

详情

展开后加载摘要…

URL PDF HTML 收藏