arXivDaily arXiv每日学术速递 周一至周五更新

高校专区

The Hong Kong University of Science and Technology(香港科技大学)

2026-01-16 至 2026-01-16 共收录 8
2601.10513 2026-01-16 cs.CL cs.HC

AEQ-Bench: Measuring Empathy of Omni-Modal Large Models

AEQ-Bench:衡量多模态大模型的共情能力

Xuan Luo, Lewei Yao, Libo Zhao, Lanqing Hong, Kai Chen, Dehua Tao, Daxin Tan, Ruifeng Xu, Jing Li

机构 * The Hong Kong Polytechnic University(香港理工大学) The Harbin Institute of Technology(哈尔滨工业大学) Huawei(华为) Hong Kong University of Science and Technology(香港理工大学) Shenzhen Loop Area Institute(深圳河套学院)

AI总结 AEQ-Bench通过评估多模态大模型在情感识别和音频响应共情判断上的能力,揭示了音频输出能力对模型性能的影响及非语言表达评估的局限性。

详情

展开后加载摘要…

URL PDF HTML 收藏
2601.10365 2026-01-16 cs.RO

FastStair: Learning to Run Up Stairs with Humanoid Robots

FastStair: 人类机器人跑步上楼梯的学习

Yan Liu, Tao Yu, Haolin Song, Hongbo Zhu, Nianzong Hu, Yuzhi Hao, Xiuyong Yao, Xizhe Zang, Hua Chen, Jie Zhao

机构 * School of Mechanics Engineering, Harbin Institute of Technology (HIT), Harbin Heilongjiang 150001, China(哈尔滨工业大学机械工程学院) LimX Dynamics, Shenzhen, China(LimX Dynamics) Zhejiang University-University of Illinois Urbana-Champaign Institute (ZJUI), Haining, China(浙江大学-伊利诺伊大学厄巴纳-香槟分校联合研究所) Department of Electronic Engineering and Information Science (EEIS), University of Science and Technology of China, Hefei 230027, China(中国科学技术大学电子工程与信息科学系) Hong Kong University of Science and Technology, Hong Kong SAR, China(香港科技大学) Department of Mechanical Engineering, National University of Singapore, Singapore 117575(新加坡国立大学机械工程系)

AI总结 FastStair通过结合基于模型的规划器和强化学习,实现仿人机器人快速稳定的楼梯上升,展示了在高速和长楼梯上的卓越性能。

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.23463 2026-01-16 cs.LG cs.CR stat.ML

Differential Privacy as a Perk: Federated Learning over Multiple-Access Fading Channels with a Multi-Antenna Base Station

差分隐私作为奖励:多接入衰落信道上的联邦学习与多天线基站

Hao Liang, Haifeng Wen, Kaishun Wu, Hong Xing

机构 * IoT Thrust, The Hong Kong University of Science and Technology (Guangzhou)(科技与应用大学信息学部,香港科学与技术大学(广州)) Department of ECE, The Hong Kong University of Science and Technology(电子与计算机工程系,香港科学与技术大学)

AI总结 本文研究多接入衰落信道上的联邦学习,通过多天线基站实现差分隐私保护,推导新型DP界并优化收敛-隐私权衡。

Comments 13 pages, 6 figures, submitted for possible publication

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.14981 2026-01-16 cs.CV

SPATIALGEN: Layout-guided 3D Indoor Scene Generation

SPATIALGEN: 布局引导的3D室内场景生成

Chuan Fang, Heng Li, Yixun Liang, Jia Zheng, Yongsen Mao, Yuan Liu, Rui Tang, Zihan Zhou, Ping Tan

机构 * Hong Kong University of Science and Technology(香港科技大学) Manycore Tech Inc(Manycore科技公司)

AI总结 SPATIALGEN通过布局引导的多视角多模态扩散模型生成高质量3D室内场景,解决现有方法在视觉质量、多样性及语义一致性方面的不足。

Comments 3D scene generation; diffusion model; Scene reconstruction and understanding

详情

展开后加载摘要…

URL PDF HTML 收藏
2601.10124 2026-01-16 cs.CV

VQ-Seg: Vector-Quantized Token Perturbation for Semi-Supervised Medical Image Segmentation

VQ-Seg: 基于向量量化令牌扰动的半监督医学图像分割

Sicheng Yang, Zhaohu Xing, Lei Zhu

机构 * The Hong Kong University of Science and Technology (Guangzhou)(香港科技大学(广州))

AI总结 VQ-Seg通过向量量化和量化扰动模块提升半监督医学图像分割性能,有效解决传统dropout正则化问题。

Comments Accepted by NeurIPS 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2512.22972 2026-01-16 cs.CV eess.SP

Wavelet-based Multi-View Fusion of 4D Radar Tensor and Camera for Robust 3D Object Detection

基于小波的4D雷达张量与相机多视图融合用于鲁棒3D目标检测

Runwei Guan, Jianan Liu, Shaofeng Liang, Fangqiang Ding, Shanliang Yao, Xiaokai Bai, Daizong Liu, Tao Huang, Guoqiang Mao, Hui Xiong

机构 * Thrust of Artificial Intelligence, Hong Kong University of Science and Technology (Guangzhou)(人工智能 thrust,香港科技大学(广州)) Momoniai AI Department of Mechanical Engineering, Massachusetts Institute of Technology(机械工程系,麻省理工学院) School of Information Engineering, Yancheng Institute of Technology(信息工程学院,盐城科技学院) College of Information Science and Electronic Engineering, Zhejiang University(信息科学与电子工程学院,浙江大学) Institute for Math & AI, Wuhan University(数学与人工智能研究所,武汉大学) College of Science and Engineering and the Centre for AI and Data Science Innovation, James Cook University(科学与工程学院及人工智能与数据科学创新中心,詹姆斯库克大学) School of Transportation, Southeast University(交通运输学院,东南大学)

AI总结 WRCFormer通过小波注意力模块和几何引导渐进融合机制,高效融合4D雷达张量与相机图像,提升3D目标检测在恶劣天气下的鲁棒性。

Comments 10 pages, 10 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2505.02064 2026-01-16 cs.CV

RTV-Bench: Benchmarking MLLM Continuous Perception, Understanding and Reasoning through Real-Time Video

RTV-Bench: 通过实时视频对多模态大语言模型的连续感知、理解和推理进行基准测试

Shuhang Xun, Sicheng Tao, Jungang Li, Yibo Shi, Zhixin Lin, Zhanhui Zhu, Yibo Yan, Hanqian Li, Linghao Zhang, Shikang Wang, Yixin Liu, Hanbo Zhang, Ying Ma, Xuming Hu

机构 * HIT(哈尔滨工业大学) HKUST (GZ)(香港科技大学(广州)) HKUST(香港科技大学) XJTU(西安交通大学) SDU(山东大学) CityU(城市大学) HUST(华中科技大学)

AI总结 RTV-Bench通过实时视频对多模态大语言模型的连续感知、理解和推理能力进行细粒度基准测试,揭示了实时模型在长时段视频处理中的性能优势与局限。

Comments Accepted by NeurIPS 2025 Datasets and Benchmarks Track;

详情

展开后加载摘要…

URL PDF HTML 收藏
2502.07115 2026-01-16 cs.LG cs.AI math.OC

Online Scheduling for LLM Inference with KV Cache Constraints

在线调度用于具有KV缓存约束的LLM推理

Patrick Jaillet, Jiashuo Jiang, Konstantina Mellou, Marco Molinaro, Chara Podimata, Zijie Zhou

机构 * Department of Electrical Engineering and Computer Science, Massachusetts Institute of Technology(麻省理工学院电子工程与计算机科学系) HKUST(香港科技大学) Microsoft Research(微软研究院) Sloan School of Management, Massachusetts Institute of Technology(斯隆管理学院,麻省理工学院) Operations Research Center, Massachusetts Institute of Technology(运营研究中心,麻省理工学院)

AI总结 本文提出了一种在线调度算法,通过理论建模和实证验证,在管理KV缓存内存的同时最小化LLM推理延迟,显著优于现有基准算法。

详情

展开后加载摘要…

URL PDF HTML 收藏