arXivDaily arXiv每日学术速递 周一至周五更新

高校专区

The University of Hong Kong(香港大学)

2025-11-18 至 2025-11-18 共收录 7
2511.13647 2025-11-18 cs.CV

Part-X-MLLM: Part-aware 3D Multimodal Large Language Model

Chunshi Wang, Junliang Ye, Yunhan Yang, Yang Li, Zizhuo Lin, Jun Zhu, Zhuo Chen, Yawei Luo, Chunchao Guo

机构 * Zhejiang University(浙江大学) Tencent Hunyuan(腾讯文言) Tsinghua University(清华大学) The University of Hong Kong(香港大学)

详情

展开后加载摘要…

URL PDF HTML 收藏
2511.12594 2025-11-18 cs.CV

Seg-VAR: Image Segmentation with Visual Autoregressive Modeling

Rongkun Zheng, Lu Qi, Xi Chen, Yi Wang, Kun Wang, Hengshuang Zhao

机构 * The University of Hong Kong(香港大学) Insta360 Shanghai Artificial Intelligence Laboratory(上海人工智能实验室) Shanghai Innovation Institute(上海创新研究院) SenseTime Research(商汤科技研究院)

Comments NeurIPS 2025, 22 pages

详情

展开后加载摘要…

URL PDF HTML 收藏
2503.12015 2025-11-18 cs.CV

QDM: Quadtree-Based Region-Adaptive Sparse Diffusion Models for Efficient Image Super-Resolution

Donglin Yang, Paul Vicol, Xiaojuan Qi, Renjie Liao, Xiaofan Zhang

机构 * The University of Hong Kong(香港大学) Google DeepMind(谷歌DeepMind) The University of British Columbia(不列颠哥伦比亚大学) Shanghai Jiao Tong University(上海交通大学)

详情

展开后加载摘要…

URL PDF HTML 收藏
2409.14830 2025-11-18 cs.CR cs.AI cs.HC cs.LG

Identify As A Human Does: A Pathfinder of Next-Generation Anti-Cheat Framework for First-Person Shooter Games

Jiayi Zhang, Chenxin Sun, Yue Gu, Qingyu Zhang, Jiayi Lin, Xiaojiang Du, Chenxiong Qian

机构 * School of Computing and Data Science, The University of Hong Kong(计算与数据科学学院,香港大学) Department of Electrical and Computer Engineering, Stevens Institute of Technology(电气与计算机工程系,史蒂文斯理工学院)

详情

展开后加载摘要…

URL PDF HTML 收藏
2511.12020 2025-11-18 cs.CV

LIHE: Linguistic Instance-Split Hyperbolic-Euclidean Framework for Generalized Weakly-Supervised Referring Expression Comprehension

Xianglong Shi, Silin Cheng, Sirui Zhao, Yunhan Jiang, Enhong Chen, Yang Liu, Sebastien Ourselin

机构 * University of Science and Technology of China(中国科学技术大学) The University of Hong Kong(香港大学) King’s College London(伦敦国王学院) Peking University(北京大学)

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.00298 2025-11-18 cs.CV

AniMer+: Unified Pose and Shape Estimation Across Mammalia and Aves via Family-Aware Transformer

Liang An, Jin Lyu, Li Lin, Pujin Cheng, Yebin Liu, Xiaoying Tang

机构 * Department of Electronic and Electrical Engineering, Southern University of Science and Technology, Shenzhen, China(南方科技大学电子与电气工程系) Department of Automation, Tsinghua University, Beijing, China(清华大学自动化系) Jiaxing Research Institute, Southern University of Science and Technology, Jiaxing, China(南方科技大学嘉兴研究所) Department of Electrical and Electronic Engineering, the University of Hong Kong, Hong Kong, China(香港大学电子与电气工程系)

Comments Accepted to TPAMI2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2505.08283 2025-11-18 cs.LG cs.CV

DPL: Decoupled Prototype Learning for Enhancing Robustness of Vision-Language Transformers to Missing Modalities

Jueqing Lu, Yuanyuan Qi, Xiaohao Yang, Shuaicheng Niu, Fucai Ke, Shujie Zhou, Wei Tan, Jionghao Lin, Wray Buntine, Hamid Rezatofighi, Lan Du

机构 * Monash University(墨尔本大学) Nanyang Technological University(南洋理工大学) The University of Hong Kong(香港大学) VinUni

Comments Updates to v1. Added new coauthors and extended the experimental section

详情

展开后加载摘要…

URL PDF HTML 收藏