arXivDaily arXiv每日学术速递 周一至周五更新

期刊&会议

Conference on Computer Vision and Pattern Recognition · 会议 · Computer Vision

共收录 11877
2411.17261 2025-06-02 cs.CV cs.AI

HEIE: MLLM-Based Hierarchical Explainable AIGC Image Implausibility Evaluator

Fan Yang, Ru Zhen, Jianing Wang, Yanhao Zhang, Haoxiang Chen, Haonan Lu, Sicheng Zhao, Guiguang Ding

机构 * Tsinghua University(清华大学) BNRist OPPO AI Center(OPPO人工智能中心) Peking University(北京大学) Hangzhou Zhuoxi Institute of Brain and Intelligence(杭州智行脑科学与人工智能研究所)

Comments Accepted by CVPR 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2410.07599 2025-06-02 cs.CV

Adventurer: Optimizing Vision Mamba Architecture Designs for Efficiency

Feng Wang, Timing Yang, Yaodong Yu, Sucheng Ren, Guoyizhe Wei, Angtian Wang, Wei Shao, Yuyin Zhou, Alan Yuille, Cihang Xie

机构 * Johns Hopkins University(约翰霍普金斯大学) UC Berkeley(伯克利大学) University of Florida(佛罗里达大学) UC Santa Cruz(圣克拉拉大学)

Comments Published in CVPR 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2405.14858 2025-06-02 cs.CV

Mamba-R: Vision Mamba ALSO Needs Registers

Feng Wang, Jiahao Wang, Sucheng Ren, Guoyizhe Wei, Jieru Mei, Wei Shao, Yuyin Zhou, Alan Yuille, Cihang Xie

机构 * Johns Hopkins University(约翰霍普金斯大学) University of Florida(佛罗里达大学) UC Santa Cruz(加州圣克拉拉大学)

Comments Published in CVPR 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2505.24026 2025-06-02 cs.CV cs.AI

MaskAdapt: Unsupervised Geometry-Aware Domain Adaptation Using Multimodal Contextual Learning and RGB-Depth Masking

Numair Nadeem, Muhammad Hamza Asad, Saeed Anwar, Abdul Bais

机构 * University of Regina(里贾纳大学) University Canada West(加拿大西部大学) The University of Western Australia(西澳大学)

Comments 11 pages, 5 figures, presented at the IEEE/CVF Conference on Computer Vision and Pattern Recognition (CVPR) 2025. Reviewer comments available upon request

详情

展开后加载摘要…

URL PDF HTML 收藏
2505.23766 2025-05-30 cs.CV

Argus: Vision-Centric Reasoning with Grounded Chain-of-Thought

Yunze Man, De-An Huang, Guilin Liu, Shiwei Sheng, Shilong Liu, Liang-Yan Gui, Jan Kautz, Yu-Xiong Wang, Zhiding Yu

机构 * University of Illinois Urbana-Champaign(伊利诺伊大学厄巴纳-香槟分校) NVIDIA

Comments CVPR 2025. Project Page: https://yunzeman.github.io/argus/

详情

展开后加载摘要…

URL PDF HTML 收藏
2505.23763 2025-05-30 cs.CV

Sketch Down the FLOPs: Towards Efficient Networks for Human Sketch

Aneeshan Sain, Subhajit Maity, Pinaki Nath Chowdhury, Subhadeep Koley, Ayan Kumar Bhunia, Yi-Zhe Song

机构 * SketchX, CVSSP, University of Surrey, United Kingdom(SketchX、CVSSP、塞夫顿大学、英国) Department of Computer Science, University of Central Florida(计算机科学系、中央佛罗里达大学)

Comments Accepted at CVPR 2025, Project Page: https://subhajitmaity.me/SketchDownTheFLOPs

详情

展开后加载摘要…

URL PDF HTML 收藏
2505.23744 2025-05-30 cs.CV cs.AI

Boosting Domain Incremental Learning: Selecting the Optimal Parameters is All You Need

Qiang Wang, Xiang Song, Yuhang He, Jizhou Han, Chenhao Ding, Xinyuan Gao, Yihong Gong

机构 * Xi’an Jiaotong University(西安交通大学) Shenzhen University of Advanced Technology(深圳先进技术大学)

Comments Accepted at CVPR 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2505.23597 2025-05-30 cs.CV

Bridging Classical and Modern Computer Vision: PerceptiveNet for Tree Crown Semantic Segmentation

Georgios Voulgaris

机构 * University of Oxford(牛津大学)

Comments Accepted for publication at the IEEE/CVF Conference on Computer Vision and Pattern Recognition (CVPR) EarthVision

详情

展开后加载摘要…

URL PDF HTML 收藏
2505.23353 2025-05-30 eess.IV cs.AI cs.CV

Synthetic Generation and Latent Projection Denoising of Rim Lesions in Multiple Sclerosis

Alexandra G. Roberts, Ha M. Luu, Mert Şişman, Alexey V. Dimov, Ceren Tozlu, Ilhami Kovanlikaya, Susan A. Gauthier, Thanh D. Nguyen, Yi Wang

Comments Accepted full paper in Synthetic Data @ CVPR 2025 12 pages, 10 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2505.23290 2025-05-30 cs.SD cs.CV eess.AS

Wav2Sem: Plug-and-Play Audio Semantic Decoupling for 3D Speech-Driven Facial Animation

Hao Li, Ju Dai, Xin Zhao, Feng Zhou, Junjun Pan, Lei Li

机构 * Beihang University(北京航空航天大学) Peng Cheng Laboratory(鹏城实验室) North China University of Technology(华北理工大学) University of Washington(华盛顿大学) University of Copenhagen(哥本哈根大学)

Comments Accepted to CVPR 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2505.23180 2025-05-30 eess.IV cs.CV

Proximal Algorithm Unrolling: Flexible and Efficient Reconstruction Networks for Single-Pixel Imaging

Ping Wang, Lishun Wang, Gang Qu, Xiaodong Wang, Yulun Zhang, Xin Yuan

机构 * Westlake University(西湖大学) Zhejiang University(浙江大学) Shanghai Jiao Tong University(上海交通大学)

Comments Accepted by CVPR 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2505.23068 2025-05-30 cs.CV

URWKV: Unified RWKV Model with Multi-state Perspective for Low-light Image Restoration

Rui Xu, Yuzhen Niu, Yuezhou Li, Huangbiao Xu, Wenxi Liu, Yuzhong Chen

机构 * Fujian Key Laboratory of Network Computing and Intelligent Information Processing, College of Computer and Data Science, Fuzhou University(福建网络计算与智能信息处理重点实验室,计算机与数据科学学院,福州大学) Engineering Research Center of Big Data Intelligence, Ministry of Education(大数据智能工程研究中心,教育部)

Comments This paper has been accepted to CVPR 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2505.22859 2025-05-30 cs.CV

4DTAM: Non-Rigid Tracking and Mapping via Dynamic Surface Gaussians

Hidenobu Matsuki, Gwangbin Bae, Andrew J. Davison

机构 * Dyson Robotics Laboratory, Imperial College London(帝京理工学院伦敦分校动态机器人实验室)

Comments CVPR 2025. Project Page: https://muskie82.github.io/4dtam/

详情

展开后加载摘要…

URL PDF HTML 收藏
2505.22858 2025-05-30 cs.CV

A Probabilistic Jump-Diffusion Framework for Open-World Egocentric Activity Recognition

Sanjoy Kundu, Shanmukha Vellamcheti, Sathyanarayanan N. Aakur

机构 * CSSE Department, Auburn University(安全科学与工程系,阿伯丁大学)

Comments Extended abstract of arXiv:2504.03948 for CVPR 2025 EgoVis Workshop

详情

展开后加载摘要…

URL PDF HTML 收藏
2503.24129 2025-05-30 cs.CV cs.LG

It's a (Blind) Match! Towards Vision-Language Correspondence without Parallel Data

Dominik Schnaus, Nikita Araslanov, Daniel Cremers

机构 * TU Munich(慕尼黑技术大学) Munich Center for Machine Learning(慕尼黑机器学习中心)

Comments Accepted to CVPR 2025, Project page: https://dominik-schnaus.github.io/itsamatch/

详情

展开后加载摘要…

URL PDF HTML 收藏
2503.12821 2025-05-30 cs.CV cs.AI

From Head to Tail: Towards Balanced Representation in Large Vision-Language Models through Adaptive Data Calibration

Mingyang Song, Xiaoye Qu, Jiawei Zhou, Yu Cheng

机构 * Fudan University(复旦大学) Shanghai Artificial Intelligence Laboratory(上海人工智能实验室) Stony Brook University(石溪大学) The Chinese University of Hong Kong(香港中文大学)

Comments Accepted by CVPR 2025. Project Page: https://vlmlt.github.io/

详情

展开后加载摘要…

URL PDF HTML 收藏
2412.04383 2025-05-30 cs.CV cs.RO

SeeGround: See and Ground for Zero-Shot Open-Vocabulary 3D Visual Grounding

Rong Li, Shijie Li, Lingdong Kong, Xulei Yang, Junwei Liang

机构 * HKUST(GZ)(香港科技大学(广州)) I 2 R, A*STAR(I2R, A*STAR) National University of Singapore(新加坡国立大学) CSE, HKUST(香港科技大学计算机科学与工程系)

Comments CVPR 2025; 21 pages, 10 figures, 10 tables; Code at https://seeground.github.io/

详情

展开后加载摘要…

URL PDF HTML 收藏
2412.00175 2025-05-30 cs.CV cs.LG cs.SD eess.AS eess.IV

Circumventing shortcuts in audio-visual deepfake detection datasets with unsupervised learning

Stefan Smeu, Dragos-Alexandru Boldisor, Dan Oneata, Elisabeta Oneata

机构 * Bitdefender Politehnica Bucharest(巴尔干理工大学)

Comments Accepted as a highlight paper at the IEEE/CVF Conference on Computer Vision and Pattern Recognition (CVPR) 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2307.07333 2025-05-30 cs.CV

SynTable: A Synthetic Data Generation Pipeline for Unseen Object Amodal Instance Segmentation of Cluttered Tabletop Scenes

Zhili Ng, Haozhe Wang, Zhengshen Zhang, Francis Tay Eng Hock, Marcelo H. Ang

机构 * Advanced Robotics Centre, National University of Singapore(新加坡国立大学先进机器人中心)

Comments Camera-ready version for SynData4CV Workshop @ CVPR 2025. 18 Pages, 11 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2505.22616 2025-05-29 cs.CV eess.IV

PS4PRO: Pixel-to-pixel Supervision for Photorealistic Rendering and Optimization

Yezhi Shen, Qiuchen Zhai, Fengqing Zhu

机构 * School of Electrical and Computer Engineering, Purdue University(电子与计算机工程学院,普渡大学)

Comments Accepted to the CVPR 2025 Workshop on Autonomous Driving (WAD)

详情

展开后加载摘要…

URL PDF HTML 收藏
2505.22429 2025-05-29 cs.CV cs.RO

Zero-Shot 3D Visual Grounding from Vision-Language Models

Rong Li, Shijie Li, Lingdong Kong, Xulei Yang, Junwei Liang

机构 * HKUST(GZ)(香港科技大学(广州)) I 2 R, A*STAR(I2R, A*STAR) NUS(国立大学) CSE, HKUST(计算机科学与工程系,香港科技大学)

Comments 3D-LLM/VLA @ CVPR 2025; Project Page at https://seeground.github.io/

详情

展开后加载摘要…

URL PDF HTML 收藏
2505.22427 2025-05-29 cs.CV

RC-AutoCalib: An End-to-End Radar-Camera Automatic Calibration Network

Van-Tin Luu, Yon-Lin Cai, Vu-Hoang Tran, Wei-Chen Chiu, Yi-Ting Chen, Ching-Chun Huang

机构 * National Yang Ming Chiao Tung University(国立阳明交通大学) Ho Chi Minh City University of Technology and Education(胡志明市技术与教育大学)

Comments Accepted to CVPR 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2505.22079 2025-05-29 cs.CV

Bringing CLIP to the Clinic: Dynamic Soft Labels and Negation-Aware Learning for Medical Analysis

Hanbin Ko, Chang-Min Park

机构 * Interdisciplinary Program in Bioengineering, Seoul National University Graduate School(生物工程跨学科项目,首尔国立大学研究生院) Integrated Major in Innovative Medical Science, Seoul National University Graduate School(创新医学联合专业,首尔国立大学研究生院) Department of Radiology, Seoul National University Hospital(放射科,首尔国立大学医院)

Comments 16 pages (8 main, 2 references, 6 appendix), 13 figures. Accepted to CVPR 2025. This author-accepted manuscript includes an expanded ethics/data user agreement section. The final version will appear in the Proceedings of CVPR 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2505.21943 2025-05-29 cs.CV

Point-to-Region Loss for Semi-Supervised Point-Based Crowd Counting

Wei Lin, Chenyang Zhao, Antoni B. Chan

机构 * Department of Computer Science, City University of Hong Kong(计算机科学系,香港城市大学)

Comments accepted by CVPR-2025(highlight)

详情

展开后加载摘要…

URL PDF HTML 收藏
2505.21561 2025-05-29 cs.CV cs.LG

Knowledge Distillation Approach for SOS Fusion Staging: Towards Fully Automated Skeletal Maturity Assessment

Omid Halimi Milani, Amanda Nikho, Marouane Tliba, Lauren Mills, Ahmet Enis Cetin, Mohammed H Elnagar

机构 * Department of Electrical and Computer Engineering, University of Illinois Chicago(伊利诺伊大学芝加哥分校电气与计算机工程系) Department of Orthodontics, College of Dentistry, University of Illinois Chicago(伊利诺伊大学芝加哥分校牙科学院正畸学系) University of Orleans(奥尔良大学)

Comments This paper has been accepted to the CVPR Workshop 2025, to be held in Nashville, Tennessee

详情

展开后加载摘要…

URL PDF HTML 收藏
2505.21513 2025-05-29 cs.CV cs.AI cs.LG

Enhancing Vision Transformer Explainability Using Artificial Astrocytes

Nicolas Echevarrieta-Catalan, Ana Ribas-Rodriguez, Francisco Cedron, Odelia Schwartz, Vanessa Aguiar-Pulido

机构 * University of Miami(迈阿密大学) University of A Coruña(阿罗乌纳大学)

Comments LXCV Workshop at IEEE / CVF Computer Vision and Pattern Recognition Conference (CVPR) 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2504.09960 2025-05-29 cs.CV

Dual-Path Enhancements in Event-Based Eye Tracking: Augmented Robustness and Adaptive Temporal Modeling

Hoang M. Truong, Vinh-Thuan Ly, Huy G. Tran, Thuan-Phat Nguyen, Tram T. Doan

机构 * University of Science, VNU-HCM(越南胡志明市国家大学) Vietnam National University(越南国家大学)

Comments Camera-ready version for CVPRW 2025. Accepted for presentation at the IEEE/CVF Conference on Computer Vision and Pattern Recognition Workshops (CVPRW 2025)

Journal ref IEEE Conference on Computer Vision and Pattern Recognition (CVPR) Workshops, Nashville, 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2411.16310 2025-05-29 cs.CV

Functionality understanding and segmentation in 3D scenes

Jaime Corsetti, Francesco Giuliari, Alice Fasoli, Davide Boscaini, Fabio Poiesi

机构 * Fondazione Bruno Kessler(布鲁诺·凯塞勒基金会) University of Trento(特伦托大学)

Comments CVPR 2025 Highlight. Camera ready version. 20 pages, 12 figures, 7 tables. Fixed typo in Eq.2

详情

展开后加载摘要…

URL PDF HTML 收藏
2505.21377 2025-05-28 cs.CV

Empowering Vector Graphics with Consistently Arbitrary Viewing and View-dependent Visibility

Yidi Li, Jun Xiao, Zhengda Lu, Yiqun Wang, Haiyong Jiang

机构 * School of Artificial Intelligence, University of Chinese Academy of Sciences(人工智能学院,中国科学院大学)

Comments CVPR 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2505.21335 2025-05-28 cs.GR cs.AI cs.CV cs.LG cs.RO

Structure from Collision

Takuhiro Kaneko

机构 * NTT Corporation(日本通信用电讯公司)

Comments Accepted to CVPR 2025 (Highlight). Project page: https://www.kecl.ntt.co.jp/people/kaneko.takuhiro/projects/sfc/

详情

展开后加载摘要…

URL PDF HTML 收藏