arXivDaily arXiv每日学术速递 周一至周五更新

期刊&会议

Conference on Computer Vision and Pattern Recognition · 会议 · Computer Vision

共收录 11876
2311.16304 2025-11-07 cs.CV

Robust Self-calibration of Focal Lengths from the Fundamental Matrix

Viktor Kocur, Daniel Kyselica, Zuzana Kukelova

机构 * Faculty of Mathematics, Physics and Informatics, Comenius University in Bratislava(数学、物理与信息学系,布拉迪斯拉瓦康门大学) Visual Recognition Group, Faculty of Electrical Engineering, Czech Technical University in Prague(视觉识别组,布拉格捷克技术大学)

Comments Pubslished in CVPR 2024. Accepted: 26.2.2024. Published: 16.6.2024. This work was funded by the Horizon-Widera-2021 European Twinning project TERAIS G.A. n. 101079338. Code available: https://github.com/kocurvik/robust_self_calibration and https://doi.org/10.5281/zenodo.14584742

详情

展开后加载摘要…

URL PDF HTML 收藏
2303.16078 2025-11-07 cs.CV

Practical solutions to the relative pose of three calibrated cameras

Charalambos Tzamos, Viktor Kocur, Yaqing Ding, Daniel Barath, Zuzana Berger Haladova, Torsten Sattler, Zuzana Kukelova

机构 * Visual Recognition Group, Faculty of Electrical Engineering, Czech Technical University in Prague(视觉识别组,电气工程系,布拉格捷克技术大学) Faculty of Mathematics, Physics and Informatics, Comenius University in Bratislava(数学、物理和信息学系,布拉迪斯拉发康门尼乌斯大学) ETH Zürich (Zurich), HUN-REN SZTAKI (Budapest)(苏黎世ETH,匈牙利-中国SZTAKI(布达佩斯)) Czech Institute of Informatics, Robotics and Cybernetics, Czech Technical University in Prague(捷克信息学、机器人学与自动控制研究所,布拉格捷克技术大学)

Comments Paper presented at CVPR 2025 (DOI: 10.1109/CVPR52734.2025.02041). Code available at https://github.com/kocurvik/threeview and https://doi.org/10.5281/zenodo.16599943. Data available at https://doi.org/10.5281/zenodo.16603086

详情

展开后加载摘要…

URL PDF HTML 收藏
2401.01887 2025-11-07 cs.CV

LEAP-VO: Long-term Effective Any Point Tracking for Visual Odometry

Weirong Chen, Le Chen, Rui Wang, Marc Pollefeys

机构 * TU Munich(慕尼黑技术大学) Munich Center for Machine Learning(慕尼黑机器学习中心) MPI for Intelligent Systems(智能系统研究所) Microsoft(微软公司)

Comments Accepted to CVPR 2024. Project page: https://wrchen530.github.io/projects/leapvo

详情

展开后加载摘要…

URL PDF HTML 收藏
2511.01425 2025-11-04 cs.AI cs.CV

Learning to Seek Evidence: A Verifiable Reasoning Agent with Causal Faithfulness Analysis

Yuhang Huang, Zekai Lin, Fan Zhong, Lei Liu

机构 * Institute of Biomedical Science, Fudan University(复旦大学生物医学研究院) Fudan University(复旦大学) Intelligent Medicine Institute, Fudan University(复旦大学智能医学研究院) Shanghai Institute of Infectious Disease and Biosecurity, Fudan University(复旦大学传染病与生物安全研究所) Shanghai Institute of Stem Cell Research and Clinical Translation, Fudan University(复旦大学干细胞研究与临床转化研究所)

Comments 12 pages, 3 figures. Under review at the Conference on Computer Vision and Pattern Recognition (CVPR) 2026

详情

展开后加载摘要…

URL PDF HTML 收藏
2502.20056 2025-11-04 cs.CV cs.AI

Enhanced Contrastive Learning with Multi-view Longitudinal Data for Chest X-ray Report Generation

Kang Liu, Zhuoqi Ma, Xiaolu Kang, Yunan Li, Kun Xie, Zhicheng Jiao, Qiguang Miao

机构 * School of Computer Science and Technology, Xidian University(西安电子科技大学计算机科学与技术学院) Xi’an Key Laboratory of Big Data and Intelligent Vision(西安大数据与智能视觉重点实验室) Key Laboratory of Collaborative Intelligence Systems, Ministry of Education, Xidian University(教育部协同智能系统重点实验室) Warren Alpert Medical School, Brown University(布朗大学沃伦·阿尔佩特医学学院)

Comments Accepted by CVPR 2025

Journal ref 2025 IEEE/CVF Conference on Computer Vision and Pattern Recognition (CVPR), Nashville, TN, USA, 2025, pp. 10348-10359

详情

展开后加载摘要…

URL PDF HTML 收藏
2412.07140 2025-11-04 cs.CV

FIRE: Robust Detection of Diffusion-Generated Images via Frequency-Guided Reconstruction Error

Beilin Chu, Xuan Xu, Xin Wang, Yufei Zhang, Weike You, Linna Zhou

机构 * School of CyberSpace Security, Beijing University of Posts and Telecommunications(网络安全学院,北京邮电大学)

Comments 14 pages, 14 figures. Accepted to CVPR 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2511.00510 2025-11-04 cs.CV cs.RO eess.IV

OmniTrack++: Omnidirectional Multi-Object Tracking by Learning Large-FoV Trajectory Feedback

Kai Luo, Hao Shi, Kunyu Peng, Fei Teng, Sheng Wu, Kaiwei Wang, Kailun Yang

机构 * Hunan University(湖南大学) Zhejiang University(浙江大学) Ant Group(蚂蚁集团) Karlsruhe Institute of Technology(卡尔斯鲁厄理工学院) INSAIT, Sofia University ``St. Kliment Ohridski''(INSAIT,索菲亚大学『圣克莱门特·奥赫里迪斯』)

Comments Extended version of CVPR 2025 paper arXiv:2503.04565. Datasets and code will be made publicly available at https://github.com/xifen523/OmniTrack

详情

展开后加载摘要…

URL PDF HTML 收藏
2402.17726 2025-11-04 cs.CV

VRP-SAM: SAM with Visual Reference Prompt

Yanpeng Sun, Jiahui Chen, Shan Zhang, Xinyu Zhang, Xiaofan Li, Qiang Chen, Gang Zhang, Errui Ding, Jingdong Wang, Zechao Li

机构 * Nanjing University of Science and Technology(南京理工大学) Baidu VIS(百度视觉部) Beihang University(北航) Australian National University(澳大利亚国立大学)

Comments Accepted by CVPR 2024; The camera-ready version

详情

展开后加载摘要…

URL PDF HTML 收藏
2411.10501 2025-10-31 cs.CV cs.LG

OnlyFlow: Optical Flow based Motion Conditioning for Video Diffusion Models

Mathis Koroglu, Hugo Caselles-Dupré, Guillaume Jeanneret Sanmiguel, Matthieu Cord

机构 * Obvious Research ISIR - Sorbonne University(ISIR - 索邦大学)

Comments 8 pages, 1 supplementary page, 9 figures

Journal ref Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition (CVPR) Workshops, 2025, pp. 6225-6235

详情

展开后加载摘要…

URL PDF HTML 收藏
1412.8070 2025-10-30 cs.CV

Functional correspondence by matrix completion

Artiom Kovnatsky, Michael M. Bronstein, Xavier Bresson, Pierre Vandergheynst

Comments "Functional Correspondence by Matrix Completion" (CVPR 2015): This paper, presented at one of the world's top AI conferences, is almost entirely fabricated, and its results are not reproducible

详情

展开后加载摘要…

URL PDF HTML 收藏
2410.09706 2025-10-29 eess.IV

ECVC: Exploiting Non-Local Correlations in Multiple Frames for Contextual Video Compression

Wei Jiang, Junru Li, Kai Zhang, Li Zhang

Comments Accepted to CVPR 2025

Journal ref Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition (CVPR), pp. 7331-7341, 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2504.01128 2025-10-29 cs.CV cs.AI

RipVIS: Rip Currents Video Instance Segmentation Benchmark for Beach Monitoring and Safety

Andrei Dumitriu, Florin Tatui, Florin Miron, Aakash Ralhan, Radu Tudor Ionescu, Radu Timofte

机构 * Computer Vision Lab, CAIDAS & IFI, University of Würzburg, Germany(计算机视觉实验室,CAIDAS与IFI,乌尔姆大学,德国) University of Bucharest, Romania(布加勒斯特大学,罗马尼亚)

Comments Accepted at CVPR 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2506.14265 2025-10-28 cs.CV

Self-supervised Representation Learning with Local Aggregation for Image-based Profiling

Siran Dai, Qianqian Xu, Peisong Wen, Yang Liu, Qingming Huang

机构 * Institute of Information Engineering, Chinese Academy of Sciences(中国科学院信息工程研究所) School of Cyber Security, University of Chinese Academy of Sciences(中国科学院大学网络空间安全学院) Institute of Computing Technology, Chinese Academy of Sciences(中国科学院计算技术研究所) School of Computer Science and Technology, University of Chinese Academy of Sciences(中国科学院大学计算机科学与技术学院)

Comments CVPR 2025 Computer Vision for Drug Discovery

详情

展开后加载摘要…

URL PDF HTML 收藏
2002.11949 2025-10-28 cs.CV cs.LG

Unbiased Scene Graph Generation from Biased Training

Kaihua Tang, Yulei Niu, Jianqiang Huang, Jiaxin Shi, Hanwang Zhang

机构 * Nanyang Technological University(南洋理工大学) Damo Academy, Alibaba Group(阿里达摩院) Renmin University of China(中国人民大学) Tsinghua University(清华大学)

Comments This paper is accepted by CVPR 2020. The code is publicly available on GitHub: https://github.com/KaihuaTang/Scene-Graph-Benchmark.pytorch

详情

展开后加载摘要…

URL PDF HTML 收藏
2506.09237 2025-10-27 cs.CV cs.LG

PatchGuard: Adversarially Robust Anomaly Detection and Localization through Vision Transformers and Pseudo Anomalies

Mojtaba Nafez, Amirhossein Koochakian, Arad Maleki, Jafar Habibi, Mohammad Hossein Rohban

机构 * Sharif University of Technology(谢里夫理工大学)

Comments Accepted to the Conference on Computer Vision and Pattern Recognition (CVPR) 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2303.08085 2025-10-27 cs.CV eess.IV

Alias-Free Convnets: Fractional Shift Invariance via Polynomial Activations

Hagay Michaeli, Tomer Michaeli, Daniel Soudry

Comments The paper was accepted to CVPR 2023. Our code is available at https://github.com/hmichaeli/alias_free_convnets/

详情

展开后加载摘要…

URL PDF HTML 收藏
2506.03885 2025-10-24 cs.CV

Video, How Do Your Tokens Merge?

Sam Pollard, Michael Wray

机构 * University of Bristol(布里斯托大学)

Comments Accepted at eLVM workshop at CVPR 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.18582 2025-10-23 cs.CV

The Photographer Eye: Teaching Multimodal Large Language Models to Understand Image Aesthetics like Photographers

Daiqing Qi, Handong Zhao, Jing Shi, Simon Jenni, Yifei Fan, Franck Dernoncourt, Scott Cohen, Sheng Li

机构 * University of Virginia(弗吉尼亚大学) Adobe(Adobe公司)

Journal ref CVPR 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2504.07744 2025-10-23 cs.CV

MMLA: Multi-Environment, Multi-Species, Low-Altitude Drone Dataset

Jenna Kline, Samuel Stevens, Guy Maalouf, Camille Rondeau Saint-Jean, Dat Nguyen Ngoc, Majid Mirmehdi, David Guerin, Tilo Burghardt, Elzbieta Pastucha, Blair Costelloe, Matthew Watson, Thomas Richardson, Ulrik Pagh Schultz Lundquist

Comments Accepted at CVPR Workshop, CV4Animals 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2412.01822 2025-10-23 cs.CV

VLsI: Verbalized Layers-to-Interactions from Large to Small Vision Language Models

Byung-Kwan Lee, Ryo Hachiuma, Yu-Chiang Frank Wang, Yong Man Ro, Yueh-Hua Wu

机构 * NVIDIA KAIST(韩国科学技术院) National Taiwan University(国立台湾大学)

Comments CVPR 2025, Project page: https://byungkwanlee.github.io/VLsI-page/

详情

展开后加载摘要…

URL PDF HTML 收藏
2505.09306 2025-10-22 cs.CV cs.LG

Predicting butterfly species presence from satellite imagery using soft contrastive regularisation

Thijs L van der Plas, Stephen Law, Michael JO Pocock

机构 * The Alan Turing Institute, the UK(英国阿尔安·图灵研究所) Wageningen University & Research, the NL(瓦赫宁根大学与研究中心) University College London, the UK(伦敦大学学院) UK Centre for Ecology & Hydrology, the UK(英国生态与水文研究中心)

Comments To be published in the 2025 CVPR FGVC12 workshop

Journal ref CVPR FGVC12 workshop (2025) pp. 2174-2183

详情

展开后加载摘要…

URL PDF HTML 收藏
2507.08217 2025-10-21 cs.AI

Quantum Federated Learning for Multimodal Data: A Modality-Agnostic Approach

Atit Pokharel, Ratun Rahman, Thomas Morris, Dinh C. Nguyen

机构 * Department of Electrical and Computer Engineering, The University of Alabama in Huntsville(电气与计算机工程系,阿拉巴马大学亨茨维尔分校)

Comments This paper was presented at BEAM with CVPR 2025

Journal ref Proceedings of the Computer Vision and Pattern Recognition Conference, pp. 545-554. 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2504.18256 2025-10-21 cs.CV

SSL4Eco: A Global Seasonal Dataset for Geospatial Foundation Models in Ecology

Elena Plekhanova, Damien Robert, Johannes Dollinger, Emilia Arens, Philipp Brun, Jan Dirk Wegner, Niklaus Zimmermann

Comments CVPR 2025, EarthVision workshop

详情

展开后加载摘要…

URL PDF HTML 收藏
2504.09697 2025-10-20 cs.GR cs.CV cs.LG

SPICE: A Synergistic, Precise, Iterative, and Customizable Image Editing Workflow

Kenan Tang, Yanhong Li, Yao Qin

机构 * University of California, Santa Barbara(加州大学圣芭芭拉分校)

Comments The paper has been accepted to NeurIPS Creative AI Track 2025. Figure 4(c) has been accepted to CVPR AI Art Gallery 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2502.08636 2025-10-20 cs.CV cs.AI

Spatial457: A Diagnostic Benchmark for 6D Spatial Reasoning of Large Multimodal Models

Xingrui Wang, Wufei Ma, Tiezheng Zhang, Celso M de Melo, Jieneng Chen, Alan Yuille

机构 * Johns Hopkins University(约翰霍普金斯大学) DEVCOM Army Research Laboratory(陆军研究实验室)

Comments Published in CVPR 2025 as Highlight. Data and code are released at https://github.com/XingruiWang/Spatial457

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.13670 2025-10-16 cs.CV

NTIRE 2025 Challenge on Low Light Image Enhancement: Methods and Results

Xiaoning Liu, Zongwei Wu, Florin-Alexandru Vasluianu, Hailong Yan, Bin Ren, Yulun Zhang, Shuhang Gu, Le Zhang, Ce Zhu, Radu Timofte, Kangbiao Shi, Yixu Feng, Tao Hu, Yu Cao, Peng Wu, Yijin Liang, Yanning Zhang, Qingsen Yan, Han Zhou, Wei Dong, Yan Min, Mohab Kishawy, Jun Chen, Pengpeng Yu, Anjin Park, Seung-Soo Lee, Young-Joon Park, Zixiao Hu, Junyv Liu, Huilin Zhang, Jun Zhang, Fei Wan, Bingxin Xu, Hongzhe Liu, Cheng Xu, Weiguo Pan, Songyin Dai, Xunpeng Yi, Qinglong Yan, Yibing Zhang, Jiayi Ma, Changhui Hu, Kerui Hu, Donghang Jing, Tiesheng Chen, Zhi Jin, Hongjun Wu, Biao Huang, Haitao Ling, Jiahao Wu, Dandan Zhan, G Gyaneshwar Rao, Vijayalaxmi Ashok Aralikatti, Nikhil Akalwadi, Ramesh Ashok Tabib, Uma Mudenagudi, Ruirui Lin, Guoxi Huang, Nantheera Anantrasirichai, Qirui Yang, Alexandru Brateanu, Ciprian Orhei, Cosmin Ancuti, Daniel Feijoo, Juan C. Benito, Álvaro García, Marcos V. Conde, Yang Qin, Raul Balmez, Anas M. Ali, Bilel Benjdira, Wadii Boulila, Tianyi Mao, Huan Zheng, Yanyan Wei, Shengeng Tang, Dan Guo, Zhao Zhang, Sabari Nathan, K Uma, A Sasithradevi, B Sathya Bama, S. Mohamed Mansoor Roomi, Ao Li, Xiangtao Zhang, Zhe Liu, Yijie Tang, Jialong Tang, Zhicheng Fu, Gong Chen, Joe Nasti, John Nicholson, Zeyu Xiao, Zhuoyuan Li, Ashutosh Kulkarni, Prashant W. Patil, Santosh Kumar Vipparthi, Subrahmanyam Murala, Duan Liu, Weile Li, Hangyuan Lu, Rixian Liu, Tengfeng Wang, Jinxing Liang, Chenxin Yu

机构 * NTIRE 2025 Challenge on Low Light Image Enhancement: Methods and Results(NTIRE 2025 挑战赛低光照图像增强:方法与结果)

Comments CVPR NTIRE 2025 Workshop, please refer to CVPR2025_workshops/NTIRE" target="_blank" rel="noopener">https://openaccess.thecvf.com/CVPR2025_workshops/NTIRE

详情

展开后加载摘要…

URL PDF HTML 收藏
2412.13443 2025-10-15 cs.CV eess.IV

DarkIR: Robust Low-Light Image Restoration

Daniel Feijoo, Juan C. Benito, Alvaro Garcia, Marcos V. Conde

机构 * Cidaut AI, Spain(Cidaut AI) Computer Vision Lab, University of Würzburg(乌尔姆大学计算机视觉实验室)

Comments CVPR 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2211.13969 2025-10-15 cs.CV cs.RO

Unsupervised Continual Semantic Adaptation through Neural Rendering

Zhizheng Liu, Francesco Milano, Jonas Frey, Roland Siegwart, Hermann Blum, Cesar Cadena

Comments Accepted by the IEEE/CVF Conference on Computer Vision and Pattern Recognition (CVPR) 2023. Zhizheng Liu and Francesco Milano share first authorship. Hermann Blum and Cesar Cadena share senior authorship. 18 pages, 8 figures, 9 tables

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.11204 2025-10-14 cs.CV

Class Prototypes based Contrastive Learning for Classifying Multi-Label and Fine-Grained Educational Videos

Rohit Gupta, Anirban Roy, Claire Christensen, Sujeong Kim, Sarah Gerard, Madeline Cincebeaux, Ajay Divakaran, Todd Grindal, Mubarak Shah

机构 * Center for Research in Computer Vision, University of Central Florida(计算机视觉研究中心,中央佛罗里达大学) SRI International(SRI国际)

Comments Published at CVPR 2023

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.10011 2025-10-14 cs.CV

MIMO: A medical vision language model with visual referring multimodal input and pixel grounding multimodal output

Yanyuan Chen, Dexuan Xu, Yu Huang, Songkun Zhan, Hanpin Wang, Dongxue Chen, Xueping Wang, Meikang Qiu, Hang Li

机构 * School of Software & Microelectronics, Peking University(软件与微电子学院,北京大学) School of Computer Science, Peking University(计算机学院,北京大学) National Engineering Research Center for Software Engineering, Peking University(软件工程国家工程研究中心,北京大学) Peking University Sixth Hospital(北京大学第六医院) Augusta University(奥古斯塔大学) Peking University First Hospital(北京大学第一医院)

Comments CVPR 2025

详情

展开后加载摘要…

URL PDF HTML 收藏