arXivDaily arXiv每日学术速递 周一至周五更新
arXiv周末暂无论文更新,休息一下吧,周末愉快~~

期刊&会议

International Journal of Computer Vision · 期刊 · Computer Vision

至 收录 556
2508.08004 2025-08-12 cs.CV

Sample-aware RandAugment: Search-free Automatic Data Augmentation for Effective Image Recognition

Anqi Xiao, Weichen Yu, Hongyuan Yu

机构 * University of Chinese Academy of Sciences(中国科学院大学) Carnegie Mellon University(卡内基梅隆大学) Xiaomi Inc.(小米公司)

Comments International Journal of Computer Vision, 2025

URL PDF HTML 收藏
2403.11735 2025-08-12 cs.CV cs.LG

LSKNet: A Foundation Lightweight Backbone for Remote Sensing

Yuxuan Li, Xiang Li, Yimian Dai, Qibin Hou, Li Liu, Yongxiang Liu, Ming-Ming Cheng, Jian Yang

机构 * Nankai University(南开大学) Academy of Advanced Technology Research of Hunan(湖南先进技术研究院) NKIARI

Comments Accepted at IJCV 2024. Project page: (https://github.com/zcablii/LSKNet)[https://github.com/zcablii/LSKNet]. arXiv admin note: substantial text overlap with arXiv:2303.09030

URL PDF HTML 收藏
2408.02123 2025-08-01 cs.CV cs.LG

FovEx: Human-Inspired Explanations for Vision Transformers and Convolutional Neural Networks

Mahadev Prasad Panda, Matteo Tiezzi, Martina Vilas, Gemma Roig, Bjoern M. Eskofier, Dario Zanca

机构 * Department AIBE, FAU Erlangen-Nürnberg(FAU埃朗根-纽伦堡大学AIBE部门) PAVIS, Istituto Italiano di Tecnologia (IIT), Genova, Italy(意大利技术研究院(IIT)帕维斯部门,热那亚,意大利) Ernst Strüngmann Institute for Neuroscience, Frankfurt, Germany(神经科学埃朗根-纽伦堡研究所,法兰克福,德国) Goethe-Universität Frankfurt am Main(法兰克福大学) Institute of AI for Health, Helmholtz Zentrum München, Munich, Germany(健康人工智能研究所,海德堡中心慕尼黑,德国)

Comments Accepted in the International Journal of Computer Vision (Springer Nature)

Journal ref Int J Comput Vis (2025)

URL PDF HTML 收藏
2407.09094 2025-07-31 eess.IV cs.CV

Beyond Image Prior: Embedding Noise Prior into Conditional Denoising Transformer

Yuanfei Huang, Hua Huang

Comments Accepted by International Journal of Computer Vision (IJCV)

URL PDF HTML 收藏
2410.07336 2025-07-31 cs.CV cs.AI cs.CL cs.MM

Positive-Augmented Contrastive Learning for Vision-and-Language Evaluation and Training

Sara Sarto, Nicholas Moratelli, Marcella Cornia, Lorenzo Baraldi, Rita Cucchiara

机构 * University of Modena and Reggio Emilia(摩德纳和雷吉奥艾米利亚大学) IIT-CNR(IIT-卡内基梅隆研究院)

Comments International Journal of Computer Vision (2025)

URL PDF HTML 收藏
2507.20548 2025-07-29 cs.CV

Annotation-Free Human Sketch Quality Assessment

Lan Yang, Kaiyue Pang, Honggang Zhang, Yi-Zhe Song

Comments Accepted by IJCV

URL PDF HTML 收藏
2408.06687 2025-07-11 cs.CV cs.AI cs.LG

Masked Image Modeling: A Survey

Vlad Hondru, Florinel Alin Croitoru, Shervin Minaee, Radu Tudor Ionescu, Nicu Sebe

Comments Accepted at the International Journal of Computer Vision

URL PDF HTML 收藏
2406.04345 2025-07-08 cs.CV

Active Stereo in the Wild through Virtual Pattern Projection

Luca Bartolomei, Matteo Poggi, Fabio Tosi, Andrea Conti, Stefano Mattoccia

Comments IJCV extended version of ICCV 2023 paper: "Active Stereo Without Pattern Projector"

URL PDF HTML 收藏
2412.11074 2025-07-04 cs.CV cs.LG

Adapter-Enhanced Semantic Prompting for Continual Learning

Baocai Yin, Ji Zhao, Huajie Jiang, Ningning Hou, Yongli Hu, Amin Beheshti, Ming-Hsuan Yang, Yuankai Qi

机构 * Beijing Key Laboratory of Multimedia and Intelligent Software Technology(北京多媒体与智能软件技术重点实验室) Faculty of Information Technology, Beijing University of Technology(信息科技学院,北京理工大学) University of California at Merced(加州大学默塞德分校) Macquarie University(麦考瑞大学)

Comments This work has been submitted to the IJCV for possible publication

URL PDF HTML 收藏
2503.22359 2025-07-01 cs.CV

Mitigating Knowledge Discrepancies among Multiple Datasets for Task-agnostic Unified Face Alignment

Jiahao Xia, Min Xu, Wenjian Huang, Jianguo Zhang, Haimin Zhang, Chunxia Xiao

Comments 24 Pages, 9 Figures, accepted to IJCV-2025

URL PDF HTML 收藏
2409.17792 2025-06-30 cs.CV

Reblurring-Guided Single Image Defocus Deblurring: A Learning Framework with Misaligned Training Pairs

Dongwei Ren, Xinya Shu, Yu Li, Xiaohe Wu, Jin Li, Wangmeng Zuo

机构 * College of Intelligence and Computing, Tianjin University(天津大学智能与计算学院) Faculty of Computing, Harbin Institute of Technology(哈尔滨工业大学计算机学院)

Comments Accepted to International Journal of Computer Vision. The source code and dataset are available at https://github.com/ssscrystal/Reblurring-guided-JDRL

URL PDF HTML 收藏
2506.20381 2025-06-26 cs.CV cs.LG

Exploiting Lightweight Hierarchical ViT and Dynamic Framework for Efficient Visual Tracking

Ben Kang, Xin Chen, Jie Zhao, Chunjuan Bo, Dong Wang, Huchuan Lu

机构 * Dalian University of Technology(大连理工大学) Dalian Minzu University(大连民族大学)

Comments This paper was accepted by International Journal of Computer Vision(IJCV)

URL PDF HTML 收藏
2506.20342 2025-06-26 cs.CV cs.AI cs.LG

Feature Hallucination for Self-supervised Action Recognition

Lei Wang, Piotr Koniusz

机构 * Griffith University(格里菲斯大学) Data61/CSIRO(Data61/澳大利亚联邦科学与工业研究组织) University of New South Wales(新南威尔士大学)

Comments Accepted for publication in International Journal of Computer Vision (IJCV)

URL PDF HTML 收藏
2307.08526 2025-06-24 cs.CV cs.AI cs.LG

Image Captions are Natural Prompts for Text-to-Image Models

Shiye Lei, Hao Chen, Sen Zhang, Bo Zhao, Dacheng Tao

Comments 31 pages, 2 figure, 15 tables. Codes are available at https://github.com/LeavesLei/Caption_in_Prompt

Journal ref International Journal of Computer Vision (June 2025)

URL PDF HTML 收藏
2301.01147 2025-06-23 cs.CV

4Seasons: Benchmarking Visual SLAM and Long-Term Localization for Autonomous Driving in Challenging Conditions

Patrick Wenzel, Nan Yang, Rui Wang, Niclas Zeller, Daniel Cremers

机构 * Technical University of Munich(慕尼黑技术大学) Reality Labs at Meta(Meta现实实验室) Microsoft Mixed Reality & AI Lab(微软混合现实与人工智能实验室) Karlsruhe University of Applied Sciences(卡尔斯鲁厄应用科学大学)

Comments Published in International Journal of Computer Vision (IJCV). arXiv admin note: substantial text overlap with arXiv:2009.06364

URL PDF HTML 收藏
2408.04223 2025-06-17 cs.CV cs.AI

VideoQA in the Era of LLMs: An Empirical Study

Junbin Xiao, Nanxin Huang, Hangyu Qin, Dongyang Li, Yicong Li, Fengbin Zhu, Zhulin Tao, Jianxing Yu, Liang Lin, Tat-Seng Chua, Angela Yao

机构 * National University of Singapore(新加坡国立大学) Communication University of China(中国传媒大学) Sun Yat-Sen University(中山大学)

Comments IJCV'25

URL PDF HTML 收藏
2506.09954 2025-06-12 cs.CV cs.AI

Vision Generalist Model: A Survey

Ziyi Wang, Yongming Rao, Shuofeng Sun, Xinrun Liu, Yi Wei, Xumin Yu, Zuyan Liu, Yanbo Wang, Hongmin Liu, Jie Zhou, Jiwen Lu

机构 * Department of Automation, Tsinghua University(自动化系,清华大学) Tencent HunyuanX(腾讯混元) Beijing University of Posts and Telecommunications(北京邮电大学) University of Science and Technology Beijing(北京科技大学)

Comments Accepted by International Journal of Computer Vision (IJCV)

URL PDF HTML 收藏
2506.04609 2025-06-06 cs.LG cs.CV

Exploring bidirectional bounds for minimax-training of Energy-based models

Cong Geng, Jia Wang, Li Chen, Zhiyong Gao, Jes Frellsen, Søren Hauberg

机构 * China Mobile Research Institute(中国移动研究院) Institute of Image Communication and Network Engineering(图像通信与网络工程研究所) Department of Applied Mathematics and Computer Science(应用数学与计算机科学系)

Comments accepted to IJCV

Journal ref International Journal of Computer Vision (2025): 1-22

URL PDF HTML 收藏
2407.19889 2025-06-06 cs.CV

Self-Supervised Learning for Text Recognition: A Critical Survey

Carlos Penarrubia, Jose J. Valero-Mas, Jorge Calvo-Zaragoza

Comments Published at International Journal of Computer Vision (IJCV)

URL PDF HTML 收藏
2409.00304 2025-06-04 cs.CV

StimuVAR: Spatiotemporal Stimuli-aware Video Affective Reasoning with Multimodal Large Language Models

Yuxiang Guo, Faizan Siddiqui, Yang Zhao, Rama Chellappa, Shao-Yuan Lo

机构 * Johns Hopkins University(约翰霍普金斯大学) Honda Research Institute USA(本田研究院美国)

Comments Paper is accepted by IJCV

URL PDF HTML 收藏
2505.16976 2025-05-23 cs.CV cs.MM

Creatively Upscaling Images with Global-Regional Priors

Yurui Qian, Qi Cai, Yingwei Pan, Ting Yao, Tao Mei

Comments International Journal of Computer Vision (IJCV) 2025

URL PDF HTML 收藏
2108.08532 2025-05-22 cs.CV

An Information Theory-inspired Strategy for Automatic Network Pruning

Xiawu Zheng, Yuexiao Ma, Teng Xi, Gang Zhang, Errui Ding, Yuchao Li, Jie Chen, Yonghong Tian, Rongrong Ji

Comments Accepted by IJCV

URL PDF HTML 收藏
2402.18134 2025-05-19 cs.CV

Learning to Deblur Polarized Images

Chu Zhou, Minggui Teng, Xinyu Zhou, Chao Xu, Imari Sato, Boxin Shi

机构 * National Institute of Informatics, Japan(日本信息机构国家研究所) National Engineering Research Center of Visual Technology, School of Computer Science, Peking University, China(视觉技术国家工程研究中心,北京大学计算机学院,中国) State Key Laboratory for Multimedia Information Processing, School of Computer Science, Peking University, China(多媒体信息处理国家重点实验室,北京大学计算机学院,中国) National Key Laboratory of General AI, School of Intelligence Science and Technology, Peking University, China(通用人工智能国家重点实验室,北京大学智能科学与技术学院,中国)

Comments This version has been accepted for publication in IJCV. This arXiv version corresponds to the final accepted manuscript

URL PDF HTML 收藏
2311.14435 2025-05-16 cs.CV cs.AI

Local Concept Embeddings for Analysis of Concept Distributions in Vision DNN Feature Spaces

Georgii Mikriukov, Gesina Schwalbe, Korinna Bade

机构 * Hochschule Anhalt(安哈尔特大学) Continental AG(Continental公司) University of Lübeck(吕贝克大学)

Comments This is the authors accepted manuscript of the article accepted for publication in the International Journal of Computer Vision (IJCV). The final version will be available via SpringerLink upon publication. To cite this work please refer to the final journal version once published

URL PDF HTML 收藏
2411.15106 2025-05-07 cs.CV cs.AI cs.LG

About Time: Advances, Challenges, and Outlooks of Action Understanding

Alexandros Stergiou, Ronald Poppe

机构 * University of Twente(特文特大学) Utrecht University(乌特雷赫大学)

Comments Accepted at the International Journal of Computer Vision (IJCV)

URL PDF HTML 收藏
2311.14284 2025-05-07 cs.CV

Paragraph-to-Image Generation with Information-Enriched Diffusion Model

Weijia Wu, Zhuang Li, Yefei He, Mike Zheng Shou, Chunhua Shen, Lele Cheng, Yan Li, Tingting Gao, Di Zhang

机构 * Kuaishou Technology(快手科技) Zhejiang University(浙江大学) Show Lab, National University of Singapore(新加坡国立大学Show实验室)

Comments The project website is at: https://weijiawu.github.io/ParaDiffusionPage/. Code: https://github.com/weijiawu/ParaDiffusion

Journal ref IJCV 2025

URL PDF HTML 收藏
2504.19244 2025-05-07 cs.CV

Semantic-Aligned Learning with Collaborative Refinement for Unsupervised VI-ReID

De Cheng, Lingfeng He, Nannan Wang, Dingwen Zhang, Xinbo Gao

机构 * Xidian University(西电大学) Northwestern Polytechnical University(西北工业大学) Chongqing University of Posts and Telecommunications(重庆邮电大学)

Comments Accepted by IJCV 2025

URL PDF HTML 收藏
2404.14671 2025-04-29 cs.CV

LaneCorrect: Self-supervised Lane Detection

Ming Nie, Xinyue Cai, Hang Xu, Li Zhang

机构 * School of Data Science, Fudan University(复旦大学数据科学学院) Huawei Noah’s Ark Lab(华为诺亚实验室)

Comments IJCV 2025

URL PDF HTML 收藏
2407.05238 2025-04-24 cs.CV

P2P: Part-to-Part Motion Cues Guide a Strong Tracking Framework for LiDAR Point Clouds

Jiahao Nie, Fei Xie, Sifan Zhou, Xueyi Zhou, Dong-Kyu Chae, Zhiwei He

机构 * Hangzhou Dianzi University(杭州电子科技大学) Shanghai Jiao Tong University(上海交通大学) Carnegie Mellon University(卡内基梅隆大学) Hanyang University(翰阳大学)

Comments Accept by IJCV

URL PDF HTML 收藏
2304.11603 2025-04-21 cs.CV

LaMD: Latent Motion Diffusion for Image-Conditional Video Generation

Yaosi Hu, Zhenzhong Chen, Chong Luo

Comments accepted by IJCV

URL PDF HTML 收藏