arXivDaily arXiv每日学术速递 周一至周五更新

期刊&会议

International Conference on Computer Vision · 会议 · Computer Vision

共收录 4772
2507.04631 2025-07-08 cs.CV cs.AI cs.RO

Learning Robust Stereo Matching in the Wild with Selective Mixture-of-Experts

Yun Wang, Longguang Wang, Chenghao Zhang, Yongjian Zhang, Zhanjie Zhang, Ao Ma, Chenyou Fan, Tin Lun Lam, Junjie Hu

机构 * City University of Hong Kong(香港城市大学) Shenzhen Campus, Sun Yat-sen University(中山大学深圳校区) Chinese Academy of Sciences(中国科学院) Zhejiang University(浙江大学) South China Normal University(华南师范大学) The Chinese University of Hong Kong, Shenzhen(香港中文大学(深圳))

Journal ref ICCV 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2507.04547 2025-07-08 eess.IV cs.CV

FB-Diff: Fourier Basis-guided Diffusion for Temporal Interpolation of 4D Medical Imaging

Xin You, Runze Yang, Chuyan Zhang, Zhongliang Jiang, Jie Yang, Nassir Navab

机构 * Computer Aided Medical Procedures, Technical University of Munich(技术大学慕尼黑计算机辅助医学程序) Institute of Medical Robotics, Shanghai Jiao Tong University(上海交通大学医学机器人研究所) Munich Center for Machine Learning, Munich(慕尼黑机器学习中心) School of Computing, Macquarie University(麦考瑞大学计算机学院)

Comments Accepted by ICCV 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2507.04408 2025-07-08 cs.CV

A View-consistent Sampling Method for Regularized Training of Neural Radiance Fields

Aoxiang Fan, Corentin Dumery, Nicolas Talabot, Pascal Fua

机构 * Computer Vision Laboratory, EPFL, Switzerland(瑞士联邦理工学院计算机视觉实验室)

Comments ICCV 2025 accepted

详情

展开后加载摘要…

URL PDF HTML 收藏
2507.04403 2025-07-08 cs.CV

Sat2City: 3D City Generation from A Single Satellite Image with Cascaded Latent Diffusion

Tongyan Hua, Lutao Jiang, Ying-Cong Chen, Wufan Zhao

机构 * HKUST(GZ)(香港科技大学(广州))

Comments ICCV 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2507.04369 2025-07-08 cs.CV

MambaFusion: Height-Fidelity Dense Global Fusion for Multi-modal 3D Object Detection

Hanshi Wang, Jin Gao, Weiming Hu, Zhipeng Zhang

机构 * State Key Laboratory of Multimodal Artificial Intelligence Systems (MAIS), CASIA(多模态人工智能系统国家重点实验室(MAIS),中国科学院自动化所) School of Artificial Intelligence, University of Chinese Academy of Sciences(中国科学院大学人工智能学院) School of Artificial Intelligence, Shanghai Jiao Tong University(上海交通大学人工智能学院) Anyverse Intelligence Beijing Key Laboratory of Super Intelligent Security of Multi-Modal Information(北京多模态信息超级智能安全重点实验室) School of Information Science and Technology, ShanghaiTech University(上海科技大学信息科学与技术学院)

Comments 10 pages

Journal ref ICCV 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2507.04302 2025-07-08 cs.CV cs.LG

Adversarial Data Augmentation for Single Domain Generalization via Lyapunov Exponent-Guided Optimization

Zuyu Zhang, Ning Chen, Yongshan Liu, Qinghua Zhang, Xu Zhang

机构 * Chongqing University of Posts and Telecommunications(重庆邮电大学)

Comments Accepted to ICCV 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2507.04051 2025-07-08 cs.CV

Generate, Refine, and Encode: Leveraging Synthesized Novel Samples for On-the-Fly Fine-Grained Category Discovery

Xiao Liu, Nan Pu, Haiyang Zheng, Wenjing Li, Nicu Sebe, Zhun Zhong

机构 * Hefei University of Technology(合肥工业大学) University of Trento(特伦托大学) Helmholtz AI(海德堡人工智能研究所)

Comments ICCV 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2507.04006 2025-07-08 cs.CV

Group-wise Scaling and Orthogonal Decomposition for Domain-Invariant Feature Extraction in Face Anti-Spoofing

Seungjin Jung, Kanghee Lee, Yonghyun Jeong, Haeun Noh, Jungmin Lee, Jongwon Choi

机构 * Chung-Ang University(Chung-Ang 大学) Naver Cloud

Comments Published at ICCV 2025. code is will be available at https://github.com/SeungjinJung/GD-FAS

详情

展开后加载摘要…

URL PDF HTML 收藏
2507.03730 2025-07-08 cs.CV cs.AI cs.HC cs.LG

Less is More: Empowering GUI Agent with Context-Aware Simplification

Gongwei Chen, Xurui Zhou, Rui Shao, Yibo Lyu, Kaiwen Zhou, Shuai Wang, Wentao Li, Yinchuan Li, Zhongang Qi, Liqiang Nie

机构 * Harbin Institute of Technology, Shenzhen(哈尔滨工业大学(深圳)) Huawei Noah’s Ark Lab(华为诺亚实验室)

Comments Accepted to ICCV 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2507.02664 2025-07-08 cs.CV

AIGI-Holmes: Towards Explainable and Generalizable AI-Generated Image Detection via Multimodal Large Language Models

Ziyin Zhou, Yunpeng Luo, Yuanchen Wu, Ke Sun, Jiayi Ji, Ke Yan, Shouhong Ding, Xiaoshuai Sun, Yunsheng Wu, Rongrong Ji

机构 * Key Laboratory of Multimedia Trusted Perception and Efficient Computing, Ministry of Education of China, Xiamen University(中国教育部多媒体可信感知与高效计算重点实验室,厦门大学) Tencent YouTu Lab(腾讯YouTu实验室)

Comments Accepted to ICCV 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2506.23537 2025-07-08 eess.IV cs.CV

AFUNet: Cross-Iterative Alignment-Fusion Synergy for HDR Reconstruction via Deep Unfolding Paradigm

Xinyue Li, Zhangkai Ni, Wenhan Yang

机构 * Tongji University(同济大学) Pengcheng Laboratory(鹏城实验室)

Comments Accepted to International Conference on Computer Vision (ICCV) 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2506.05344 2025-07-08 cs.CV

SparseMM: Head Sparsity Emerges from Visual Concept Responses in MLLMs

Jiahui Wang, Zuyan Liu, Yongming Rao, Jiwen Lu

机构 * Tsinghua University(清华大学) Tencent Hunyuan X(腾讯混元实验室)

Comments Accepted to ICCV 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2507.03578 2025-07-08 cs.CV cs.AI cs.LG

SciVid: Cross-Domain Evaluation of Video Models in Scientific Applications

Yana Hasson, Pauline Luc, Liliane Momeni, Maks Ovsjanikov, Guillaume Le Moing, Alina Kuznetsova, Ira Ktena, Jennifer J. Sun, Skanda Koppula, Dilara Gokay, Joseph Heyward, Etienne Pot, Andrew Zisserman

机构 * Google(谷歌) DeepMind(深度思维)

Comments ICCV 2025, GitHub repo: https://github.com/google-deepmind/scivid

详情

展开后加载摘要…

URL PDF HTML 收藏
2507.03434 2025-07-08 cs.CV cs.MM

Unlearning the Noisy Correspondence Makes CLIP More Robust

Haochen Han, Alex Jinpeng Wang, Peijun Ye, Fangming Liu

机构 * Peng Cheng Laboratory(鹏城实验室) Central South University(中南大学)

Comments ICCV 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2507.03306 2025-07-08 cs.CV

MGSfM: Multi-Camera Geometry Driven Global Structure-from-Motion

Peilin Tao, Hainan Cui, Diantao Tu, Shuhan Shen

Comments Accepted at ICCV 2025, The code is available at https://github.com/3dv-casia/MGSfM/

详情

展开后加载摘要…

URL PDF HTML 收藏
2507.03304 2025-07-08 cs.CV

Bridging Domain Generalization to Multimodal Domain Generalization via Unified Representations

Hai Huang, Yan Xia, Sashuai Zhou, Hanting Wang, Shulei Wang, Zhou Zhao

机构 * Zhejiang University(浙江大学)

Comments Accepted by ICCV 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2507.03292 2025-07-08 cs.CV

Zero-shot Inexact CAD Model Alignment from a Single Image

Pattaramanee Arsomngern, Sasikarn Khwanmuang, Matthias Nießner, Supasorn Suwajanakorn

机构 * VISTEC, Thailand(泰国VISTEC) Technical University of Munich, Germany(德国慕尼黑技术大学)

Comments ICCV 2025. Project page: https://zerocad9d.github.io/

详情

展开后加载摘要…

URL PDF HTML 收藏
2507.02929 2025-07-08 cs.CV cs.AI cs.LG stat.ML

OBSER: Object-Based Sub-Environment Recognition for Zero-Shot Environmental Inference

Won-Seok Choi, Dong-Sig Han, Suhyung Choi, Hyeonseo Yang, Byoung-Tak Zhang

机构 * Seoul National University(首尔国立大学)

Comments This manuscript was initially submitted to ICCV 2025 and is now made available as a preprint

详情

展开后加载摘要…

URL PDF HTML 收藏
2507.01417 2025-07-08 cs.CV cs.LG

Gradient Short-Circuit: Efficient Out-of-Distribution Detection via Feature Intervention

Jiawei Gu, Ziyue Qiao, Zechao Li

Comments Accepted to ICCV 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2507.00505 2025-07-08 cs.CV

LLaVA-SP: Enhancing Visual Representation with Visual Spatial Tokens for MLLMs

Haoran Lou, Chunxiao Fan, Ziyan Liu, Yuexin Wu, Xinliang Wang

机构 * Beijing University of Posts and Telecommunications(北京邮电大学) Beihang University(北航)

Comments Accepted to ICCV 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2411.15235 2025-07-08 cs.LG cs.AI cs.CV cs.NE

CODE-CL: Conceptor-Based Gradient Projection for Deep Continual Learning

Marco Paul E. Apolinario, Sakshi Choudhary, Kaushik Roy

Comments Accepted to the IEEE/CVF International Conference on Computer Vision (ICCV) 2025, 12 pages, 5 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2507.02857 2025-07-04 cs.CV

AnyI2V: Animating Any Conditional Image with Motion Control

Ziye Li, Hao Luo, Xincheng Shuai, Henghui Ding

机构 * Fudan University(复旦大学) DAMO Academy, Alibaba group(阿里达摩院) Hupan Lab(虎扑实验室)

Comments ICCV 2025, Project Page: https://henghuiding.com/AnyI2V/

详情

展开后加载摘要…

URL PDF HTML 收藏
2507.02748 2025-07-04 cs.CV cs.AI cs.LG

Linear Attention with Global Context: A Multipole Attention Mechanism for Vision and Physics

Alex Colagrande, Paul Caillon, Eva Feillet, Alexandre Allauzen

机构 * Miles Team, LAMSADE, Université Paris Dauphine-PSL(巴黎南大学LAMSADE研究所)

Comments Accepted at ECLR Workshop at ICCV 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2507.02714 2025-07-04 cs.CV cs.AI

FairHuman: Boosting Hand and Face Quality in Human Image Generation with Minimum Potential Delay Fairness in Diffusion Models

Yuxuan Wang, Tianwei Cao, Huayu Zhang, Zhongjiang He, Kongming Liang, Zhanyu Ma

机构 * Beijing University of Posts and Telecommunications(北京邮电大学) Institute of Artificial Intelligence, China Telecom(中国电信人工智能研究院)

Comments ICCV 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2507.02691 2025-07-04 cs.CV

CanonSwap: High-Fidelity and Consistent Video Face Swapping via Canonical Space Modulation

Xiangyang Luo, Ye Zhu, Yunfei Liu, Lijian Lin, Cong Wan, Zijian Cai, Shao-Lun Huang, Yu Li

机构 * Tsinghua Shenzhen International Graduate School, Tsinghua University(清华大学深圳国际研究生院,清华大学) International Digital Economy Academy (IDEA)(国际数字经济学院) Xi’an Jiaotong University(西安交通大学)

Comments ICCV Accepted

详情

展开后加载摘要…

URL PDF HTML 收藏
2507.02363 2025-07-04 cs.CV

LocalDyGS: Multi-view Global Dynamic Scene Modeling via Adaptive Local Implicit Feature Decoupling

Jiahao Wu, Rui Peng, Jianbo Jiao, Jiayu Yang, Luyang Tang, Kaiqiang Xiong, Jie Liang, Jinbo Yan, Runling Liu, Ronggang Wang

机构 * Guangdong Provincial Key Laboratory of Ultra High Definition Immersive Media Technology Shenzhen Graduate School, Peking University(广东省超高清沉浸媒体技术重点实验室 香港大学深圳研究生院) Pengcheng Lab(鹏城实验室) University of Birmingham(伯明翰大学)

Comments Accepted by ICCV 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2503.03259 2025-07-04 cs.CV

BANet: Bilateral Aggregation Network for Mobile Stereo Matching

Gangwei Xu, Jiaxin Liu, Xianqi Wang, Junda Cheng, Yong Deng, Jinliang Zang, Yurui Chen, Xin Yang

机构 * Huazhong University of Science and Technology(华中科技大学) Autel Robotics(Autel机器人公司) Optics Valley Laboratory(光学谷实验室)

Comments Accepted by ICCV 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2410.16236 2025-07-04 cs.CV

LLaVA-KD: A Framework of Distilling Multimodal Large Language Models

Yuxuan Cai, Jiangning Zhang, Haoyang He, Xinwei He, Ao Tong, Zhenye Gan, Chengjie Wang, Zhucun Xue, Yong Liu, Xiang Bai

机构 * Huazhong University of Science and Technology(华中科技大学) Zhejiang University(浙江大学) Youtu Lab, Tencent(腾讯优图实验室) Huazhong Agricultural University(华中农业大学)

Comments ICCV'25

详情

展开后加载摘要…

URL PDF HTML 收藏
2507.01953 2025-07-03 cs.CV

FreeMorph: Tuning-Free Generalized Image Morphing with Diffusion Model

Yukang Cao, Chenyang Si, Jinghao Wang, Ziwei Liu

机构 * S-Lab, Nanyang Technological University(南洋理工大学S实验室) Nanjing University(南京大学) The Chinese University of Hong Kong(香港中文大学)

Comments ICCV 2025. Project page: https://yukangcao.github.io/FreeMorph/

详情

展开后加载摘要…

URL PDF HTML 收藏
2507.01838 2025-07-03 cs.CV

MobileIE: An Extremely Lightweight and Effective ConvNet for Real-Time Image Enhancement on Mobile Devices

Hailong Yan, Ao Li, Xiangtao Zhang, Zhe Liu, Zenglin Shi, Ce Zhu, Le Zhang

机构 * UESTC

Comments Accepted by ICCV 2025

详情

展开后加载摘要…

URL PDF HTML 收藏