arXivDaily arXiv每日学术速递 周一至周五更新

期刊&会议

International Conference on Computer Vision · 会议 · Computer Vision

共收录 4770
2503.07235 2025-08-04 cs.CV

Retinex-MEF: Retinex-based Glare Effects Aware Unsupervised Multi-Exposure Image Fusion

Haowen Bai, Jiangshe Zhang, Zixiang Zhao, Lilun Deng, Yukun Cui, Shuang Xu

机构 * Xi’an Jiaotong University(西安交通大学) ETH Zürich(苏黎世联邦理工学院) Northwestern Polytechnical University(西北工业大学)

Comments Accepted to ICCV 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2503.06101 2025-08-04 cs.LG cs.AI

ULTHO: Ultra-Lightweight yet Efficient Hyperparameter Optimization in Deep Reinforcement Learning

Mingqi Yuan, Bo Li, Xin Jin, Wenjun Zeng

机构 * Department of Computing, The Hong Kong Polytechnic University(香港理工大学计算机系) Ningbo Institute of Digital Twin, Eastern Institute of Technology(宁波数字孪生研究所) Ningbo Key Laboratory of Spatial Intelligence and Digital Derivative(宁波空间智能与数字衍生关键实验室)

Comments 24 pages, 25 figures

Journal ref 2025 IEEE/CVF International Conference on Computer Vision (ICCV)

详情

展开后加载摘要…

URL PDF HTML 收藏
2412.17812 2025-08-04 cs.CV cs.GR

FaceLift: Learning Generalizable Single Image 3D Face Reconstruction from Synthetic Heads

Weijie Lyu, Yi Zhou, Ming-Hsuan Yang, Zhixin Shu

机构 * University of California, Merced(加州大学梅尔塞德斯分校) Adobe Research(Adobe研究)

Comments ICCV 2025 Camera-Ready Version. Project Page: https://weijielyu.github.io/FaceLift

详情

展开后加载摘要…

URL PDF HTML 收藏
2412.13180 2025-08-04 cs.CV

Feather the Throttle: Revisiting Visual Token Pruning for Vision-Language Model Acceleration

Mark Endo, Xiaohan Wang, Serena Yeung-Levy

机构 * Stanford University(斯坦福大学)

Comments ICCV 2025, project page: https://web.stanford.edu/~markendo/projects/feather

详情

展开后加载摘要…

URL PDF HTML 收藏
2412.05101 2025-08-04 cs.CV

The Silent Assistant: NoiseQuery as Implicit Guidance for Goal-Driven Image Generation

Ruoyu Wang, Huayang Huang, Ye Zhu, Olga Russakovsky, Yu Wu

机构 * School of Computer Science, Wuhan University(武汉大学计算机学院) Department of Computer Science, Princeton University(普林斯顿大学计算机科学系)

Comments ICCV 2025 Highlight

详情

展开后加载摘要…

URL PDF HTML 收藏
2412.02837 2025-08-04 cs.CV

$\texttt{BATCLIP}$: Bimodal Online Test-Time Adaptation for CLIP

Sarthak Kumar Maharana, Baoming Zhang, Leonid Karlinsky, Rogerio Feris, Yunhui Guo

机构 * The University of Texas at Dallas(德克萨斯大学达拉斯分校) MIT-IBM Watson AI Lab(麻省理工-IBM沃森人工智能实验室)

Comments ICCV 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2411.13587 2025-08-04 cs.RO cs.AI

Exploring the Adversarial Vulnerabilities of Vision-Language-Action Models in Robotics

Taowen Wang, Cheng Han, James Chenhao Liang, Wenhao Yang, Dongfang Liu, Luna Xinyu Zhang, Qifan Wang, Jiebo Luo, Ruixiang Tang

机构 * Rochester Institute of Technology(罗切斯特技术研究所) University of Missouri - Kansas City(密苏里大学-凯撒城分校) U.S. Naval Research Laboratory(美国海军研究实验室) Lamar University(拉马尔大学) Meta AI University of Rochester(罗切斯特大学) Rutgers University(新泽西罗格斯大学)

Comments ICCV camera ready; Github: https://github.com/William-wAng618/roboticAttack Homepage: https://vlaattacker.github.io/

详情

展开后加载摘要…

URL PDF HTML 收藏
2411.10086 2025-08-04 cs.CV

CorrCLIP: Reconstructing Patch Correlations in CLIP for Open-Vocabulary Semantic Segmentation

Dengke Zhang, Fagui Liu, Quan Tang

机构 * South China University of Technology(南方科技大学) Pengcheng Laboratory(鹏城实验室)

Comments Accepted to ICCV 2025 Oral

详情

展开后加载摘要…

URL PDF HTML 收藏
2406.08451 2025-08-04 cs.CV

GUIOdyssey: A Comprehensive Dataset for Cross-App GUI Navigation on Mobile Devices

Quanfeng Lu, Wenqi Shao, Zitao Liu, Lingxiao Du, Fanqing Meng, Boxuan Li, Botong Chen, Siyuan Huang, Kaipeng Zhang, Ping Luo

机构 * Shanghai AI Laboratory(上海人工智能实验室) The University of Hong Kong(香港大学) Nanjing University(南京大学) Shanghai Jiao Tong University(上海交通大学) Harbin Institute of Technology, Shenzhen(哈尔滨工业大学(深圳))

Comments 22 pages, 14 figures, ICCV 2025, a cross-app GUI navigation dataset

详情

展开后加载摘要…

URL PDF HTML 收藏
2507.23785 2025-08-01 cs.CV

Gaussian Variation Field Diffusion for High-fidelity Video-to-4D Synthesis

Bowen Zhang, Sicheng Xu, Chuxin Wang, Jiaolong Yang, Feng Zhao, Dong Chen, Baining Guo

机构 * University of Science and Technology of China(中国科学技术大学) Microsoft Research Asia(微软亚洲研究院)

Comments ICCV 2025. Project page: https://gvfdiffusion.github.io/

详情

展开后加载摘要…

URL PDF HTML 收藏
2507.23784 2025-08-01 cs.CV cs.AI cs.LG

SUB: Benchmarking CBM Generalization via Synthetic Attribute Substitutions

Jessica Bader, Leander Girrbach, Stephan Alaniz, Zeynep Akata

机构 * Technical University of Munich, Helmholtz Munich, Munich Center for Machine Learning (MCML)(慕尼黑技术大学、亥姆霍兹慕尼黑、慕尼黑机器学习中心) LTCI, Télécom Paris, Institut Polytechnique de Paris(LTCI、巴黎电信、巴黎理工学院)

Comments Accepted at ICCV 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2507.23771 2025-08-01 cs.LG cs.AI cs.CV

Consensus-Driven Active Model Selection

Justin Kay, Grant Van Horn, Subhransu Maji, Daniel Sheldon, Sara Beery

机构 * MIT(麻省理工学院) UMass Amherst(马萨诸塞大学阿默斯特分校)

Comments ICCV 2025 Highlight. 16 pages, 8 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2507.23734 2025-08-01 cs.CV cs.RO

RAGNet: Large-scale Reasoning-based Affordance Segmentation Benchmark towards General Grasping

Dongming Wu, Yanping Fu, Saike Huang, Yingfei Liu, Fan Jia, Nian Liu, Feng Dai, Tiancai Wang, Rao Muhammad Anwer, Fahad Shahbaz Khan, Jianbing Shen

机构 * The Chinese University of Hong Kong(香港中文大学) Institute of Computing Technology, Chinese Academy of Sciences(中国科学院计算技术研究所) Dexmal Mohamed bin Zayed University of Artificial Intelligence(穆罕默德·本·扎耶德人工智能大学) SKL-IOTSC, CIS, University of Macau(澳门科学馆-物联网与智能系统研究中心,澳门大学)

Comments Accepted by ICCV 2025. The code is at https://github.com/wudongming97/AffordanceNet

详情

展开后加载摘要…

URL PDF HTML 收藏
2507.23715 2025-08-01 cs.CV

DiffuMatch: Category-Agnostic Spectral Diffusion Priors for Robust Non-rigid Shape Matching

Emery Pierson, Lei Li, Angela Dai, Maks Ovsjanikov

机构 * LIX, Ecole Polytechnique(巴黎政治学院LIX实验室) Technical University of Munich(慕尼黑技术大学)

Comments Presented at ICCV 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2507.23509 2025-08-01 cs.CV cs.AI

I Am Big, You Are Little; I Am Right, You Are Wrong

David A. Kelly, Akchunya Chanchal, Nathan Blake

机构 * King’s College London(伦敦国王学院)

Comments 10 pages, International Conference on Computer Vision, ICCV 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2507.23483 2025-08-01 cs.CV

Stable-Sim2Real: Exploring Simulation of Real-Captured 3D Data with Two-Stage Depth Diffusion

Mutian Xu, Chongjie Ye, Haolin Liu, Yushuang Wu, Jiahao Chang, Xiaoguang Han

机构 * SSE, CUHKSZ(香港科技大学信息学院) FNii-Shenzhen(深圳FNii) Guangdong Provincial Key Laboratory of Future Networks of Intelligence, CUHKSZ(广东省未来网络智能重点实验室) Tencent Hunyuan3D(腾讯 Hunyuan3D) ByteDance Games(字节跳动游戏)

Comments ICCV 2025 (Highlight). Project page: https://mutianxu.github.io/stable-sim2real/

详情

展开后加载摘要…

URL PDF HTML 收藏
2507.23480 2025-08-01 cs.CV

FastPoint: Accelerating 3D Point Cloud Model Inference via Sample Point Distance Prediction

Donghyun Lee, Dawoon Jeong, Jae W. Lee, Hongil Yoon

机构 * Seoul National University(首尔国立大学) Google(谷歌)

Comments Accepted to ICCV 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2507.01631 2025-08-01 cs.CV cs.AI cs.GR cs.LG

Tile and Slide : A New Framework for Scaling NeRF from Local to Global 3D Earth Observation

Camille Billouard, Dawa Derksen, Alexandre Constantin, Bruno Vallet

机构 * CNES(法国国家太空研究中心) Univ Gustave Eiffel(巴黎高等工程师学院) IGN(法国国家地理信息研究所)

Comments Accepted at ICCV 2025 Workshop 3D-VAST (From street to space: 3D Vision Across Altitudes). Our code will be made public after the conference at https://github.com/Ellimac0/Snake-NeRF

详情

展开后加载摘要…

URL PDF HTML 收藏
2503.19480 2025-08-01 cs.CV

GenHancer: Imperfect Generative Models are Secretly Strong Vision-Centric Enhancers

Shijie Ma, Yuying Ge, Teng Wang, Yuxin Guo, Yixiao Ge, Ying Shan

机构 * ARC Lab, Tencent PCG(腾讯PCG ARC实验室) Institute of Automation, CAS(中国科学院自动化研究所)

Comments ICCV 2025. Project released at: https://mashijie1028.github.io/GenHancer/

详情

展开后加载摘要…

URL PDF HTML 收藏
2503.17856 2025-08-01 cs.CV

ClaraVid: A Holistic Scene Reconstruction Benchmark From Aerial Perspective With Delentropy-Based Complexity Profiling

Radu Beche, Sergiu Nedevschi

机构 * Technical University of Cluj-Napoca(克卢日-纳波卡技术大学)

Comments Accepted ICCV 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2503.15621 2025-08-01 cs.CV cs.AI cs.CL cs.MM

LLaVA-MORE: A Comparative Study of LLMs and Visual Backbones for Enhanced Visual Instruction Tuning

Federico Cocchi, Nicholas Moratelli, Davide Caffagni, Sara Sarto, Lorenzo Baraldi, Marcella Cornia, Rita Cucchiara

机构 * University of Modena and Reggio Emilia(摩德纳和雷吉奥艾米利亚大学) University of Pisa(比萨大学) IIT-CNR(意大利国家研究 council(IIT))

Comments ICCV 2025 Workshop on What is Next in Multimodal Foundation Models

详情

展开后加载摘要…

URL PDF HTML 收藏
2409.16178 2025-08-01 cs.CV

SDFit: 3D Object Pose and Shape by Fitting a Morphable SDF to a Single Image

Dimitrije Antić, Georgios Paschalidis, Shashank Tripathi, Theo Gevers, Sai Kumar Dwivedi, Dimitrios Tzionas

机构 * University of Amsterdam(阿姆斯特丹大学) Max Planck Institute for Intelligent Systems(智能系统马克斯·普朗克研究所)

Comments ICCV'25 Camera Ready; 12 pages, 11 figures, 5 tables

详情

展开后加载摘要…

URL PDF HTML 收藏
2405.05846 2025-08-01 cs.CR cs.CV

An Inversion-based Measure of Memorization for Diffusion Models

Zhe Ma, Qingming Li, Xuhong Zhang, Tianyu Du, Ruixiao Lin, Zonghui Wang, Shouling Ji, Wenzhi Chen

机构 * Zhejiang University(浙江大学)

Comments Accepted by ICCV 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2507.23374 2025-08-01 cs.CV

NeRF Is a Valuable Assistant for 3D Gaussian Splatting

Shuangkang Fang, I-Chao Shen, Takeo Igarashi, Yufeng Wang, ZeSheng Wang, Yi Yang, Wenrui Ding, Shuchang Zhou

机构 * Beihang University(北京航空航天大学) The University of Tokyo(东京大学) StepFun

Comments Accepted by ICCV

详情

展开后加载摘要…

URL PDF HTML 收藏
2507.23162 2025-08-01 cs.CV

Neural Multi-View Self-Calibrated Photometric Stereo without Photometric Stereo Cues

Xu Cao, Takafumi Taketomi

机构 * CyberAgent, Japan(日本CyberAgent公司)

Comments Accepted to ICCV 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2507.23134 2025-08-01 cs.CV

Details Matter for Indoor Open-vocabulary 3D Instance Segmentation

Sanghun Jung, Jingjing Zheng, Ke Zhang, Nan Qiao, Albert Y. C. Chen, Lu Xia, Chi Liu, Yuyin Sun, Xiao Zeng, Hsiang-Wei Huang, Byron Boots, Min Sun, Cheng-Hao Kuo

机构 * University of Washington(华盛顿大学) Amazon Lab126(亚马逊实验室126) National Tsing Hua University(国立清华大学)

Comments ICCV 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2507.23021 2025-08-01 cs.CV cs.AI

Modeling Human Gaze Behavior with Diffusion Models for Unified Scanpath Prediction

Giuseppe Cartella, Vittorio Cuculo, Alessandro D'Amelio, Marcella Cornia, Giuseppe Boccignone, Rita Cucchiara

机构 * University of Modena and Reggio Emilia(摩德纳和雷吉奥艾米利亚大学) University of Milan(米兰大学)

Comments Proceedings of the IEEE/CVF International Conference on Computer Vision (ICCV), 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2507.22886 2025-08-01 cs.CV

Towards Omnimodal Expressions and Reasoning in Referring Audio-Visual Segmentation

Kaining Ying, Henghui Ding, Guangquan Jie, Yu-Gang Jiang

机构 * Fudan University, China(复旦大学)

Comments ICCV 2025, Project Page: https://henghuiding.com/OmniAVS/

详情

展开后加载摘要…

URL PDF HTML 收藏
2505.24329 2025-08-01 cs.CV

DisTime: Distribution-based Time Representation for Video Large Language Models

Yingsen Zeng, Zepeng Huang, Yujie Zhong, Chengjian Feng, Jie Hu, Lin Ma, Yang Liu

机构 * Meituan Inc.(美团公司) Wangxuan Institute of Computer Technology, Peking University(北京大学王轩计算机技术研究所)

Comments Accepted by ICCV 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2505.02178 2025-08-01 cs.CV

Sparfels: Fast Reconstruction from Sparse Unposed Imagery

Shubhendu Jena, Amine Ouasfi, Mae Younes, Adnane Boukhayma

机构 * Inria(法国国家信息与自动化技术研究院) Univ. Rennes(雷恩大学) CNRS(法国国家科学研究中心) IRISA(信息科学与自动化研究院)

Comments ICCV 2025. Project page : https://shubhendu-jena.github.io/Sparfels-web/

详情

展开后加载摘要…

URL PDF HTML 收藏