arXivDaily arXiv每日学术速递 周一至周五更新

期刊&会议

International Conference on Computer Vision · 会议 · Computer Vision

2025-08-01 至 2025-08-01 共收录 28
2507.23785 2025-08-01 cs.CV

Gaussian Variation Field Diffusion for High-fidelity Video-to-4D Synthesis

Bowen Zhang, Sicheng Xu, Chuxin Wang, Jiaolong Yang, Feng Zhao, Dong Chen, Baining Guo

机构 * University of Science and Technology of China(中国科学技术大学) Microsoft Research Asia(微软亚洲研究院)

Comments ICCV 2025. Project page: https://gvfdiffusion.github.io/

详情

展开后加载摘要…

URL PDF HTML 收藏
2507.23784 2025-08-01 cs.CV cs.AI cs.LG

SUB: Benchmarking CBM Generalization via Synthetic Attribute Substitutions

Jessica Bader, Leander Girrbach, Stephan Alaniz, Zeynep Akata

机构 * Technical University of Munich, Helmholtz Munich, Munich Center for Machine Learning (MCML)(慕尼黑技术大学、亥姆霍兹慕尼黑、慕尼黑机器学习中心) LTCI, Télécom Paris, Institut Polytechnique de Paris(LTCI、巴黎电信、巴黎理工学院)

Comments Accepted at ICCV 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2507.23771 2025-08-01 cs.LG cs.AI cs.CV

Consensus-Driven Active Model Selection

Justin Kay, Grant Van Horn, Subhransu Maji, Daniel Sheldon, Sara Beery

机构 * MIT(麻省理工学院) UMass Amherst(马萨诸塞大学阿默斯特分校)

Comments ICCV 2025 Highlight. 16 pages, 8 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2507.23734 2025-08-01 cs.CV cs.RO

RAGNet: Large-scale Reasoning-based Affordance Segmentation Benchmark towards General Grasping

Dongming Wu, Yanping Fu, Saike Huang, Yingfei Liu, Fan Jia, Nian Liu, Feng Dai, Tiancai Wang, Rao Muhammad Anwer, Fahad Shahbaz Khan, Jianbing Shen

机构 * The Chinese University of Hong Kong(香港中文大学) Institute of Computing Technology, Chinese Academy of Sciences(中国科学院计算技术研究所) Dexmal Mohamed bin Zayed University of Artificial Intelligence(穆罕默德·本·扎耶德人工智能大学) SKL-IOTSC, CIS, University of Macau(澳门科学馆-物联网与智能系统研究中心,澳门大学)

Comments Accepted by ICCV 2025. The code is at https://github.com/wudongming97/AffordanceNet

详情

展开后加载摘要…

URL PDF HTML 收藏
2507.23715 2025-08-01 cs.CV

DiffuMatch: Category-Agnostic Spectral Diffusion Priors for Robust Non-rigid Shape Matching

Emery Pierson, Lei Li, Angela Dai, Maks Ovsjanikov

机构 * LIX, Ecole Polytechnique(巴黎政治学院LIX实验室) Technical University of Munich(慕尼黑技术大学)

Comments Presented at ICCV 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2507.23509 2025-08-01 cs.CV cs.AI

I Am Big, You Are Little; I Am Right, You Are Wrong

David A. Kelly, Akchunya Chanchal, Nathan Blake

机构 * King’s College London(伦敦国王学院)

Comments 10 pages, International Conference on Computer Vision, ICCV 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2507.23483 2025-08-01 cs.CV

Stable-Sim2Real: Exploring Simulation of Real-Captured 3D Data with Two-Stage Depth Diffusion

Mutian Xu, Chongjie Ye, Haolin Liu, Yushuang Wu, Jiahao Chang, Xiaoguang Han

机构 * SSE, CUHKSZ(香港科技大学信息学院) FNii-Shenzhen(深圳FNii) Guangdong Provincial Key Laboratory of Future Networks of Intelligence, CUHKSZ(广东省未来网络智能重点实验室) Tencent Hunyuan3D(腾讯 Hunyuan3D) ByteDance Games(字节跳动游戏)

Comments ICCV 2025 (Highlight). Project page: https://mutianxu.github.io/stable-sim2real/

详情

展开后加载摘要…

URL PDF HTML 收藏
2507.23480 2025-08-01 cs.CV

FastPoint: Accelerating 3D Point Cloud Model Inference via Sample Point Distance Prediction

Donghyun Lee, Dawoon Jeong, Jae W. Lee, Hongil Yoon

机构 * Seoul National University(首尔国立大学) Google(谷歌)

Comments Accepted to ICCV 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2507.01631 2025-08-01 cs.CV cs.AI cs.GR cs.LG

Tile and Slide : A New Framework for Scaling NeRF from Local to Global 3D Earth Observation

Camille Billouard, Dawa Derksen, Alexandre Constantin, Bruno Vallet

机构 * CNES(法国国家太空研究中心) Univ Gustave Eiffel(巴黎高等工程师学院) IGN(法国国家地理信息研究所)

Comments Accepted at ICCV 2025 Workshop 3D-VAST (From street to space: 3D Vision Across Altitudes). Our code will be made public after the conference at https://github.com/Ellimac0/Snake-NeRF

详情

展开后加载摘要…

URL PDF HTML 收藏
2503.19480 2025-08-01 cs.CV

GenHancer: Imperfect Generative Models are Secretly Strong Vision-Centric Enhancers

Shijie Ma, Yuying Ge, Teng Wang, Yuxin Guo, Yixiao Ge, Ying Shan

机构 * ARC Lab, Tencent PCG(腾讯PCG ARC实验室) Institute of Automation, CAS(中国科学院自动化研究所)

Comments ICCV 2025. Project released at: https://mashijie1028.github.io/GenHancer/

详情

展开后加载摘要…

URL PDF HTML 收藏
2503.17856 2025-08-01 cs.CV

ClaraVid: A Holistic Scene Reconstruction Benchmark From Aerial Perspective With Delentropy-Based Complexity Profiling

Radu Beche, Sergiu Nedevschi

机构 * Technical University of Cluj-Napoca(克卢日-纳波卡技术大学)

Comments Accepted ICCV 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2503.15621 2025-08-01 cs.CV cs.AI cs.CL cs.MM

LLaVA-MORE: A Comparative Study of LLMs and Visual Backbones for Enhanced Visual Instruction Tuning

Federico Cocchi, Nicholas Moratelli, Davide Caffagni, Sara Sarto, Lorenzo Baraldi, Marcella Cornia, Rita Cucchiara

机构 * University of Modena and Reggio Emilia(摩德纳和雷吉奥艾米利亚大学) University of Pisa(比萨大学) IIT-CNR(意大利国家研究 council(IIT))

Comments ICCV 2025 Workshop on What is Next in Multimodal Foundation Models

详情

展开后加载摘要…

URL PDF HTML 收藏
2409.16178 2025-08-01 cs.CV

SDFit: 3D Object Pose and Shape by Fitting a Morphable SDF to a Single Image

Dimitrije Antić, Georgios Paschalidis, Shashank Tripathi, Theo Gevers, Sai Kumar Dwivedi, Dimitrios Tzionas

机构 * University of Amsterdam(阿姆斯特丹大学) Max Planck Institute for Intelligent Systems(智能系统马克斯·普朗克研究所)

Comments ICCV'25 Camera Ready; 12 pages, 11 figures, 5 tables

详情

展开后加载摘要…

URL PDF HTML 收藏
2405.05846 2025-08-01 cs.CR cs.CV

An Inversion-based Measure of Memorization for Diffusion Models

Zhe Ma, Qingming Li, Xuhong Zhang, Tianyu Du, Ruixiao Lin, Zonghui Wang, Shouling Ji, Wenzhi Chen

机构 * Zhejiang University(浙江大学)

Comments Accepted by ICCV 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2507.23374 2025-08-01 cs.CV

NeRF Is a Valuable Assistant for 3D Gaussian Splatting

Shuangkang Fang, I-Chao Shen, Takeo Igarashi, Yufeng Wang, ZeSheng Wang, Yi Yang, Wenrui Ding, Shuchang Zhou

机构 * Beihang University(北京航空航天大学) The University of Tokyo(东京大学) StepFun

Comments Accepted by ICCV

详情

展开后加载摘要…

URL PDF HTML 收藏
2507.23162 2025-08-01 cs.CV

Neural Multi-View Self-Calibrated Photometric Stereo without Photometric Stereo Cues

Xu Cao, Takafumi Taketomi

机构 * CyberAgent, Japan(日本CyberAgent公司)

Comments Accepted to ICCV 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2507.23134 2025-08-01 cs.CV

Details Matter for Indoor Open-vocabulary 3D Instance Segmentation

Sanghun Jung, Jingjing Zheng, Ke Zhang, Nan Qiao, Albert Y. C. Chen, Lu Xia, Chi Liu, Yuyin Sun, Xiao Zeng, Hsiang-Wei Huang, Byron Boots, Min Sun, Cheng-Hao Kuo

机构 * University of Washington(华盛顿大学) Amazon Lab126(亚马逊实验室126) National Tsing Hua University(国立清华大学)

Comments ICCV 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2507.23021 2025-08-01 cs.CV cs.AI

Modeling Human Gaze Behavior with Diffusion Models for Unified Scanpath Prediction

Giuseppe Cartella, Vittorio Cuculo, Alessandro D'Amelio, Marcella Cornia, Giuseppe Boccignone, Rita Cucchiara

机构 * University of Modena and Reggio Emilia(摩德纳和雷吉奥艾米利亚大学) University of Milan(米兰大学)

Comments Proceedings of the IEEE/CVF International Conference on Computer Vision (ICCV), 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2507.22886 2025-08-01 cs.CV

Towards Omnimodal Expressions and Reasoning in Referring Audio-Visual Segmentation

Kaining Ying, Henghui Ding, Guangquan Jie, Yu-Gang Jiang

机构 * Fudan University, China(复旦大学)

Comments ICCV 2025, Project Page: https://henghuiding.com/OmniAVS/

详情

展开后加载摘要…

URL PDF HTML 收藏
2505.24329 2025-08-01 cs.CV

DisTime: Distribution-based Time Representation for Video Large Language Models

Yingsen Zeng, Zepeng Huang, Yujie Zhong, Chengjian Feng, Jie Hu, Lin Ma, Yang Liu

机构 * Meituan Inc.(美团公司) Wangxuan Institute of Computer Technology, Peking University(北京大学王轩计算机技术研究所)

Comments Accepted by ICCV 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2505.02178 2025-08-01 cs.CV

Sparfels: Fast Reconstruction from Sparse Unposed Imagery

Shubhendu Jena, Amine Ouasfi, Mae Younes, Adnane Boukhayma

机构 * Inria(法国国家信息与自动化技术研究院) Univ. Rennes(雷恩大学) CNRS(法国国家科学研究中心) IRISA(信息科学与自动化研究院)

Comments ICCV 2025. Project page : https://shubhendu-jena.github.io/Sparfels-web/

详情

展开后加载摘要…

URL PDF HTML 收藏
2504.05164 2025-08-01 cs.CV

Balancing Task-invariant Interaction and Task-specific Adaptation for Unified Image Fusion

Xingyu Hu, Junjun Jiang, Chenyang Wang, Kui Jiang, Xianming Liu, Jiayi Ma

机构 * Harbin Institute of Technology(哈尔滨理工大学) Wuhan University(武汉大学)

Comments Accepted by ICCV 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2503.22351 2025-08-01 cs.CV

One Look is Enough: Seamless Patchwise Refinement for Zero-Shot Monocular Depth Estimation on High-Resolution Images

Byeongjun Kwon, Munchurl Kim

机构 * KAIST(韩国科学技术院)

Comments ICCV 2025 (camera-ready version). [Project page](https://kaist-viclab.github.io/One-Look-is-Enough_site)

详情

展开后加载摘要…

URL PDF HTML 收藏
2503.15897 2025-08-01 cs.CV cs.LG

Learning 3D Scene Analogies with Neural Contextual Scene Maps

Junho Kim, Gwangtak Bae, Eun Sun Lee, Young Min Kim

机构 * Dept. of Electrical and Computer Engineering, Seoul National University(电子与计算机工程系,首尔国立大学) Interdisciplinary Program in Artificial Intelligence and INMC, Seoul National University(人工智能跨学科项目及INMC,首尔国立大学)

Comments Accepted to ICCV 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2503.14939 2025-08-01 cs.CV

VisNumBench: Evaluating Number Sense of Multimodal Large Language Models

Tengjin Weng, Jingyi Wang, Wenhao Jiang, Zhong Ming

机构 * College of Computer Science and Software Engineering, Shenzhen University(深圳大学计算机科学与软件工程学院) Guangdong Laboratory of Artificial Intelligence and Digital Economy (SZ)(广东省人工智能与数字经济发展实验室) Shenzhen International Graduate School, Tsinghua University(清华大学深圳国际 Graduate School) Shenzhen Technology University(深圳技术大学)

Comments accepted by ICCV 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2502.20760 2025-08-01 cs.CV

VRM: Knowledge Distillation via Virtual Relation Matching

Weijia Zhang, Fei Xie, Weidong Cai, Chao Ma

机构 * Shanghai Jiao Tong University(上海交通大学) The University of Sydney(悉尼大学)

Comments Accepted by ICCV 2025 (Highlight)

详情

展开后加载摘要…

URL PDF HTML 收藏
2501.08727 2025-08-01 cs.LG

Transformed Low-rank Adaptation via Tensor Decomposition and Its Applications to Text-to-image Models

Zerui Tao, Yuhta Takida, Naoki Murata, Qibin Zhao, Yuki Mitsufuji

机构 * RIKEN AIP(RIKEN人工智能研究所) Sony AI(索尼人工智能) Sony Group Corporation(索尼集团)

Comments ICCV 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2412.06458 2025-08-01 cs.CV

Pruning All-Rounder: Rethinking and Improving Inference Efficiency for Large Vision Language Models

Wei Suo, Ji Ma, Mengyang Sun, Lin Yuanbo Wu, Peng Wang, Yanning Zhang

机构 * Northwestern Polytechnical University(西北工业大学) Swansea University(斯旺西大学)

Comments Accepted by ICCV 25

详情

展开后加载摘要…

URL PDF HTML 收藏