arXivDaily arXiv每日学术速递 周一至周五更新

期刊&会议

International Conference on Computer Vision · 会议 · Computer Vision

共收录 4770
2504.05164 2025-08-01 cs.CV

Balancing Task-invariant Interaction and Task-specific Adaptation for Unified Image Fusion

Xingyu Hu, Junjun Jiang, Chenyang Wang, Kui Jiang, Xianming Liu, Jiayi Ma

机构 * Harbin Institute of Technology(哈尔滨理工大学) Wuhan University(武汉大学)

Comments Accepted by ICCV 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2503.22351 2025-08-01 cs.CV

One Look is Enough: Seamless Patchwise Refinement for Zero-Shot Monocular Depth Estimation on High-Resolution Images

Byeongjun Kwon, Munchurl Kim

机构 * KAIST(韩国科学技术院)

Comments ICCV 2025 (camera-ready version). [Project page](https://kaist-viclab.github.io/One-Look-is-Enough_site)

详情

展开后加载摘要…

URL PDF HTML 收藏
2503.15897 2025-08-01 cs.CV cs.LG

Learning 3D Scene Analogies with Neural Contextual Scene Maps

Junho Kim, Gwangtak Bae, Eun Sun Lee, Young Min Kim

机构 * Dept. of Electrical and Computer Engineering, Seoul National University(电子与计算机工程系,首尔国立大学) Interdisciplinary Program in Artificial Intelligence and INMC, Seoul National University(人工智能跨学科项目及INMC,首尔国立大学)

Comments Accepted to ICCV 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2503.14939 2025-08-01 cs.CV

VisNumBench: Evaluating Number Sense of Multimodal Large Language Models

Tengjin Weng, Jingyi Wang, Wenhao Jiang, Zhong Ming

机构 * College of Computer Science and Software Engineering, Shenzhen University(深圳大学计算机科学与软件工程学院) Guangdong Laboratory of Artificial Intelligence and Digital Economy (SZ)(广东省人工智能与数字经济发展实验室) Shenzhen International Graduate School, Tsinghua University(清华大学深圳国际 Graduate School) Shenzhen Technology University(深圳技术大学)

Comments accepted by ICCV 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2502.20760 2025-08-01 cs.CV

VRM: Knowledge Distillation via Virtual Relation Matching

Weijia Zhang, Fei Xie, Weidong Cai, Chao Ma

机构 * Shanghai Jiao Tong University(上海交通大学) The University of Sydney(悉尼大学)

Comments Accepted by ICCV 2025 (Highlight)

详情

展开后加载摘要…

URL PDF HTML 收藏
2501.08727 2025-08-01 cs.LG

Transformed Low-rank Adaptation via Tensor Decomposition and Its Applications to Text-to-image Models

Zerui Tao, Yuhta Takida, Naoki Murata, Qibin Zhao, Yuki Mitsufuji

机构 * RIKEN AIP(RIKEN人工智能研究所) Sony AI(索尼人工智能) Sony Group Corporation(索尼集团)

Comments ICCV 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2412.06458 2025-08-01 cs.CV

Pruning All-Rounder: Rethinking and Improving Inference Efficiency for Large Vision Language Models

Wei Suo, Ji Ma, Mengyang Sun, Lin Yuanbo Wu, Peng Wang, Yanning Zhang

机构 * Northwestern Polytechnical University(西北工业大学) Swansea University(斯旺西大学)

Comments Accepted by ICCV 25

详情

展开后加载摘要…

URL PDF HTML 收藏
2507.22872 2025-07-31 cs.CV

TR-PTS: Task-Relevant Parameter and Token Selection for Efficient Tuning

Siqi Luo, Haoran Yang, Yi Xin, Mingyang Yi, Guangyang Wu, Guangtao Zhai, Xiaohong Liu

机构 * Shanghai Jiao Tong University(上海交通大学) Shanghai Innovation Institute(上海创新研究院) Renmin University of China(中国人民大学) Suzhou Key Laboratory of Artificial Intelligence(苏州人工智能重点实验室)

Comments Accepted by ICCV 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2507.22825 2025-07-31 cs.CV

DepR: Depth Guided Single-view Scene Reconstruction with Instance-level Diffusion

Qingcheng Zhao, Xiang Zhang, Haiyang Xu, Zeyuan Chen, Jianwen Xie, Yuan Gao, Zhuowen Tu

机构 * ShanghaiTech University(上海科技大学) UC San Diego(加州大学圣地亚哥分校) Lambda, Inc.(Lambda公司) Stanford University(斯坦福大学)

Comments ICCV 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2507.22813 2025-07-31 cs.CV

DISTIL: Data-Free Inversion of Suspicious Trojan Inputs via Latent Diffusion

Hossein Mirzaei, Zeinab Taghavi, Sepehr Rezaee, Masoud Hadi, Moein Madadi, Mackenzie W. Mathis

机构 * École Polytechnique Fédérale de Lausanne(洛桑联邦理工学院)

Comments ICCV 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2507.22692 2025-07-31 cs.CV

Zero-Shot Image Anomaly Detection Using Generative Foundation Models

Lemar Abdi, Amaan Valiuddin, Francisco Caetano, Christiaan Viviers, Fons van der Sommen

机构 * Eindhoven University of Technology(埃因霍温理工大学)

Comments Accepted at the workshop of Anomaly Detection with Foundation Models, ICCV 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2507.22617 2025-07-31 cs.CR cs.CV

Hate in Plain Sight: On the Risks of Moderating AI-Generated Hateful Illusions

Yiting Qu, Ziqing Yang, Yihan Ma, Michael Backes, Savvas Zannettou, Yang Zhang

机构 * CISPA Helmholtz Center for Information Security(CISPA 欧洲信息安全研究中心) TU Delft(代尔夫特理工大学)

Comments Accepted at ICCV 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2507.22615 2025-07-31 cs.CV

Generative Active Learning for Long-tail Trajectory Prediction via Controllable Diffusion Model

Daehee Park, Monu Surana, Pranav Desai, Ashish Mehta, Reuben MV John, Kuk-Jin Yoon

机构 * Intelligent Systems and Learning Lab., DGIST, Korea(智能系统与学习实验室,DGIST,韩国) Qualcomm Research, USA(高通研究,美国) Visual Intelligence Lab., KAIST, Korea(视觉智能实验室,KAIST,韩国)

Comments Accepted at ICCV 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2507.22604 2025-07-31 cs.CV

ShortFT: Diffusion Model Alignment via Shortcut-based Fine-Tuning

Xiefan Guo, Miaomiao Cui, Liefeng Bo, Di Huang

机构 * State Key Laboratory of Complex and Critical Software Environment(复杂与关键软件环境国家重点实验室) Beihang University(北京航空航天大学) School of Computer Science and Engineering(计算机科学与工程学院)

Comments Accepted by ICCV 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2507.22553 2025-07-31 cs.CV cs.AI cs.LG

RainbowPrompt: Diversity-Enhanced Prompt-Evolving for Continual Learning

Kiseong Hong, Gyeong-hyeon Kim, Eunwoo Kim

机构 * Department of AI(人工智能系) School of CSE(计算机科学与工程学院)

Comments Accepted by the 2025 IEEE/CVF International Conference on Computer Vision (ICCV 2025)

详情

展开后加载摘要…

URL PDF HTML 收藏
2507.22480 2025-07-31 cs.CV

Estimating 2D Camera Motion with Hybrid Motion Basis

Haipeng Li, Tianhao Zhou, Zhanglei Yang, Yi Wu, Yan Chen, Zijing Mao, Shen Cheng, Bing Zeng, Shuaicheng Liu

机构 * University of Electronic Science and Technology of China(电子科技大学)

Comments ICCV 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2507.22412 2025-07-31 cs.CV

UAVScenes: A Multi-Modal Dataset for UAVs

Sijie Wang, Siqi Li, Yawei Zhang, Shangshu Yu, Shenghai Yuan, Rui She, Quanjiang Guo, JinXuan Zheng, Ong Kang Howe, Leonrich Chandra, Shrivarshann Srijeyan, Aditya Sivadas, Toshan Aggarwal, Heyuan Liu, Hongming Zhang, Chujie Chen, Junyu Jiang, Lihua Xie, Wee Peng Tay

机构 * Nanyang Technological University(南洋理工大学) School of Computer Science and Engineering, Northeastern University(东北大学计算机科学与工程学院) Beihang University(北航大学) University of Electronic Science and Technology of China(电子科技大学)

Comments Accepted by ICCV 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2507.22404 2025-07-31 cs.CV cs.AI cs.LG

MINR: Implicit Neural Representations with Masked Image Modelling

Sua Lee, Joonhun Lee, Myungjoo Kang

机构 * Seoul National University(首尔国立大学)

Comments Accepted to the ICCV 2023 workshop on Out-of-Distribution Generalization in Computer Vision

详情

展开后加载摘要…

URL PDF HTML 收藏
2507.21391 2025-07-31 cs.CV cs.AI cs.CL

Multimodal LLMs as Customized Reward Models for Text-to-Image Generation

Shijie Zhou, Ruiyi Zhang, Huaisheng Zhu, Branislav Kveton, Yufan Zhou, Jiuxiang Gu, Jian Chen, Changyou Chen

机构 * University at Buffalo(布法罗大学) Adobe Research(Adobe研究) Pennsylvania State University(宾夕法尼亚州立大学)

Comments Accepted at ICCV 2025. Code available at https://github.com/sjz5202/LLaVA-Reward

详情

展开后加载摘要…

URL PDF HTML 收藏
2507.18082 2025-07-31 cs.CV cs.AI

TextSAM-EUS: Text Prompt Learning for SAM to Accurately Segment Pancreatic Tumor in Endoscopic Ultrasound

Pascal Spiegler, Taha Koleilat, Arash Harirpoush, Corey S. Miller, Hassan Rivaz, Marta Kersten-Oertel, Yiming Xiao

机构 * Concordia University(康科迪亚大学) Jewish General Hospital(犹太总医院) McGill University Faculty of Medicine(麦吉尔大学医学院) Lady Davis Institute for Medical Research(拉德克利夫医学研究 institute)

Comments Accepted to ICCV 2025 Workshop CVAMD

详情

展开后加载摘要…

URL PDF HTML 收藏
2507.17332 2025-07-31 cs.CV

PARTE: Part-Guided Texturing for 3D Human Reconstruction from a Single Image

Hyeongjin Nam, Donghwan Kim, Gyeongsik Moon, Kyoung Mu Lee

机构 * Dept. of ECE&ASRI, Seoul National University(电子工程与先进科学研究所,首尔国立大学) Dept. of CSE, Korea University(计算机科学与工程系,韩国大学)

Comments Published at ICCV 2025, 22 pages including the supplementary material

详情

展开后加载摘要…

URL PDF HTML 收藏
2507.04047 2025-07-31 cs.CV

Move to Understand a 3D Scene: Bridging Visual Grounding and Exploration for Efficient and Versatile Embodied Navigation

Ziyu Zhu, Xilin Wang, Yixuan Li, Zhuofan Zhang, Xiaojian Ma, Yixin Chen, Baoxiong Jia, Wei Liang, Qian Yu, Zhidong Deng, Siyuan Huang, Qing Li

机构 * Tsinghua University(清华大学) Beijing Institute of Technology(北京理工大学) Beihang University(北航) State Key Laboratory of General Artificial Intelligence, BIGAI, China(国家一般人工智能重点实验室, BIGAI, 中国)

Comments Embodied AI; 3D Vision Language Understanding; ICCV 2025 Highlight; https://mtu3d.github.io; Spatial intelligence

详情

展开后加载摘要…

URL PDF HTML 收藏
2507.01439 2025-07-31 cs.CV

TurboReg: TurboClique for Robust and Efficient Point Cloud Registration

Shaocheng Yan, Pengcheng Shi, Zhenjun Zhao, Kaixin Wang, Kuang Cao, Ji Wu, Jiayuan Li

机构 * School of Remote Sensing and Information Engineering, Wuhan University(武汉大学遥感与信息工程学院) Department of Computer and Systems Engineering, University of Zaragoza(阿拉维扎大学计算机与系统工程系) College of Computer Science, Beijing University of Technology(北京理工大学计算机学院) School of Computer Science, Wuhan University(武汉大学计算机学院)

Comments ICCV-2025 Accepted Paper

详情

展开后加载摘要…

URL PDF HTML 收藏
2506.21541 2025-07-31 cs.CV

StruMamba3D: Exploring Structural Mamba for Self-supervised Point Cloud Representation Learning

Chuxin Wang, Yixin Zha, Wenfei Yang, Tianzhu Zhang

机构 * University of Science and Technology of China(中国科学技术大学)

Comments Accepted by ICCV 2025, website: https://chuxwa.github.io/project_StruMamba3D/

详情

展开后加载摘要…

URL PDF HTML 收藏
2506.21132 2025-07-31 cs.CV

Learning to See in the Extremely Dark

Hai Jiang, Binhao Guan, Zhen Liu, Xiaohong Liu, Jian Yu, Zheng Liu, Songchen Han, Shuaicheng Liu

机构 * School of Aeronautics and Astronautics, Sichuan University(四川大学航空宇航学院) University of Electronic Science and Technology of China(电子科技大学) Shanghai Jiao Tong University(上海交通大学) National Innovation Center for UHD Video Technology(超高清视频技术国家创新中心)

Comments Accepted by ICCV 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2504.10487 2025-07-31 cs.CV cs.LG

FLOSS: Free Lunch in Open-vocabulary Semantic Segmentation

Yasser Benigmim, Mohammad Fahes, Tuan-Hung Vu, Andrei Bursuc, Raoul de Charette

机构 * Inria(法国国家信息与自动化技术研究所)

Comments ICCV 2025; Project Page: https://yasserben.github.io/FLOSS/

详情

展开后加载摘要…

URL PDF HTML 收藏
2504.06504 2025-07-31 cs.CV

STaR: Seamless Spatial-Temporal Aware Motion Retargeting with Penetration and Consistency Constraints

Xiaohang Yang, Qing Wang, Jiahao Yang, Gregory Slabaugh, Shanxin Yuan

机构 * Queen Mary University of London(伦敦大学皇家亨利学院)

Comments Accepted by ICCV 2025, 13 pages, 9 figures; Code page: https://github.com/XiaohangYang829/STaR

详情

展开后加载摘要…

URL PDF HTML 收藏
2504.02612 2025-07-31 cs.CV

Fine-Tuning Visual Autoregressive Models for Subject-Driven Generation

Jiwoo Chung, Sangeek Hyun, Hyunjun Kim, Eunseo Koh, MinKyu Lee, Jae-Pil Heo

机构 * Sungkyunkwan University(成均馆大学)

Comments Accepted to ICCV 2025. Project page: https://jiwoogit.github.io/ARBooth/

详情

展开后加载摘要…

URL PDF HTML 收藏
2503.11026 2025-07-31 eess.AS cs.CV cs.LG cs.MM

MAVFlow: Preserving Paralinguistic Elements with Conditional Flow Matching for Zero-Shot AV2AV Multilingual Translation

Sungwoo Cho, Jeongsoo Choi, Sungnyun Kim, Se-Young Yun

机构 * KAIST AI(韩国科学技术院人工智能研究所) KAIST EE(韩国科学技术院电子工程系)

Comments Accepted to ICCV 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2503.07360 2025-07-31 cs.RO

AffordDexGrasp: Open-set Language-guided Dexterous Grasp with Generalizable-Instructive Affordance

Yi-Lin Wei, Mu Lin, Yuhao Lin, Jian-Jian Jiang, Xiao-Ming Wu, Ling-An Zeng, Wei-Shi Zheng

机构 * School of Computer Science and Engineering, Sun Yat-sen University, China(中山大学计算机科学与工程学院) Key Laboratory of Machine Intelligence and Advanced Computing, Ministry of Education, China(教育部人工智能与高级计算重点实验室)

Comments Accepted by ICCV 2025.Project page: https://isee-laboratory.github.io/AffordDexGrasp/

详情

展开后加载摘要…

URL PDF HTML 收藏