arXivDaily arXiv每日学术速递 周一至周五更新

期刊&会议

International Conference on Computer Vision · 会议 · Computer Vision

2025-07-31 至 2025-07-31 共收录 30
2507.22872 2025-07-31 cs.CV

TR-PTS: Task-Relevant Parameter and Token Selection for Efficient Tuning

Siqi Luo, Haoran Yang, Yi Xin, Mingyang Yi, Guangyang Wu, Guangtao Zhai, Xiaohong Liu

机构 * Shanghai Jiao Tong University(上海交通大学) Shanghai Innovation Institute(上海创新研究院) Renmin University of China(中国人民大学) Suzhou Key Laboratory of Artificial Intelligence(苏州人工智能重点实验室)

Comments Accepted by ICCV 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2507.22825 2025-07-31 cs.CV

DepR: Depth Guided Single-view Scene Reconstruction with Instance-level Diffusion

Qingcheng Zhao, Xiang Zhang, Haiyang Xu, Zeyuan Chen, Jianwen Xie, Yuan Gao, Zhuowen Tu

机构 * ShanghaiTech University(上海科技大学) UC San Diego(加州大学圣地亚哥分校) Lambda, Inc.(Lambda公司) Stanford University(斯坦福大学)

Comments ICCV 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2507.22813 2025-07-31 cs.CV

DISTIL: Data-Free Inversion of Suspicious Trojan Inputs via Latent Diffusion

Hossein Mirzaei, Zeinab Taghavi, Sepehr Rezaee, Masoud Hadi, Moein Madadi, Mackenzie W. Mathis

机构 * École Polytechnique Fédérale de Lausanne(洛桑联邦理工学院)

Comments ICCV 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2507.22692 2025-07-31 cs.CV

Zero-Shot Image Anomaly Detection Using Generative Foundation Models

Lemar Abdi, Amaan Valiuddin, Francisco Caetano, Christiaan Viviers, Fons van der Sommen

机构 * Eindhoven University of Technology(埃因霍温理工大学)

Comments Accepted at the workshop of Anomaly Detection with Foundation Models, ICCV 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2507.22617 2025-07-31 cs.CR cs.CV

Hate in Plain Sight: On the Risks of Moderating AI-Generated Hateful Illusions

Yiting Qu, Ziqing Yang, Yihan Ma, Michael Backes, Savvas Zannettou, Yang Zhang

机构 * CISPA Helmholtz Center for Information Security(CISPA 欧洲信息安全研究中心) TU Delft(代尔夫特理工大学)

Comments Accepted at ICCV 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2507.22615 2025-07-31 cs.CV

Generative Active Learning for Long-tail Trajectory Prediction via Controllable Diffusion Model

Daehee Park, Monu Surana, Pranav Desai, Ashish Mehta, Reuben MV John, Kuk-Jin Yoon

机构 * Intelligent Systems and Learning Lab., DGIST, Korea(智能系统与学习实验室,DGIST,韩国) Qualcomm Research, USA(高通研究,美国) Visual Intelligence Lab., KAIST, Korea(视觉智能实验室,KAIST,韩国)

Comments Accepted at ICCV 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2507.22604 2025-07-31 cs.CV

ShortFT: Diffusion Model Alignment via Shortcut-based Fine-Tuning

Xiefan Guo, Miaomiao Cui, Liefeng Bo, Di Huang

机构 * State Key Laboratory of Complex and Critical Software Environment(复杂与关键软件环境国家重点实验室) Beihang University(北京航空航天大学) School of Computer Science and Engineering(计算机科学与工程学院)

Comments Accepted by ICCV 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2507.22553 2025-07-31 cs.CV cs.AI cs.LG

RainbowPrompt: Diversity-Enhanced Prompt-Evolving for Continual Learning

Kiseong Hong, Gyeong-hyeon Kim, Eunwoo Kim

机构 * Department of AI(人工智能系) School of CSE(计算机科学与工程学院)

Comments Accepted by the 2025 IEEE/CVF International Conference on Computer Vision (ICCV 2025)

详情

展开后加载摘要…

URL PDF HTML 收藏
2507.22480 2025-07-31 cs.CV

Estimating 2D Camera Motion with Hybrid Motion Basis

Haipeng Li, Tianhao Zhou, Zhanglei Yang, Yi Wu, Yan Chen, Zijing Mao, Shen Cheng, Bing Zeng, Shuaicheng Liu

机构 * University of Electronic Science and Technology of China(电子科技大学)

Comments ICCV 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2507.22412 2025-07-31 cs.CV

UAVScenes: A Multi-Modal Dataset for UAVs

Sijie Wang, Siqi Li, Yawei Zhang, Shangshu Yu, Shenghai Yuan, Rui She, Quanjiang Guo, JinXuan Zheng, Ong Kang Howe, Leonrich Chandra, Shrivarshann Srijeyan, Aditya Sivadas, Toshan Aggarwal, Heyuan Liu, Hongming Zhang, Chujie Chen, Junyu Jiang, Lihua Xie, Wee Peng Tay

机构 * Nanyang Technological University(南洋理工大学) School of Computer Science and Engineering, Northeastern University(东北大学计算机科学与工程学院) Beihang University(北航大学) University of Electronic Science and Technology of China(电子科技大学)

Comments Accepted by ICCV 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2507.22404 2025-07-31 cs.CV cs.AI cs.LG

MINR: Implicit Neural Representations with Masked Image Modelling

Sua Lee, Joonhun Lee, Myungjoo Kang

机构 * Seoul National University(首尔国立大学)

Comments Accepted to the ICCV 2023 workshop on Out-of-Distribution Generalization in Computer Vision

详情

展开后加载摘要…

URL PDF HTML 收藏
2507.21391 2025-07-31 cs.CV cs.AI cs.CL

Multimodal LLMs as Customized Reward Models for Text-to-Image Generation

Shijie Zhou, Ruiyi Zhang, Huaisheng Zhu, Branislav Kveton, Yufan Zhou, Jiuxiang Gu, Jian Chen, Changyou Chen

机构 * University at Buffalo(布法罗大学) Adobe Research(Adobe研究) Pennsylvania State University(宾夕法尼亚州立大学)

Comments Accepted at ICCV 2025. Code available at https://github.com/sjz5202/LLaVA-Reward

详情

展开后加载摘要…

URL PDF HTML 收藏
2507.18082 2025-07-31 cs.CV cs.AI

TextSAM-EUS: Text Prompt Learning for SAM to Accurately Segment Pancreatic Tumor in Endoscopic Ultrasound

Pascal Spiegler, Taha Koleilat, Arash Harirpoush, Corey S. Miller, Hassan Rivaz, Marta Kersten-Oertel, Yiming Xiao

机构 * Concordia University(康科迪亚大学) Jewish General Hospital(犹太总医院) McGill University Faculty of Medicine(麦吉尔大学医学院) Lady Davis Institute for Medical Research(拉德克利夫医学研究 institute)

Comments Accepted to ICCV 2025 Workshop CVAMD

详情

展开后加载摘要…

URL PDF HTML 收藏
2507.17332 2025-07-31 cs.CV

PARTE: Part-Guided Texturing for 3D Human Reconstruction from a Single Image

Hyeongjin Nam, Donghwan Kim, Gyeongsik Moon, Kyoung Mu Lee

机构 * Dept. of ECE&ASRI, Seoul National University(电子工程与先进科学研究所,首尔国立大学) Dept. of CSE, Korea University(计算机科学与工程系,韩国大学)

Comments Published at ICCV 2025, 22 pages including the supplementary material

详情

展开后加载摘要…

URL PDF HTML 收藏
2507.04047 2025-07-31 cs.CV

Move to Understand a 3D Scene: Bridging Visual Grounding and Exploration for Efficient and Versatile Embodied Navigation

Ziyu Zhu, Xilin Wang, Yixuan Li, Zhuofan Zhang, Xiaojian Ma, Yixin Chen, Baoxiong Jia, Wei Liang, Qian Yu, Zhidong Deng, Siyuan Huang, Qing Li

机构 * Tsinghua University(清华大学) Beijing Institute of Technology(北京理工大学) Beihang University(北航) State Key Laboratory of General Artificial Intelligence, BIGAI, China(国家一般人工智能重点实验室, BIGAI, 中国)

Comments Embodied AI; 3D Vision Language Understanding; ICCV 2025 Highlight; https://mtu3d.github.io; Spatial intelligence

详情

展开后加载摘要…

URL PDF HTML 收藏
2507.01439 2025-07-31 cs.CV

TurboReg: TurboClique for Robust and Efficient Point Cloud Registration

Shaocheng Yan, Pengcheng Shi, Zhenjun Zhao, Kaixin Wang, Kuang Cao, Ji Wu, Jiayuan Li

机构 * School of Remote Sensing and Information Engineering, Wuhan University(武汉大学遥感与信息工程学院) Department of Computer and Systems Engineering, University of Zaragoza(阿拉维扎大学计算机与系统工程系) College of Computer Science, Beijing University of Technology(北京理工大学计算机学院) School of Computer Science, Wuhan University(武汉大学计算机学院)

Comments ICCV-2025 Accepted Paper

详情

展开后加载摘要…

URL PDF HTML 收藏
2506.21541 2025-07-31 cs.CV

StruMamba3D: Exploring Structural Mamba for Self-supervised Point Cloud Representation Learning

Chuxin Wang, Yixin Zha, Wenfei Yang, Tianzhu Zhang

机构 * University of Science and Technology of China(中国科学技术大学)

Comments Accepted by ICCV 2025, website: https://chuxwa.github.io/project_StruMamba3D/

详情

展开后加载摘要…

URL PDF HTML 收藏
2506.21132 2025-07-31 cs.CV

Learning to See in the Extremely Dark

Hai Jiang, Binhao Guan, Zhen Liu, Xiaohong Liu, Jian Yu, Zheng Liu, Songchen Han, Shuaicheng Liu

机构 * School of Aeronautics and Astronautics, Sichuan University(四川大学航空宇航学院) University of Electronic Science and Technology of China(电子科技大学) Shanghai Jiao Tong University(上海交通大学) National Innovation Center for UHD Video Technology(超高清视频技术国家创新中心)

Comments Accepted by ICCV 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2504.10487 2025-07-31 cs.CV cs.LG

FLOSS: Free Lunch in Open-vocabulary Semantic Segmentation

Yasser Benigmim, Mohammad Fahes, Tuan-Hung Vu, Andrei Bursuc, Raoul de Charette

机构 * Inria(法国国家信息与自动化技术研究所)

Comments ICCV 2025; Project Page: https://yasserben.github.io/FLOSS/

详情

展开后加载摘要…

URL PDF HTML 收藏
2504.06504 2025-07-31 cs.CV

STaR: Seamless Spatial-Temporal Aware Motion Retargeting with Penetration and Consistency Constraints

Xiaohang Yang, Qing Wang, Jiahao Yang, Gregory Slabaugh, Shanxin Yuan

机构 * Queen Mary University of London(伦敦大学皇家亨利学院)

Comments Accepted by ICCV 2025, 13 pages, 9 figures; Code page: https://github.com/XiaohangYang829/STaR

详情

展开后加载摘要…

URL PDF HTML 收藏
2504.02612 2025-07-31 cs.CV

Fine-Tuning Visual Autoregressive Models for Subject-Driven Generation

Jiwoo Chung, Sangeek Hyun, Hyunjun Kim, Eunseo Koh, MinKyu Lee, Jae-Pil Heo

机构 * Sungkyunkwan University(成均馆大学)

Comments Accepted to ICCV 2025. Project page: https://jiwoogit.github.io/ARBooth/

详情

展开后加载摘要…

URL PDF HTML 收藏
2503.11026 2025-07-31 eess.AS cs.CV cs.LG cs.MM

MAVFlow: Preserving Paralinguistic Elements with Conditional Flow Matching for Zero-Shot AV2AV Multilingual Translation

Sungwoo Cho, Jeongsoo Choi, Sungnyun Kim, Se-Young Yun

机构 * KAIST AI(韩国科学技术院人工智能研究所) KAIST EE(韩国科学技术院电子工程系)

Comments Accepted to ICCV 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2503.07360 2025-07-31 cs.RO

AffordDexGrasp: Open-set Language-guided Dexterous Grasp with Generalizable-Instructive Affordance

Yi-Lin Wei, Mu Lin, Yuhao Lin, Jian-Jian Jiang, Xiao-Ming Wu, Ling-An Zeng, Wei-Shi Zheng

机构 * School of Computer Science and Engineering, Sun Yat-sen University, China(中山大学计算机科学与工程学院) Key Laboratory of Machine Intelligence and Advanced Computing, Ministry of Education, China(教育部人工智能与高级计算重点实验室)

Comments Accepted by ICCV 2025.Project page: https://isee-laboratory.github.io/AffordDexGrasp/

详情

展开后加载摘要…

URL PDF HTML 收藏
2501.06838 2025-07-31 eess.IV cs.CV

Generalized and Efficient 2D Gaussian Splatting for Arbitrary-scale Super-Resolution

Du Chen, Liyi Chen, Zhengqiang Zhang, Lei Zhang

机构 * The Hong Kong Polytechnic University(香港理工大学) OPPO Research Institute(OPPO研究院)

Comments Accepted by ICCV 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2411.17765 2025-07-31 cs.CV

I2VControl: Disentangled and Unified Video Motion Synthesis Control

Wanquan Feng, Tianhao Qi, Jiawei Liu, Mingzhen Sun, Pengqi Tu, Tianxiang Ma, Fei Dai, Songtao Zhao, Siyu Zhou, Qian He

机构 * Intelligent Creation Team, ByteDance(字节跳动智能创作团队) University of Science and Technology of China (USTC)(中国科学技术大学) Institute of Automation, Chinese Academy of Sciences (CASIA)(中国科学院自动化研究所)

Comments Accepted to ICCV 2025. Project page: https://wanquanf.github.io/I2VControl

详情

展开后加载摘要…

URL PDF HTML 收藏
2411.16072 2025-07-31 cs.CV

Language Driven Occupancy Prediction

Zhu Yu, Bowen Pang, Lizhe Liu, Runmin Zhang, Qiang Li, Si-Yuan Cao, Maochun Luo, Mingxia Chen, Sheng Yang, Hui-Liang Shen

机构 * Zhejiang University(浙江大学) Unmanned Vehicle Dept., CaiNiao Inc., Alibaba Group(无人车辆部门,菜鸟公司,阿里巴巴集团) Ningbo Global Innovation Center, Zhejiang University(宁波全球创新中心,浙江大学)

Comments ICCV 2025; Project Page: https://github.com/pkqbajng/LOcc

详情

展开后加载摘要…

URL PDF HTML 收藏
2406.05400 2025-07-31 cs.CV math.DG

Metric Convolutions: A Unifying Theory to Adaptive Image Convolutions

Thomas Dagès, Michael Lindenbaum, Alfred M. Bruckstein

机构 * Technion – Israel Institute of Technology(以色列技术学院) Technical University of Munich(慕尼黑技术大学) Munich Center for Machine Learning(慕尼黑机器学习中心)

Comments Updated version, Accepted for publication at the IEEE/CVF International Conference on Computer Vision (ICCV) 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2507.22100 2025-07-31 cs.CV

Trade-offs in Image Generation: How Do Different Dimensions Interact?

Sicheng Zhang, Binzhu Xie, Zhonghao Yan, Yuli Zhang, Donghao Zhou, Xiaofei Chen, Shi Qiu, Jiaqi Liu, Guoyang Xie, Zhichao Lu

机构 * Khalifa University(卡利法大学) The Chinese University of Hong Kong(香港中文大学) Queen Mary University of London(伦敦大学玛丽女王学院) Xi’an Jiaotong-Liverpool University(西安交通大学利物浦大学) City University of Hong Kong(香港城市大学)

Comments Accepted in ICCV 2025, Codebase: https://github.com/fesvhtr/TRIG

详情

展开后加载摘要…

URL PDF HTML 收藏
2507.22076 2025-07-31 cs.LG

Test-time Prompt Refinement for Text-to-Image Models

Mohammad Abdul Hafeez Khan, Yash Jain, Siddhartha Bhattacharyya, Vibhav Vineet

机构 * Florida Institute of Technology(佛罗里达理工学院) Microsoft Research(微软研究院)

Comments Accepted to ICCV 2025, MARS2 Workshop. Total 14 pages, 12 figures and 3 tables

详情

展开后加载摘要…

URL PDF HTML 收藏
2412.20104 2025-07-31 cs.CV cs.AI cs.LG cs.RO

SyncDiff: Synchronized Motion Diffusion for Multi-Body Human-Object Interaction Synthesis

Wenkun He, Yun Liu, Ruitao Liu, Li Yi

机构 * Tsinghua University(清华大学) Shanghai Qi Zhi Institute(上海启智研究院) Shanghai Artificial Intelligence Laboratory(上海人工智能实验室)

Comments 27 pages, 10 figures, 20 tables. Accepted by ICCV 2025

详情

展开后加载摘要…

URL PDF HTML 收藏