arXivDaily arXiv每日学术速递 周一至周五更新

期刊&会议

International Conference on Computer Vision · 会议 · Computer Vision

共收录 4770
2503.12095 2025-08-20 cs.CV

Towards Vision Zero: The TUM Traffic Accid3nD Dataset

Walter Zimmer, Ross Greer, Daniel Lehmberg, Marc Pavel, Holger Caesar, Xingcheng Zhou, Ahmed Ghita, Mohan Trivedi, Rui Song, Hu Cao, Akshay Gopalkrishnan, Alois C. Knoll

Comments Accepted for the IEEE/CVF International Conference on Computer Vision Workshops (ICCVW)

详情

展开后加载摘要…

URL PDF HTML 收藏
2410.02458 2025-08-20 eess.IV cs.CL cs.CV

MedVisionLlama: Leveraging Pre-Trained Large Language Model Layers to Enhance Medical Image Segmentation

Gurucharan Marthi Krishna Kumar, Aman Chadha, Janine Mendola, Amir Shmuel

机构 * Montreal Neurological Institute, McGill University(蒙特利尔神经科学研究所,麦吉尔大学)

Comments Accepted to the CVAMD Workshop (Computer Vision for Automated Medical Diagnosis) at the 2025 IEEE/CVF International Conference on Computer Vision (ICCVW 2025)

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.13104 2025-08-19 cs.CV cs.RO

Precise Action-to-Video Generation Through Visual Action Prompts

Yuang Wang, Chao Wen, Haoyu Guo, Sida Peng, Minghan Qin, Hujun Bao, Xiaowei Zhou, Ruizhen Hu

机构 * Xiangjiang Lab(湘江实验室) Zhejiang University(浙江大学) Fudan University(复旦大学) Tsinghua University(清华大学) Shenzhen University(深圳大学)

Comments Accepted to ICCV 2025. Project page: https://zju3dv.github.io/VAP/

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.13007 2025-08-19 cs.CV

SlimComm: Doppler-Guided Sparse Queries for Bandwidth-Efficient Cooperative 3-D Perception

Melih Yazgan, Qiyuan Wu, Iramm Hamdard, Shiqi Li, J. Marius Zoellner

机构 * FZI Research Center for Information Technology(FZI信息技术研究所以) Karlsruhe Institute of Technology(卡尔斯鲁厄大学)

Comments Accepted by ICCV - Drive2X Workshop

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.12813 2025-08-19 cs.CV cs.LG

SIS-Challenge: Event-based Spatio-temporal Instance Segmentation Challenge at the CVPR 2025 Event-based Vision Workshop

Friedhelm Hamann, Emil Mededovic, Fabian Gülhan, Yuli Wu, Johannes Stegmaier, Jing He, Yiqing Wang, Kexin Zhang, Lingling Li, Licheng Jiao, Mengru Ma, Hongxiang Huang, Yuhao Yan, Hongwei Ren, Xiaopeng Lin, Yulong Huang, Bojun Cheng, Se Hyun Lee, Gyu Sung Ham, Kanghan Oh, Gi Hyun Lim, Boxuan Yang, Bowen Du, Guillermo Gallego

机构 * TU Berlin, SCIoI, ECDF(柏林技术大学、SCIoI、ECDF) RWTH Aachen(亚琛工业大学) Xidian University(西安电子科技大学) Hong Kong University of Science and Technology(香港科技大学) Sun Yat-sen University(中山大学) Wonkwang University(Wonkwang大学) Tongji University(同济大学)

Comments 13 pages, 7 figures, 7 tables

Journal ref IEEE/CVF International Conference on Computer Vision (ICCV) Workshops, 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.12668 2025-08-19 cs.CV

WP-CLIP: Leveraging CLIP to Predict Wölfflin's Principles in Visual Art

Abhijay Ghildyal, Li-Yun Wang, Feng Liu

机构 * Portland State University(波特兰州立大学)

Comments ICCV 2025 AI4VA workshop (oral), Code: https://github.com/abhijay9/wpclip

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.12330 2025-08-19 cs.CV

DoppDrive: Doppler-Driven Temporal Aggregation for Improved Radar Object Detection

Yuval Haitman, Oded Bialer

机构 * General Motors(通用汽车公司) Ben Gurion University of the Negev(本· Gurion 内盖夫大学)

Comments ICCV 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.12163 2025-08-19 cs.CV cs.AI cs.HC cs.LG

RealTalk: Realistic Emotion-Aware Lifelike Talking-Head Synthesis

Wenqing Wang, Yun Fu

机构 * Northeastern University(东北大学)

Comments Accepted to the ICCV 2025 Workshop on Artificial Social Intelligence

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.12131 2025-08-19 cs.CV

DualFit: A Two-Stage Virtual Try-On via Warping and Synthesis

Minh Tran, Johnmark Clements, Annie Prasanna, Tri Nguyen, Ngan Le

机构 * University of Arkansas(亚拉巴马大学) Coupang, Inc.(Coupang公司)

Comments Retail Vision, ICCV 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.12084 2025-08-19 cs.CV cs.AI

Generic Event Boundary Detection via Denoising Diffusion

Jaejun Hwang, Dayoung Gong, Manjin Kim, Minsu Cho

机构 * Pohang University of Science and Technology (POSTECH)(釜山科学技术大学) GenGenAI

Comments Accepted to ICCV 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.10672 2025-08-19 cs.CV cs.AI

Hybrid Generative Fusion for Efficient and Privacy-Preserving Face Recognition Dataset Generation

Feiran Li, Qianqian Xu, Shilong Bao, Boyu Han, Zhiyong Yang, Qingming Huang

机构 * Institute of Information Engineering, Chinese Academy of Sciences(中国科学院信息工程研究所) School of Cyber Security, University of Chinese Academy of Sciences(中国科学院大学网络安全学院) Institute of Computing Technology, Chinese Academy of Sciences(中国科学院计算技术研究所) School of Computer Science and Technology, University of Chinese Academy of Sciences(中国科学院大学计算机科学与技术学院) BDKM, University of Chinese Academy of Sciences(中国科学院大学BDKM)

Comments This paper has been accpeted to ICCV 2025 DataCV Workshop

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.05123 2025-08-19 cs.CV cs.AI

Latent Expression Generation for Referring Image Segmentation and Grounding

Seonghoon Yu, Junbeom Hong, Joonseok Lee, Jeany Son

机构 * GIST(韩国科学技术院) Seoul National University(首尔国立大学) POSTECH

Comments Accepted to ICCV 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2506.22908 2025-08-19 cs.CV

Attention to the Burstiness in Visual Prompt Tuning!

Yuzhu Wang, Manni Duan, Shu Kong

机构 * Zhejiang Lab(浙江实验室) University of Macau(澳门大学)

Comments ICCV 2025; v2: camera ready

详情

展开后加载摘要…

URL PDF HTML 收藏
2505.14414 2025-08-19 cs.CV

Diving into the Fusion of Monocular Priors for Generalized Stereo Matching

Chengtang Yao, Lidong Yu, Zhidan Liu, Jiaxi Zeng, Yuwei Wu, Yunde Jia

机构 * Beijing Key Laboratory of Intelligent Information Technology, School of Computer Science & Technology, Beijing Institute of Technology, China(北京智能信息科技重点实验室,计算机科学与技术学院,北京理工大学,中国) Guangdong Laboratory of Machine Perception and Intelligent Computing, Shenzhen MSU-BIT University, China(广东机器感知与智能计算实验室,深圳MSU-BIT大学,中国) NVIDIA

Comments Code: https://github.com/YaoChengTang/Diving-into-the-Fusion-of-Monocular-Priors-for-Generalized-Stereo-Matching

Journal ref ICCV 2025 Oral

详情

展开后加载摘要…

URL PDF HTML 收藏
2505.04813 2025-08-19 cs.GR cs.CV

WIR3D: Visually-Informed and Geometry-Aware 3D Shape Abstraction

Richard Liu, Daniel Fu, Noah Tan, Itai Lang, Rana Hanocka

机构 * University of Chicago(芝加哥大学)

Comments ICCV 2025 Oral Project page: https://threedle.github.io/wir3d/

详情

展开后加载摘要…

URL PDF HTML 收藏
2504.21847 2025-08-19 cs.CV cs.SD

Differentiable Room Acoustic Rendering with Multi-View Vision Priors

Derong Jin, Ruohan Gao

机构 * University of Maryland, College Park(马里兰大学学院公园分校)

Comments ICCV 2025 (Oral); Project Page: https://humathe.github.io/avdar/

详情

展开后加载摘要…

URL PDF HTML 收藏
2503.16867 2025-08-19 cs.CV

ETVA: Evaluation of Text-to-Video Alignment via Fine-grained Question Generation and Answering

Kaisi Guan, Zhengfeng Lai, Yuchong Sun, Peng Zhang, Wei Liu, Kieran Liu, Meng Cao, Ruihua Song

机构 * Renmin University of China(中国人民大学) Apple(苹果公司)

Comments International Conference on Computer Vision 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2503.14482 2025-08-19 cs.CV

ICE-Bench: A Unified and Comprehensive Benchmark for Image Creating and Editing

Yulin Pan, Xiangteng He, Chaojie Mao, Zhen Han, Zeyinzi Jiang, Jingfeng Zhang, Yu Liu

Comments 17 pages. Accepted by ICCV 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2502.20511 2025-08-19 cs.CV

Best Foot Forward: Robust Foot Reconstruction in-the-wild

Kyle Fogarty, Jing Yang, Chayan Kumar Patodi, Jack Foster, Aadi Bhanti, Steven Chacko, Cengiz Oztireli, Ujwal Bonde

机构 * University of Cambridge(剑桥大学) Hike Medical(Hike医疗)

Comments ICCV 2025 Workshop on Advanced Perception for Autonomous Healthcare

详情

展开后加载摘要…

URL PDF HTML 收藏
2412.05934 2025-08-19 cs.CR cs.AI

Heuristic-Induced Multimodal Risk Distribution Jailbreak Attack for Multimodal Large Language Models

Ma Teng, Jia Xiaojun, Duan Ranjie, Li Xinfeng, Huang Yihao, Jia Xiaoshuang, Chu Zhixuan, Ren Wenqi

Comments ICCV 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.11903 2025-08-19 cs.CV

OVG-HQ: Online Video Grounding with Hybrid-modal Queries

Runhao Zeng, Jiaqi Mao, Minghao Lai, Minh Hieu Phan, Yanjie Dong, Wei Wang, Qi Chen, Xiping Hu

机构 * Artificial Intelligence Research Institute, Shenzhen MSU-BIT University(人工智能研究院,深圳MSU-BIT大学) University of Adelaide(阿德莱德大学)

Comments Accepted to ICCV 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.11616 2025-08-18 cs.CV cs.AI cs.CL cs.LG

Controlling Multimodal LLMs via Reward-guided Decoding

Oscar Mañas, Pierluca D'Oro, Koustuv Sinha, Adriana Romero-Soriano, Michal Drozdzal, Aishwarya Agrawal

机构 * Mila - Quebec AI Institute(魁北克AI研究院) Université de Montréal(蒙特利尔大学) McGill University(麦吉尔大学) Meta FAIR Canada CIFAR AI Chair(加拿大CIFAR人工智能主席)

Comments Published at ICCV 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.11502 2025-08-18 cs.CV

AIM: Amending Inherent Interpretability via Self-Supervised Masking

Eyad Alshami, Shashank Agnihotri, Bernt Schiele, Margret Keuper

机构 * Max-Planck-Institute for Informatics(马克斯·普朗克研究所(信息学)) RTG Neuroexplicit Models of Language, Vision, and Action(神经显性语言、视觉和行动模型研究组) Data and Web Science Group(数据与网络科学小组) University of Mannheim(曼海姆大学)

Comments Accepted at International Conference on Computer Vision (ICCV) 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.11446 2025-08-18 cs.CV cs.AI

Inside Knowledge: Graph-based Path Generation with Explainable Data Augmentation and Curriculum Learning for Visual Indoor Navigation

Daniel Airinei, Elena Burceanu, Marius Leordeanu

机构 * National University of Science and Technology POLITEHNICA Bucharest(波兰堡国立科学与技术大学) Bitdefender(Bitdefender公司)

Comments Accepted at the International Conference on Computer Vision Workshops 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.11265 2025-08-18 cs.CV

Domain-aware Category-level Geometry Learning Segmentation for 3D Point Clouds

Pei He, Lingling Li, Licheng Jiao, Ronghua Shang, Fang Liu, Shuang Wang, Xu Liu, Wenping Ma

机构 * Xidian University(西电大学)

Comments to be published in International Conference on Computer Vision, ICCV 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.11185 2025-08-18 cs.CV cs.LG

CHARM3R: Towards Unseen Camera Height Robust Monocular 3D Detector

Abhinav Kumar, Yuliang Guo, Zhihao Zhang, Xinyu Huang, Liu Ren, Xiaoming Liu

机构 * Michigan State University(密歇根州立大学) Bosch Research North America(博世北美研究部) Bosch Center for AI(博世人工智能中心)

Comments ICCV 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.11074 2025-08-18 cs.SD cs.AI cs.CV eess.AS

LD-LAudio-V1: Video-to-Long-Form-Audio Generation Extension with Dual Lightweight Adapters

Haomin Zhang, Kristin Qi, Shuxin Yang, Zihao Chen, Chaofan Ding, Xinhan Di

机构 * Giant Network, China(中国巨网) Computer Science, University of Massachusetts Boston(马萨诸塞大学波士顿分校计算机科学系)

Comments Gen4AVC@ICCV: 1st Workshop on Generative AI for Audio-Visual Content Creation

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.11049 2025-08-18 cs.RO cs.CV

GenFlowRL: Shaping Rewards with Generative Object-Centric Flow in Visual Reinforcement Learning

Kelin Yu, Sheng Zhang, Harshit Soora, Furong Huang, Heng Huang, Pratap Tokekar, Ruohan Gao

机构 * University of Maryland, College Park(马里兰大学 College Park分校)

Comments Published at ICCV 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.09824 2025-08-18 cs.CV

Reverse Convolution and Its Applications to Image Restoration

Xuhong Huang, Shiqi Liu, Kai Zhang, Ying Tai, Jian Yang, Hui Zeng, Lei Zhang

机构 * Nanjing University(南京大学) The Hong Kong Polytechnic University(香港理工大学) OPPO Research Institute(OPPO研究院)

Comments ICCV 2025; https://github.com/cszn/ConverseNet

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.09597 2025-08-18 cs.CV

SVG-Head: Hybrid Surface-Volumetric Gaussians for High-Fidelity Head Reconstruction and Real-Time Editing

Heyi Sun, Cong Wang, Tian-Xing Xu, Jingwei Huang, Di Kang, Chunchao Guo, Song-Hai Zhang

机构 * Tsinghua University(清华大学) Tencent Hunyuan project page(腾讯混元项目)

Comments Accepted by ICCV 2025. Project page: https://heyy-sun.github.io/SVG-Head/

详情

展开后加载摘要…

URL PDF HTML 收藏