arXivDaily arXiv每日学术速递 周一至周五更新

期刊&会议

Conference on Computer Vision and Pattern Recognition · 会议 · Computer Vision

共收录 11877
2501.09898 2025-04-07 cs.CV cs.LG cs.RO

FoundationStereo: Zero-Shot Stereo Matching

Bowen Wen, Matthew Trepte, Joseph Aribido, Jan Kautz, Orazio Gallo, Stan Birchfield

Comments CVPR 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2412.20596 2025-04-07 cs.CV

Zero-Shot Image Restoration Using Few-Step Guidance of Consistency Models (and Beyond)

Tomer Garber, Tom Tirer

Comments CVPR 2025 (camera-ready). Code can be found at: https://github.com/tirer-lab/CM4IR

详情

展开后加载摘要…

URL PDF HTML 收藏
2412.19331 2025-04-07 cs.CV cs.AI cs.LG

CALICO: Part-Focused Semantic Co-Segmentation with Large Vision-Language Models

Kiet A. Nguyen, Adheesh Juvekar, Tianjiao Yu, Muntasir Wahed, Ismini Lourentzou

Comments Accepted to CVPR 2025. Project page: https://plan-lab.github.io/calico/

详情

展开后加载摘要…

URL PDF HTML 收藏
2412.16915 2025-04-07 cs.CV cs.AI cs.GR cs.SD eess.AS

FADA: Fast Diffusion Avatar Synthesis with Mixed-Supervised Multi-CFG Distillation

Tianyun Zhong, Chao Liang, Jianwen Jiang, Gaojie Lin, Jiaqi Yang, Zhou Zhao

Comments CVPR 2025, Homepage https://fadavatar.github.io/

详情

展开后加载摘要…

URL PDF HTML 收藏
2412.10153 2025-04-07 cs.CV cs.MM cs.NE

EVOS: Efficient Implicit Neural Training via EVOlutionary Selector

Weixiang Zhang, Shuzhao Xie, Chengwei Ren, Siyi Xie, Chen Tang, Shijia Ge, Mingzi Wang, Zhi Wang

Comments Accepted by CVPR 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2412.06978 2025-04-07 cs.CV

Edge-SD-SR: Low Latency and Parameter Efficient On-device Super-Resolution with Stable Diffusion via Bidirectional Conditioning

Mehdi Noroozi, Isma Hadji, Victor Escorcia, Anestis Zaganidis, Brais Martinez, Georgios Tzimiropoulos

Comments Accepted to CVPR 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2412.06786 2025-04-07 cs.CV

Retrieving Semantics from the Deep: an RAG Solution for Gesture Synthesis

M. Hamza Mughal, Rishabh Dabral, Merel C. J. Scholman, Vera Demberg, Christian Theobalt

Comments CVPR 2025. Project page: https://vcai.mpi-inf.mpg.de/projects/RAG-Gesture/

详情

展开后加载摘要…

URL PDF HTML 收藏
2412.00081 2025-04-07 cs.LG stat.ML

Task Singular Vectors: Reducing Task Interference in Model Merging

Antonio Andrea Gargiulo, Donato Crisostomi, Maria Sofia Bucarelli, Simone Scardapane, Fabrizio Silvestri, Emanuele Rodolà

Comments In Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition 2025 (CVPR)

详情

展开后加载摘要…

URL PDF HTML 收藏
2411.16198 2025-04-07 cs.CV

Interpreting Object-level Foundation Models via Visual Precision Search

Ruoyu Chen, Siyuan Liang, Jingzhi Li, Shiming Liu, Maosen Li, Zhen Huang, Hua Zhang, Xiaochun Cao

Comments Accepted to CVPR 2025 (Highlight)

详情

展开后加载摘要…

URL PDF HTML 收藏
2411.12593 2025-04-07 cs.CV cs.AI

AdaCM$^2$: On Understanding Extremely Long-Term Video with Adaptive Cross-Modality Memory Reduction

Yuanbin Man, Ying Huang, Chengming Zhang, Bingzhe Li, Wei Niu, Miao Yin

Comments CVPR 2025 Highlight

详情

展开后加载摘要…

URL PDF HTML 收藏
2411.09921 2025-04-07 cs.CV cs.AI

Motion-Grounded Video Reasoning: Understanding and Perceiving Motion at Pixel Level

Andong Deng, Tongjia Chen, Shoubin Yu, Taojiannan Yang, Lincoln Spencer, Yapeng Tian, Ajmal Saeed Mian, Mohit Bansal, Chen Chen

Comments CVPR 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2410.23132 2025-04-07 cs.CV cs.AI cs.LG

Revisiting MAE pre-training for 3D medical image segmentation

Tassilo Wald, Constantin Ulrich, Stanislav Lukyanenko, Andrei Goncharov, Alberto Paderno, Maximilian Miller, Leander Maerkisch, Paul F. Jäger, Klaus Maier-Hein

Comments CVPR 2025. Update to Camera-Ready

详情

展开后加载摘要…

URL PDF HTML 收藏
2410.16290 2025-04-07 eess.IV cs.CV

A Unified Model for Compressed Sensing MRI Across Undersampling Patterns

Armeet Singh Jatyani, Jiayun Wang, Aditi Chandrashekar, Zihui Wu, Miguel Liu-Schiaffini, Bahareh Tolooshams, Anima Anandkumar

Comments Accepted at 2025 Conference on Computer Vision and Pattern Recognition

详情

展开后加载摘要…

URL PDF HTML 收藏
2410.07838 2025-04-07 cs.CV cs.AI cs.LG

Minority-Focused Text-to-Image Generation via Prompt Optimization

Soobin Um, Jong Chul Ye

Comments CVPR 2025 (Oral), 21 pages, 10 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2404.10620 2025-04-07 cs.CV cs.LG

PyTorchGeoNodes: Enabling Differentiable Shape Programs for 3D Shape Reconstruction

Sinisa Stekovic, Arslan Artykov, Stefan Ainetter, Mattia D'Urso, Friedrich Fraundorfer

Comments Accepted at CVPR

详情

展开后加载摘要…

URL PDF HTML 收藏
2404.00916 2025-04-07 cs.CV

Gyro-based Neural Single Image Deblurring

Heemin Yang, Jaesung Rim, Seungyong Lee, Seung-Hwan Baek, Sunghyun Cho

Comments 10 pages, 10 figures, CVPR 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2504.02828 2025-04-04 cs.CV cs.AI cs.CL

Concept Lancet: Image Editing with Compositional Representation Transplant

Jinqi Luo, Tianjiao Ding, Kwan Ho Ryan Chan, Hancheng Min, Chris Callison-Burch, René Vidal

Comments Accepted in CVPR 2025. Project page at https://peterljq.github.io/project/colan

详情

展开后加载摘要…

URL PDF HTML 收藏
2504.02823 2025-04-04 cs.CV eess.IV

STING-BEE: Towards Vision-Language Model for Real-World X-ray Baggage Security Inspection

Divya Velayudhan, Abdelfatah Ahmed, Mohamad Alansari, Neha Gour, Abderaouf Behouch, Taimur Hassan, Syed Talal Wasim, Nabil Maalej, Muzammal Naseer, Juergen Gall, Mohammed Bennamoun, Ernesto Damiani, Naoufel Werghi

Comments Accepted at CVPR 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2504.02764 2025-04-04 cs.CV cs.AI

Scene Splatter: Momentum 3D Scene Generation from Single Image with Video Diffusion Model

Shengjun Zhang, Jinzhao Li, Xin Fei, Hao Liu, Yueqi Duan

Comments CVPR 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2504.02579 2025-04-04 eess.IV

Bridging the Gap between Gaussian Diffusion Models and Universal Quantization for Image Compression

Lucas Relic, Roberto Azevedo, Yang Zhang, Markus Gross, Christopher Schroers

Comments To appear at CVPR 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2504.02558 2025-04-04 cs.CV cs.AI

Rip Current Segmentation: A Novel Benchmark and YOLOv8 Baseline Results

Andrei Dumitriu, Florin Tatui, Florin Miron, Radu Tudor Ionescu, Radu Timofte

Comments Accepted at CVPR 2023 NTIRE Workshop

Journal ref 2023 IEEE/CVF Conference on Computer Vision and Pattern Recognition Workshops (CVPRW), pp. 1261-1271, June 2023

详情

展开后加载摘要…

URL PDF HTML 收藏
2504.02515 2025-04-04 cs.CV

Exploration-Driven Generative Interactive Environments

Nedko Savov, Naser Kazemi, Mohammad Mahdi, Danda Pani Paudel, Xi Wang, Luc Van Gool

Comments Accepted at CVPR 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2504.02508 2025-04-04 cs.CV

APHQ-ViT: Post-Training Quantization with Average Perturbation Hessian Based Reconstruction for Vision Transformers

Zhuguanyu Wu, Jiayi Zhang, Jiaxin Chen, Jinyang Guo, Di Huang, Yunhong Wang

Comments Accepted to CVPR 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2504.02465 2025-04-04 cs.GR cs.CV

RASP: Revisiting 3D Anamorphic Art for Shadow-Guided Packing of Irregular Objects

Soumyaratna Debnath, Ashish Tiwari, Kaustubh Sadekar, Shanmuganathan Raman

Comments Conference on Computer Vision and Pattern Recognition (CVPR) 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2504.02397 2025-04-04 cs.CV

Learning Audio-guided Video Representation with Gated Attention for Video-Text Retrieval

Boseung Jeong, Jicheol Park, Sungyeon Kim, Suha Kwak

Comments Accepted to CVPR 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2504.02351 2025-04-04 cs.CV cs.AI

Agglomerating Large Vision Encoders via Distillation for VFSS Segmentation

Chengxi Zeng, Yuxuan Jiang, Fan Zhang, Alberto Gambaruto, Tilo Burghardt

Journal ref IEEE / CVF Computer Vision and Pattern Recognition Conference (CVPR) 2025, 2nd Efficient Large Vision Models Workshop

详情

展开后加载摘要…

URL PDF HTML 收藏
2504.01957 2025-04-04 cs.CV

Toward Real-world BEV Perception: Depth Uncertainty Estimation via Gaussian Splatting

Shu-Wei Lu, Yi-Hsuan Tsai, Yi-Ting Chen

Comments Accepted to CVPR'25. https://hcis-lab.github.io/GaussianLSS/

详情

展开后加载摘要…

URL PDF HTML 收藏
2504.01956 2025-04-04 cs.CV

VideoScene: Distilling Video Diffusion Model to Generate 3D Scenes in One Step

Hanyang Wang, Fangfu Liu, Jiawei Chi, Yueqi Duan

Comments Accepted by CVPR 2025; Project Page: https://hanyang-21.github.io/VideoScene

详情

展开后加载摘要…

URL PDF HTML 收藏
2504.01298 2025-04-04 cs.CV

Direction-Aware Hybrid Representation Learning for 3D Hand Pose and Shape Estimation

Shiyong Liu, Zhihao Li, Xiao Tang, Jianzhuang Liu

Comments Accepted to CVPR 2025 workshop

详情

展开后加载摘要…

URL PDF HTML 收藏
2503.20871 2025-04-04 cs.CV cs.AI cs.CL

VinaBench: Benchmark for Faithful and Consistent Visual Narratives

Silin Gao, Sheryl Mathew, Li Mi, Sepideh Mamooler, Mengjie Zhao, Hiromi Wakaki, Yuki Mitsufuji, Syrielle Montariol, Antoine Bosselut

Comments Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition (CVPR 2025)

详情

展开后加载摘要…

URL PDF HTML 收藏