arXivDaily arXiv每日学术速递 周一至周五更新

期刊&会议

Conference on Computer Vision and Pattern Recognition · 会议 · Computer Vision

共收录 11877
2504.04956 2025-04-09 cs.GR cs.CV

REWIND: Real-Time Egocentric Whole-Body Motion Diffusion with Exemplar-Based Identity Conditioning

Jihyun Lee, Weipeng Xu, Alexander Richard, Shih-En Wei, Shunsuke Saito, Shaojie Bai, Te-Li Wang, Minhyuk Sung, Tae-Kyun Kim, Jason Saragih

Comments Accepted to CVPR 2025, project page: https://jyunlee.github.io/projects/rewind/

详情

展开后加载摘要…

URL PDF HTML 收藏
2504.02971 2025-04-09 cs.CV cs.CL

QID: Efficient Query-Informed ViTs in Data-Scarce Regimes for OCR-free Visual Document Understanding

Binh M. Le, Shaoyuan Xu, Jinmiao Fu, Zhishen Huang, Moyan Li, Yanhui Guo, Hongdong Li, Sameera Ramasinghe, Bryan Wang

Comments 8 pages, accepted by CVPR 2025 MULA

详情

展开后加载摘要…

URL PDF HTML 收藏
2501.09333 2025-04-09 cs.CV cs.AI

Prompt-CAM: Making Vision Transformers Interpretable for Fine-Grained Analysis

Arpita Chowdhury, Dipanjyoti Paul, Zheda Mai, Jianyang Gu, Ziheng Zhang, Kazi Sajeed Mehrab, Elizabeth G. Campolongo, Daniel Rubenstein, Charles V. Stewart, Anuj Karpatne, Tanya Berger-Wolf, Yu Su, Wei-Lun Chao

Comments Accepted by CVPR 2025 Main Conference

详情

展开后加载摘要…

URL PDF HTML 收藏
2501.05446 2025-04-09 cs.CV

Relative Pose Estimation through Affine Corrections of Monocular Depth Priors

Yifan Yu, Shaohui Liu, Rémi Pautrat, Marc Pollefeys, Viktor Larsson

Comments CVPR 2025 (Highlight)

详情

展开后加载摘要…

URL PDF HTML 收藏
2412.15322 2025-04-09 cs.CV cs.LG cs.SD eess.AS

MMAudio: Taming Multimodal Joint Training for High-Quality Video-to-Audio Synthesis

Ho Kei Cheng, Masato Ishii, Akio Hayakawa, Takashi Shibuya, Alexander Schwing, Yuki Mitsufuji

Comments Accepted to CVPR 2025. Project page: https://hkchengrex.github.io/MMAudio

详情

展开后加载摘要…

URL PDF HTML 收藏
2412.09680 2025-04-09 cs.CV

PBR-NeRF: Inverse Rendering with Physics-Based Neural Fields

Sean Wu, Shamik Basu, Tim Broedermann, Luc Van Gool, Christos Sakaridis

Comments CVPR 2025. 16 pages, 7 figures. Code is publicly available at https://github.com/s3anwu/pbrnerf

详情

展开后加载摘要…

URL PDF HTML 收藏
2412.05725 2025-04-09 cs.CV cs.AI

Black Swan: Abductive and Defeasible Video Reasoning in Unpredictable Events

Aditya Chinchure, Sahithya Ravi, Raymond Ng, Vered Shwartz, Boyang Li, Leonid Sigal

Comments CVPR 2025. For data, visit https://blackswan.cs.ubc.ca

详情

展开后加载摘要…

URL PDF HTML 收藏
2408.08461 2025-04-09 cs.CV

Style-Editor: Text-driven object-centric style editing

Jihun Park, Jongmin Gim, Kyoungmin Lee, Seunghun Lee, Sunghoon Im

Comments 22 pages, 19 figures, CVPR 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2407.05712 2025-04-09 cs.CV cs.AI

MobilePortrait: Real-Time One-Shot Neural Head Avatars on Mobile Devices

Jianwen Jiang, Gaojie Lin, Zhengkun Rong, Chao Liang, Yongming Zhu, Jiaqi Yang, Tianyun Zhong

Comments CVPR 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2406.02659 2025-04-09 q-bio.NC cs.AI cs.CV

Reanimating Images using Neural Representations of Dynamic Stimuli

Jacob Yeung, Andrew F. Luo, Gabriel Sarch, Margaret M. Henderson, Deva Ramanan, Michael J. Tarr

Comments Project Page: https://brain-nrds.github.io

Journal ref CVPR 2025 (oral)

详情

展开后加载摘要…

URL PDF HTML 收藏
2405.10577 2025-04-09 cs.CV cs.RO

DuoSpaceNet: Leveraging Both Bird's-Eye-View and Perspective View Representations for 3D Object Detection

Zhe Huang, Yizhe Zhao, Hao Xiao, Chenyan Wu, Lingting Ge

Comments CVPR 2025 Workshop on Autonomous Driving (WAD)

详情

展开后加载摘要…

URL PDF HTML 收藏
2504.05303 2025-04-08 cs.CV

InteractVLM: 3D Interaction Reasoning from 2D Foundational Models

Sai Kumar Dwivedi, Dimitrije Antić, Shashank Tripathi, Omid Taheri, Cordelia Schmid, Michael J. Black, Dimitrios Tzionas

Comments CVPR 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2504.05298 2025-04-08 cs.CV

One-Minute Video Generation with Test-Time Training

Karan Dalal, Daniel Koceja, Gashon Hussein, Jiarui Xu, Yue Zhao, Youjin Song, Shihao Han, Ka Chun Cheung, Jan Kautz, Carlos Guestrin, Tatsunori Hashimoto, Sanmi Koyejo, Yejin Choi, Yu Sun, Xiaolong Wang

Comments CVPR 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2504.05265 2025-04-08 cs.CV

From Sparse Signal to Smooth Motion: Real-Time Motion Generation with Rolling Prediction Models

German Barquero, Nadine Bertsch, Manojkumar Marramreddy, Carlos Chacón, Filippo Arcadu, Ferran Rigual, Nicky Sijia He, Cristina Palmero, Sergio Escalera, Yuting Ye, Robin Kips

Comments Published in CVPR'25. Webpage: https://barquerogerman.github.io/RPM/

详情

展开后加载摘要…

URL PDF HTML 收藏
2504.05152 2025-04-08 cs.CV

PanoDreamer: Consistent Text to 360-Degree Scene Generation

Zhexiao Xiong, Zhang Chen, Zhong Li, Yi Xu, Nathan Jacobs

Comments Accepted by CVPR 2025 Workshop on Computer Vision for Metaverse

详情

展开后加载摘要…

URL PDF HTML 收藏
2412.10128 2025-04-08 cs.LG stat.AP

Feature Selection for Latent Factor Models

Rittwika Kansabanik, Adrian Barbu

Comments Accepted in the CVPR conference 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2305.17929 2025-04-08 cs.CV cs.AI cs.GR

Factored-NeuS: Reconstructing Surfaces, Illumination, and Materials of Possibly Glossy Objects

Yue Fan, Ningjing Fan, Ivan Skorokhodov, Oleg Voynov, Savva Ignatyev, Evgeny Burnaev, Peter Wonka, Yiqun Wang

Comments CVPR 2025; 22 Pages; Project page: https://yiqun-wang.github.io/Factored-NeuS/

详情

展开后加载摘要…

URL PDF HTML 收藏
2504.04781 2025-04-08 cs.CV

OCC-MLLM-CoT-Alpha: Towards Multi-stage Occlusion Recognition Based on Large Language Models via 3D-Aware Supervision and Chain-of-Thoughts Guidance

Chaoyi Wang, Baoqing Li, Xinhan Di

Comments This work has been accepted to the Multimodal Algorithmic Reasoning (MAR) Workshop at CVPR 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2504.04744 2025-04-08 cs.CV cs.AI cs.RO

Grounding 3D Object Affordance with Language Instructions, Visual Observations and Interactions

He Zhu, Quyu Kong, Kechun Xu, Xunlong Xia, Bing Deng, Jieping Ye, Rong Xiong, Yue Wang

Comments CVPR 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2504.04701 2025-04-08 cs.CV

DFormerv2: Geometry Self-Attention for RGBD Semantic Segmentation

Bo-Wen Yin, Jiao-Long Cao, Ming-Ming Cheng, Qibin Hou

Comments Accepted by CVPR 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2504.04679 2025-04-08 cs.CV

DeclutterNeRF: Generative-Free 3D Scene Recovery for Occlusion Removal

Wanzhou Liu, Zhexiao Xiong, Xinyu Li, Nathan Jacobs

Comments Accepted by CVPR 2025 4th CV4Metaverse Workshop. 15 pages, 10 figures. Code and data at: https://github.com/wanzhouliu/declutter-nerf

详情

展开后加载摘要…

URL PDF HTML 收藏
2504.04566 2025-04-08 cs.CV

DyCON: Dynamic Uncertainty-aware Consistency and Contrastive Learning for Semi-supervised Medical Image Segmentation

Maregu Assefa, Muzammal Naseer, Iyyakutti Iyappan Ganapathi, Syed Sadaf Ali, Mohamed L Seghier, Naoufel Werghi

Comments Accepted to CVPR 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2504.04423 2025-04-08 cs.CV cs.AI

UniToken: Harmonizing Multimodal Understanding and Generation through Unified Visual Encoding

Yang Jiao, Haibo Qiu, Zequn Jie, Shaoxiang Chen, Jingjing Chen, Lin Ma, Yu-Gang Jiang

Comments Accpeted to CVPR 2025 workshop

详情

展开后加载摘要…

URL PDF HTML 收藏
2504.04340 2025-04-08 cs.CV

AnomalyHybrid: A Domain-agnostic Generative Framework for General Anomaly Detection

Ying Zhao

Comments Accepted to CVPR 2025 workshop on Harnessing Generative Models for Synthetic Visual Datasets (SyntaGen)

详情

展开后加载摘要…

URL PDF HTML 收藏
2504.04338 2025-04-08 cs.RO cs.CV cs.LG

Data Scaling Laws for End-to-End Autonomous Driving

Alexander Naumann, Xunjiang Gu, Tolga Dimlioglu, Mariusz Bojarski, Alperen Degirmenci, Alexander Popov, Devansh Bisla, Marco Pavone, Urs Müller, Boris Ivanovic

Comments 15 pages, 11 figures, 4 tables, CVPR 2025 Workshop on Autonomous Driving

详情

展开后加载摘要…

URL PDF HTML 收藏
2504.04156 2025-04-08 cs.CV

CoMBO: Conflict Mitigation via Branched Optimization for Class Incremental Segmentation

Kai Fang, Anqi Zhang, Guangyu Gao, Jianbo Jiao, Chi Harold Liu, Yunchao Wei

Comments Accepted by CVPR 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2504.04085 2025-04-08 cs.CV cs.AI

DocSAM: Unified Document Image Segmentation via Query Decomposition and Heterogeneous Mixed Learning

Xiao-Hui Li, Fei Yin, Cheng-Lin Liu

Comments This paper has been accepted by CVPR 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2504.03800 2025-04-08 cs.LG cs.AI cs.NE

Decision SpikeFormer: Spike-Driven Transformer for Decision Making

Wei Huang, Qinying Gu, Nanyang Ye

Comments This work has been accepted to CVPR 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2504.03438 2025-04-08 cs.CV

ZFusion: An Effective Fuser of Camera and 4D Radar for 3D Object Perception in Autonomous Driving

Sheng Yang, Tong Zhan, Shichen Qiao, Jicheng Gong, Qing Yang, Jian Wang, Yanfeng Lu

Comments CVPR 2025 WDFM-AD

详情

展开后加载摘要…

URL PDF HTML 收藏
2504.02862 2025-04-08 cs.CV cs.LG

Towards Understanding How Knowledge Evolves in Large Vision-Language Models

Sudong Wang, Yunjian Zhang, Yao Zhu, Jianing Li, Zizhe Wang, Yanwei Liu, Xiangyang Ji

Comments Accepted by CVPR 2025

详情

展开后加载摘要…

URL PDF HTML 收藏