arXivDaily arXiv每日学术速递 周一至周五更新

期刊&会议

Conference on Computer Vision and Pattern Recognition · 会议 · Computer Vision

共收录 11877
2504.03193 2025-04-16 cs.CV

Mamba as a Bridge: Where Vision Foundation Models Meet Vision Language Models for Domain-Generalized Semantic Segmentation

Xin Zhang, Robby T. Tan

Comments Accepted to CVPR 2025 (Highlight)

详情

展开后加载摘要…

URL PDF HTML 收藏
2503.08306 2025-04-16 cs.RO cs.CV cs.LG

Reasoning in visual navigation of end-to-end trained agents: a dynamical systems approach

Steeven Janny, Hervé Poirier, Leonid Antsfeld, Guillaume Bono, Gianluca Monaci, Boris Chidlovskii, Francesco Giuliari, Alessio Del Bue, Christian Wolf

Journal ref Computer Vision and Pattern Recognition Conference (CVPR) 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2411.19235 2025-04-16 cs.CV

InstanceGaussian: Appearance-Semantic Joint Gaussian Representation for 3D Instance-Level Perception

Haijie Li, Yanmin Wu, Jiarui Meng, Qiankun Gao, Zhiyao Zhang, Ronggang Wang, Jian Zhang

Comments 14 pages, accepted by CVPR 2025 as poster

详情

展开后加载摘要…

URL PDF HTML 收藏
2404.03632 2025-04-16 cs.CV

Reference-Based 3D-Aware Image Editing with Triplanes

Bahri Batuhan Bilecen, Yigit Yalin, Ning Yu, Aysegul Dundar

Comments CVPR 2025 Highlight. Includes supplementary material

详情

展开后加载摘要…

URL PDF HTML 收藏
2403.17224 2025-04-16 cs.LG cs.AI

Uncertainty Quantification for Gradient-based Explanations in Neural Networks

Mihir Mulye, Matias Valdenegro-Toro

Comments 13 pages, 11 figures, UNCV @ CVPR 2025 Camera ready

详情

展开后加载摘要…

URL PDF HTML 收藏
2303.13913 2025-04-16 cs.CV cs.RO

GarmentTracking: Category-Level Garment Pose Tracking

Han Xue, Wenqiang Xu, Jieyi Zhang, Tutian Tang, Yutong Li, Wenxin Du, Ruolin Ye, Cewu Lu

Comments CVPR 2023

详情

展开后加载摘要…

URL PDF HTML 收藏
2504.07210 2025-04-15 cs.GR cs.CV cs.LG

MESA: Text-Driven Terrain Generation Using Latent Diffusion and Global Copernicus Data

Paul Borne--Pons, Mikolaj Czerkawski, Rosalie Martin, Romain Rouffet

Comments Accepted at CVPR 2025 Workshop MORSE

详情

展开后加载摘要…

URL PDF HTML 收藏
2504.10275 2025-04-15 cs.CV cs.LG

LMFormer: Lane based Motion Prediction Transformer

Harsh Yadav, Maximilian Schaefer, Kun Zhao, Tobias Meisen

Comments Accepted: Autonomous Driving Workshop, CVPR 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2504.10041 2025-04-15 cs.RO cs.CV

Prior Does Matter: Visual Navigation via Denoising Diffusion Bridge Models

Hao Ren, Yiming Zeng, Zetong Bi, Zhaoliang Wan, Junlong Huang, Hui Cheng

Journal ref The IEEE/CVF Conference on Computer Vision and Pattern Recognition (CVPR) 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2504.10000 2025-04-15 cs.CR cs.AI cs.CL cs.CV cs.LG

Do We Really Need Curated Malicious Data for Safety Alignment in Multi-modal Large Language Models?

Yanbo Wang, Jiyang Guan, Jian Liang, Ran He

Comments Accepted to CVPR 2025, codes in process

详情

展开后加载摘要…

URL PDF HTML 收藏
2504.09953 2025-04-15 cs.CV

Efficient 2D to Full 3D Human Pose Uplifting including Joint Rotations

Katja Ludwig, Yuliia Oksymets, Robin Schön, Daniel Kienzle, Rainer Lienhart

Comments accepted at CVSports@CVPR'25

详情

展开后加载摘要…

URL PDF HTML 收藏
2504.09927 2025-04-15 cs.RO

Efficient Task-specific Conditional Diffusion Policies: Shortcut Model Acceleration and SO(3) Optimization

Haiyong Yu, Yanqiong Jin, Yonghao He, Wei Sui

Comments Accepted to CVPR 2025 Workshop on 2nd MEIS

详情

展开后加载摘要…

URL PDF HTML 收藏
2504.09914 2025-04-15 cs.CV

Improving Multimodal Hateful Meme Detection Exploiting LMM-Generated Knowledge

Maria Tzelepi, Vasileios Mezaris

Comments Accepted for publication, Multimodal Learning and Applications Workshop (MULA 2025) @ IEEE/CVF CVPR 2025, Nashville, TN, USA, June 2025. This is the authors' "accepted version"

详情

展开后加载摘要…

URL PDF HTML 收藏
2504.09795 2025-04-15 cs.CL cs.AI cs.CV cs.IR

VDocRAG: Retrieval-Augmented Generation over Visually-Rich Documents

Ryota Tanaka, Taichi Iki, Taku Hasegawa, Kyosuke Nishida, Kuniko Saito, Jun Suzuki

Comments Accepted by CVPR 2025; project page: https://vdocrag.github.io

详情

展开后加载摘要…

URL PDF HTML 收藏
2504.09623 2025-04-15 cs.CV cs.AI cs.MM

Ges3ViG: Incorporating Pointing Gestures into Language-Based 3D Visual Grounding for Embodied Reference Understanding

Atharv Mahesh Mane, Dulanga Weerakoon, Vigneshwaran Subbaraju, Sougata Sen, Sanjay E. Sarma, Archan Misra

Comments Accepted to the IEEE/CVF Conference on Computer Vision and Pattern Recognition (CVPR) 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2504.09621 2025-04-15 cs.CV

Tokenize Image Patches: Global Context Fusion for Effective Haze Removal in Large Images

Jiuchen Chen, Xinyu Yan, Qizhi Xu, Kaiqi Li

Comments Accepted by CVPR 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2504.09606 2025-04-15 cs.CV

Early-Bird Diffusion: Investigating and Leveraging Timestep-Aware Early-Bird Tickets in Diffusion Models for Efficient Training

Lexington Whalen, Zhenbang Du, Haoran You, Chaojian Li, Sixu Li, Yingyan, Lin

Comments 10 pages, 5 figures. Accepted to the IEEE/CVF Conference on Computer Vision and Pattern Recognition 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2504.09601 2025-04-15 cs.CV cs.LG cs.MM eess.IV physics.med-ph

Mixture-of-Shape-Experts (MoSE): End-to-End Shape Dictionary Framework to Prompt SAM for Generalizable Medical Segmentation

Jia Wei, Xiaoqi Zhao, Jonghye Woo, Jinsong Ouyang, Georges El Fakhri, Qingyu Chen, Xiaofeng Liu

Comments Accepted to CVPR 2025 workshop

详情

展开后加载摘要…

URL PDF HTML 收藏
2504.09491 2025-04-15 cs.CV

DropoutGS: Dropping Out Gaussians for Better Sparse-view Rendering

Yexing Xu, Longguang Wang, Minglin Chen, Sheng Ao, Li Li, Yulan Guo

Comments Accepted by CVPR 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2504.09203 2025-04-15 cs.CV cs.AI

AerOSeg: Harnessing SAM for Open-Vocabulary Segmentation in Remote Sensing Images

Saikat Dutta, Akhil Vasim, Siddhant Gole, Hamid Rezatofighi, Biplab Banerjee

Comments Accepted at EarthVision workshop, CVPR 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2504.09195 2025-04-15 cs.CV cs.AI

ReferGPT: Towards Zero-Shot Referring Multi-Object Tracking

Tzoulio Chamiti, Leandro Di Bella, Adrian Munteanu, Nikos Deligiannis

Comments Accepted CVPR 2025 Workshop on Distillation of Foundation Models for Autonomous Driving

详情

展开后加载摘要…

URL PDF HTML 收藏
2504.09160 2025-04-15 cs.CV

SCFlow2: Plug-and-Play Object Pose Refiner with Shape-Constraint Scene Flow

Qingyuan Wang, Rui Song, Jiaojiao Li, Kerui Cheng, David Ferstl, Yinlin Hu

Comments Accepted by CVPR 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2504.09097 2025-04-15 cs.CV

BIGS: Bimanual Category-agnostic Interaction Reconstruction from Monocular Videos via 3D Gaussian Splatting

Jeongwan On, Kyeonghwan Gwak, Gunyoung Kang, Junuk Cha, Soohyun Hwang, Hyein Hwang, Seungryul Baek

Comments Accepted to CVPR 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2504.09086 2025-04-15 cs.CV

RICCARDO: Radar Hit Prediction and Convolution for Camera-Radar 3D Object Detection

Yunfei Long, Abhinav Kumar, Xiaoming Liu, Daniel Morris

Comments CVPR 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2503.22268 2025-04-15 cs.CV

Segment Any Motion in Videos

Nan Huang, Wenzhao Zheng, Chenfeng Xu, Kurt Keutzer, Shanghang Zhang, Angjoo Kanazawa, Qianqian Wang

Comments CVPR 2025. Website: https://motion-seg.github.io/

详情

展开后加载摘要…

URL PDF HTML 收藏
2503.22121 2025-04-15 cs.CV

Detecting Localized Deepfake Manipulations Using Action Unit-Guided Video Representations

Tharun Anand, Siva Sankar Sajeev, Pravin Nair

Comments Accepted to CVPR-W 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2412.11423 2025-04-15 cs.CV

Nearly Zero-Cost Protection Against Mimicry by Personalized Diffusion Models

Namhyuk Ahn, KiYoon Yoo, Wonhyuk Ahn, Daesik Kim, Seung-Hun Nam

Comments CVPR 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2412.00905 2025-04-15 cs.CV cs.GR

Ref-GS: Directional Factorization for 2D Gaussian Splatting

Youjia Zhang, Anpei Chen, Yumin Wan, Zikai Song, Junqing Yu, Yawei Luo, Wei Yang

Comments CVPR 2025. Project page: https://ref-gs.github.io/

详情

展开后加载摘要…

URL PDF HTML 收藏
2409.17671 2025-04-15 cs.CV

Leveraging Anthropometric Measurements to Improve Human Mesh Estimation and Ensure Consistent Body Shapes

Katja Ludwig, Julian Lorenz, Daniel Kienzle, Tuan Bui, Rainer Lienhart

Comments accepted for CVSports@CVPR'25

详情

展开后加载摘要…

URL PDF HTML 收藏
2406.01867 2025-04-15 cs.CV

MoLA: Motion Generation and Editing with Latent Diffusion Enhanced by Adversarial Training

Kengo Uchida, Takashi Shibuya, Yuhta Takida, Naoki Murata, Julian Tanke, Shusuke Takahashi, Yuki Mitsufuji

Comments CVPR 2025 HuMoGen Workshop

详情

展开后加载摘要…

URL PDF HTML 收藏