arXivDaily arXiv每日学术速递 周一至周五更新

期刊&会议

Conference on Computer Vision and Pattern Recognition · 会议 · Computer Vision

共收录 11877
2412.09401 2025-03-25 cs.CV

SLAM3R: Real-Time Dense Scene Reconstruction from Monocular RGB Videos

Yuzheng Liu, Siyan Dong, Shuzhe Wang, Yingda Yin, Yanchao Yang, Qingnan Fan, Baoquan Chen

Comments CVPR 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2412.00505 2025-03-25 cs.CV eess.IV

Good, Cheap, and Fast: Overfitted Image Compression with Wasserstein Distortion

Jona Ballé, Luca Versari, Emilien Dupont, Hyunjik Kim, Matthias Bauer

Comments 16 pages, 12 figures. Accepted for presentation at CVPR 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2411.19654 2025-03-25 cs.CV cs.GR

TexGaussian: Generating High-quality PBR Material via Octree-based 3D Gaussian Splatting

Bojun Xiong, Jialun Liu, Jiakui Hu, Chenming Wu, Jinbo Wu, Xing Liu, Chen Zhao, Errui Ding, Zhouhui Lian

Comments CVPR 2025. Project Page: https://3d-aigc.github.io/TexGaussian

详情

展开后加载摘要…

URL PDF HTML 收藏
2411.17687 2025-03-25 cs.CV

GenDeg: Diffusion-based Degradation Synthesis for Generalizable All-In-One Image Restoration

Sudarshan Rajagopalan, Nithin Gopalakrishnan Nair, Jay N. Paranjape, Vishal M. Patel

Comments Accepted to CVPR 2025. Project Page: https://sudraj2002.github.io/gendegpage/

详情

展开后加载摘要…

URL PDF HTML 收藏
2411.15648 2025-03-25 cs.CV

Sample- and Parameter-Efficient Auto-Regressive Image Models

Elad Amrani, Leonid Karlinsky, Alex Bronstein

Comments CVPR 2025 camera-ready with supplementary. For code see https://github.com/elad-amrani/xtra

详情

展开后加载摘要…

URL PDF HTML 收藏
2411.15540 2025-03-25 cs.CV cs.AI cs.LG eess.IV

Optical-Flow Guided Prompt Optimization for Coherent Video Generation

Hyelin Nam, Jaemin Kim, Dohun Lee, Jong Chul Ye

Comments CVPR 2025 (poster); project page: https://motionprompt.github.io/

详情

展开后加载摘要…

URL PDF HTML 收藏
2411.15241 2025-03-25 cs.CV

EfficientViM: Efficient Vision Mamba with Hidden State Mixer based State Space Duality

Sanghyeok Lee, Joonmyung Choi, Hyunwoo J. Kim

Comments Conference on Computer Vision and Pattern Recognition (CVPR), 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2411.15224 2025-03-25 cs.LG cs.AI

Parameter Efficient Mamba Tuning via Projector-targeted Diagonal-centric Linear Transformation

Seokil Ham, Hee-Seon Kim, Sangmin Woo, Changick Kim

Comments accepted in CVPR 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2411.13623 2025-03-25 cs.CV

Unsupervised Foundation Model-Agnostic Slide-Level Representation Learning

Tim Lenz, Peter Neidlinger, Marta Ligero, Georg Wölflein, Marko van Treeck, Jakob Nikolas Kather

Comments Got accepted at CVPR 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2411.13281 2025-03-25 cs.CV cs.AI cs.CL cs.MM

VideoAutoArena: An Automated Arena for Evaluating Large Multimodal Models in Video Analysis through User Simulation

Ziyang Luo, Haoning Wu, Dongxu Li, Jing Ma, Mohan Kankanhalli, Junnan Li

Comments CVPR 2025, Project Page: https://videoautoarena.github.io/

详情

展开后加载摘要…

URL PDF HTML 收藏
2411.11911 2025-03-25 cs.LG cs.AI cs.CV cs.RO

ModeSeq: Taming Sparse Multimodal Motion Prediction with Sequential Mode Modeling

Zikang Zhou, Hengjian Zhou, Haibo Hu, Zihao Wen, Jianping Wang, Yung-Hui Li, Yu-Kai Huang

Comments CVPR 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2411.09703 2025-03-25 cs.CV

MagicQuill: An Intelligent Interactive Image Editing System

Zichen Liu, Yue Yu, Hao Ouyang, Qiuyu Wang, Ka Leong Cheng, Wen Wang, Zhiheng Liu, Qifeng Chen, Yujun Shen

Comments Accepted to CVPR 2025. Code and demo available at https://magic-quill.github.io

详情

展开后加载摘要…

URL PDF HTML 收藏
2411.07975 2025-03-25 cs.CV cs.AI cs.CL

JanusFlow: Harmonizing Autoregression and Rectified Flow for Unified Multimodal Understanding and Generation

Yiyang Ma, Xingchao Liu, Xiaokang Chen, Wen Liu, Chengyue Wu, Zhiyu Wu, Zizheng Pan, Zhenda Xie, Haowei Zhang, Xingkai yu, Liang Zhao, Yisong Wang, Jiaying Liu, Chong Ruan

Comments Accepted by CVPR 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2410.19324 2025-03-25 cs.CV cs.LG stat.ML

Simpler Diffusion (SiD2): 1.5 FID on ImageNet512 with pixel-space diffusion

Emiel Hoogeboom, Thomas Mensink, Jonathan Heek, Kay Lamerigts, Ruiqi Gao, Tim Salimans

Comments Accepted to CVPR 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2409.19601 2025-03-25 cs.CR

Infighting in the Dark: Multi-Label Backdoor Attack in Federated Learning

Ye Li, Yanchao Zhao, Chengcheng Zhu, Jiale Zhang

Comments Accepted by CVPR 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2409.19425 2025-03-25 cs.CV

Harnessing Frozen Unimodal Encoders for Flexible Multimodal Alignment

Mayug Maniparambil, Raiymbek Akshulakov, Yasser Abdelaziz Dahou Djilali, Sanath Narayan, Ankit Singh, Noel E. O'Connor

Comments Accepted CVPR 2025; First two authors contributed equally;

详情

展开后加载摘要…

URL PDF HTML 收藏
2407.07052 2025-03-25 eess.IV cs.CV

Latent Space Imaging

Matheus Souza, Yidan Zheng, Kaizhang Kang, Yogeshwar Nath Mishra, Qiang Fu, Wolfgang Heidrich

Comments Accepted to CVPR 2025; see http://github.com/vccimaging/latent-imaging

详情

展开后加载摘要…

URL PDF HTML 收藏
2406.20099 2025-03-25 cs.CV

Odd-One-Out: Anomaly Detection by Comparing with Neighbors

Ankan Bhunia, Changjian Li, Hakan Bilen

Comments Accepted at CVPR 2025. Codes & Dataset at https://github.com/VICO-UoE/OddOneOutAD

详情

展开后加载摘要…

URL PDF HTML 收藏
2406.18158 2025-03-25 cs.RO cs.CV

3D-MVP: 3D Multiview Pretraining for Robotic Manipulation

Shengyi Qian, Kaichun Mo, Valts Blukis, David F. Fouhey, Dieter Fox, Ankit Goyal

Comments CVPR 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2406.11148 2025-03-25 cs.CV cs.AI cs.LG

Few-Shot Recognition via Stage-Wise Retrieval-Augmented Finetuning

Tian Liu, Huixin Zhang, Shubham Parashar, Shu Kong

Comments Accepted to CVPR 2025. Website and code: https://tian1327.github.io/SWAT/

详情

展开后加载摘要…

URL PDF HTML 收藏
2405.18029 2025-03-25 cs.CV cs.LG

Are Images Indistinguishable to Humans Also Indistinguishable to Classifiers?

Zebin You, Xinyu Zhang, Hanzhong Guo, Jingdong Wang, Chongxuan Li

Comments Accepted to CVPR 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2405.01373 2025-03-25 cs.CV

ATOM: Attention Mixer for Efficient Dataset Distillation

Samir Khaki, Ahmad Sajedi, Kai Wang, Lucy Z. Liu, Yuri A. Lawryshyn, Konstantinos N. Plataniotis

Comments Accepted for an oral presentation in CVPR-DD 2024

详情

展开后加载摘要…

URL PDF HTML 收藏
2404.04584 2025-03-25 cs.CV

D$^3$: Scaling Up Deepfake Detection by Learning from Discrepancy

Yongqi Yang, Zhihao Qian, Ye Zhu, Olga Russakovsky, Yu Wu

Comments 13 pages, 3 figures, accepted by CVPR 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2503.17316 2025-03-24 cs.CV

Pow3R: Empowering Unconstrained 3D Reconstruction with Camera and Scene Priors

Wonbong Jang, Philippe Weinzaepfel, Vincent Leroy, Lourdes Agapito, Jerome Revaud

Comments CVPR 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2503.17276 2025-03-24 cs.CV

HyperNVD: Accelerating Neural Video Decomposition via Hypernetworks

Maria Pilligua, Danna Xue, Javier Vazquez-Corral

Comments CVPR 2025, project page: https://hypernvd.github.io/

详情

展开后加载摘要…

URL PDF HTML 收藏
2503.17261 2025-03-24 eess.IV cs.CV

Cross-Modal Interactive Perception Network with Mamba for Lung Tumor Segmentation in PET-CT Images

Jie Mei, Chenyu Lin, Yu Qiu, Yaonan Wang, Hui Zhang, Ziyang Wang, Dong Dai

Comments Accepted to CVPR 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2503.17117 2025-03-24 astro-ph.IM astro-ph.EP cs.CV cs.LG stat.AP

A New Statistical Model of Star Speckles for Learning to Detect and Characterize Exoplanets in Direct Imaging Observations

Théo Bodrito, Olivier Flasseur, Julien Mairal, Jean Ponce, Maud Langlois, Anne-Marie Lagrange

Comments Accepted to CVPR 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2311.15965 2025-03-24 cs.CV

FALCON: Fairness Learning via Contrastive Attention Approach to Continual Semantic Scene Understanding

Thanh-Dat Truong, Utsav Prabhu, Bhiksha Raj, Jackson Cothren, Khoa Luu

Comments Accepted to CVPR'25

详情

展开后加载摘要…

URL PDF HTML 收藏
2503.17198 2025-03-24 cs.CR cs.CV cs.LG

Jailbreaking the Non-Transferable Barrier via Test-Time Data Disguising

Yongli Xiang, Ziming Hong, Lina Yao, Dadong Wang, Tongliang Liu

Comments Code is released at CVPR_JailNTL" target="_blank" rel="noopener">https://github.com/tmllab/2025_CVPR_JailNTL

详情

展开后加载摘要…

URL PDF HTML 收藏
2503.17197 2025-03-24 cs.CV

FreeUV: Ground-Truth-Free Realistic Facial UV Texture Recovery via Cross-Assembly Inference Strategy

Xingchao Yang, Takafumi Taketomi, Yuki Endo, Yoshihiro Kanamori

Comments CVPR 2025. Project: https://yangxingchao.github.io/FreeUV-page/

详情

展开后加载摘要…

URL PDF HTML 收藏