arXivDaily arXiv每日学术速递 周一至周五更新

期刊&会议

Conference on Computer Vision and Pattern Recognition · 会议 · Computer Vision

共收录 11877
2412.18609 2025-03-28 cs.CV

Video-Panda: Parameter-efficient Alignment for Encoder-free Video-Language Models

Jinhui Yi, Syed Talal Wasim, Yanan Luo, Muzammal Naseer, Juergen Gall

Comments CVPR 2025 camera-ready version

详情

展开后加载摘要…

URL PDF HTML 收藏
2412.16822 2025-03-28 cs.CV cs.AI cs.LG

Layer- and Timestep-Adaptive Differentiable Token Compression Ratios for Efficient Diffusion Transformers

Haoran You, Connelly Barnes, Yuqian Zhou, Yan Kang, Zhenbang Du, Wei Zhou, Lingzhi Zhang, Yotam Nitzan, Xiaoyang Liu, Zhe Lin, Eli Shechtman, Sohrab Amirghodsi, Yingyan Celine Lin

Comments Accepted by CVPR 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2412.11890 2025-03-28 cs.CV

SegMAN: Omni-scale Context Modeling with State Space Models and Local Attention for Semantic Segmentation

Yunxiang Fu, Meng Lou, Yizhou Yu

Comments CVPR 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2412.08503 2025-03-28 cs.CV

StyleStudio: Text-Driven Style Transfer with Selective Control of Style Elements

Mingkun Lei, Xue Song, Beier Zhu, Hao Wang, Chi Zhang

Comments Accepted by CVPR 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2412.07776 2025-03-28 cs.CV cs.AI cs.LG

Video Motion Transfer with Diffusion Transformers

Alexander Pondaven, Aliaksandr Siarohin, Sergey Tulyakov, Philip Torr, Fabio Pizzati

Comments CVPR 2025 - Project page: https://ditflow.github.io/

详情

展开后加载摘要…

URL PDF HTML 收藏
2412.00493 2025-03-28 cs.CV cs.CL

Video-3D LLM: Learning Position-Aware Video Representation for 3D Scene Understanding

Duo Zheng, Shijia Huang, Liwei Wang

Comments Accepted by CVPR 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2408.12692 2025-03-28 cs.AI

Rethinking Training for De-biasing Text-to-Image Generation: Unlocking the Potential of Stable Diffusion

Eunji Kim, Siwon Kim, Minjun Park, Rahim Entezari, Sungroh Yoon

Comments 19 pages; First two authors contributed equally; Accepted at CVPR 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2404.06511 2025-03-28 cs.CV cs.AI cs.LG

MoReVQA: Exploring Modular Reasoning Models for Video Question Answering

Juhong Min, Shyamal Buch, Arsha Nagrani, Minsu Cho, Cordelia Schmid

Comments CVPR 2024; updated NExT-GQA results in Appendix

详情

展开后加载摘要…

URL PDF HTML 收藏
2402.06289 2025-03-28 cs.LG cs.CR

FedMIA: An Effective Membership Inference Attack Exploiting "All for One" Principle in Federated Learning

Gongxi Zhu, Donghao Li, Hanlin Gu, Yuan Yao, Lixin Fan, Yuxing Han

Comments 14 pages, 6 figures; Accepted by CVPR 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2503.21150 2025-03-28 cs.CV cs.AI

The Devil is in Low-Level Features for Cross-Domain Few-Shot Segmentation

Yuhan Liu, Yixiong Zou, Yuhua Li, Ruixuan Li

Comments Accepted by CVPR 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2503.21140 2025-03-28 cs.CV

Recurrent Feature Mining and Keypoint Mixup Padding for Category-Agnostic Pose Estimation

Junjie Chen, Weilong Chen, Yifan Zuo, Yuming Fang

Journal ref Published in CVPR 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2503.21076 2025-03-28 cs.CV cs.LG

KAC: Kolmogorov-Arnold Classifier for Continual Learning

Yusong Hu, Zichen Liang, Fei Yang, Qibin Hou, Xialei Liu, Ming-Ming Cheng

Comments CVPR 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2503.21003 2025-03-28 cs.CV

Forensic Self-Descriptions Are All You Need for Zero-Shot Detection, Open-Set Source Attribution, and Clustering of AI-generated Images

Tai D. Nguyen, Aref Azizpour, Matthew C. Stamm

Comments Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition (CVPR) 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2503.20998 2025-03-28 cs.GR cs.CV

CoMapGS: Covisibility Map-based Gaussian Splatting for Sparse Novel View Synthesis

Youngkyoon Jang, Eduardo Pérez-Pellitero

Comments Accepted to CVPR 2025, Mistakenly submitted as a replacement for arXiv:2402.11057

详情

展开后加载摘要…

URL PDF HTML 收藏
2503.20824 2025-03-28 eess.IV cs.AI cs.LG

Exploiting Temporal State Space Sharing for Video Semantic Segmentation

Syed Ariff Syed Hesham, Yun Liu, Guolei Sun, Henghui Ding, Jing Yang, Ender Konukoglu, Xue Geng, Xudong Jiang

Comments IEEE/CVF Conference on Computer Vision and Pattern Recognition 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2503.20235 2025-03-28 cs.CV

Leveraging 3D Geometric Priors in 2D Rotation Symmetry Detection

Ahyun Seo, Minsu Cho

Comments Accepted to CVPR 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2503.11240 2025-03-28 cs.CV cs.LG

Towards Better Alignment: Training Diffusion Models with Reinforcement Learning Against Sparse Rewards

Zijing Hu, Fengda Zhang, Long Chen, Kun Kuang, Jiahui Li, Kaifeng Gao, Jun Xiao, Xin Wang, Wenwu Zhu

Comments Accepted to CVPR 2025, add references

详情

展开后加载摘要…

URL PDF HTML 收藏
2409.12259 2025-03-28 cs.CV

WiLoR: End-to-end 3D Hand Localization and Reconstruction in-the-wild

Rolandos Alexandros Potamias, Jinglei Zhang, Jiankang Deng, Stefanos Zafeiriou

Comments CVPR 2025, Project Page https://rolpotamias.github.io/WiLoR

详情

展开后加载摘要…

URL PDF HTML 收藏
2503.20748 2025-03-27 cs.CV

UniSTD: Towards Unified Spatio-Temporal Learning across Diverse Disciplines

Chen Tang, Xinzhu Ma, Encheng Su, Xiufeng Song, Xiaohong Liu, Wei-Hong Li, Lei Bai, Wanli Ouyang, Xiangyu Yue

Comments Accepted to CVPR 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2503.20746 2025-03-27 cs.CV

PhysGen3D: Crafting a Miniature Interactive World from a Single Image

Boyuan Chen, Hanxiao Jiang, Shaowei Liu, Saurabh Gupta, Yunzhu Li, Hao Zhao, Shenlong Wang

Comments CVPR 2025, Project page: https://by-luckk.github.io/PhysGen3D

详情

展开后加载摘要…

URL PDF HTML 收藏
2503.20483 2025-03-27 cs.CV cs.LG

Dissecting and Mitigating Diffusion Bias via Mechanistic Interpretability

Yingdong Shi, Changming Li, Yifan Wang, Yongxiang Zhao, Anqi Pang, Sibei Yang, Jingyi Yu, Kan Ren

Comments CVPR 2025; Project Page: https://foundation-model-research.github.io/difflens

详情

展开后加载摘要…

URL PDF HTML 收藏
2503.20354 2025-03-27 cs.CV cs.LG

SURGEON: Memory-Adaptive Fully Test-Time Adaptation via Dynamic Activation Sparsity

Ke Ma, Jiaqi Tang, Bin Guo, Fan Dang, Sicong Liu, Zhui Zhu, Lei Wu, Cheng Fang, Ying-Cong Chen, Zhiwen Yu, Yunhao Liu

Comments Accepted to CVPR 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2503.20301 2025-03-27 cs.CV

Attribute-formed Class-specific Concept Space: Endowing Language Bottleneck Model with Better Interpretability and Scalability

Jianyang Zhang, Qianli Luo, Guowu Yang, Wenjing Yang, Weide Liu, Guosheng Lin, Fengmao Lv

Comments This paper has been accepted to CVPR 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2503.20248 2025-03-27 cs.CV cs.LG

Incremental Object Keypoint Learning

Mingfu Liang, Jiahuan Zhou, Xu Zou, Ying Wu

Comments The IEEE/CVF Conference on Computer Vision and Pattern Recognition (CVPR) 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2503.20172 2025-03-27 cs.CV

Guiding Human-Object Interactions with Rich Geometry and Relations

Mengqing Xue, Yifei Liu, Ling Guo, Shaoli Huang, Changxing Ding

Comments CVPR 2025.Project website: https://lalalfhdh.github.io/rog_page/

详情

展开后加载摘要…

URL PDF HTML 收藏
2503.20101 2025-03-27 cs.CV

EBS-EKF: Accurate and High Frequency Event-based Star Tracking

Albert W Reed, Connor Hashemi, Dennis Melamed, Nitesh Menon, Keigo Hirakawa, Scott McCloskey

Comments Accepted into the proceedings of the Conference on Computer Vision and Pattern Recognition (CVPR) for 2025. Link to code and dataset is https://gitlab.kitware.com/nest-public/kw_ebs_star_tracking#

详情

展开后加载摘要…

URL PDF HTML 收藏
2503.20011 2025-03-27 cs.CV cs.RO

Hyperdimensional Uncertainty Quantification for Multimodal Uncertainty Fusion in Autonomous Vehicles Perception

Luke Chen, Junyao Wang, Trier Mortlock, Pramod Khargonekar, Mohammad Abdullah Al Faruque

Comments Accepted at CVPR 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2503.19976 2025-03-27 cs.GR cs.CV cs.LG

Thin-Shell-SfT: Fine-Grained Monocular Non-rigid 3D Surface Tracking with Neural Deformation Fields

Navami Kairanda, Marc Habermann, Shanthika Naik, Christian Theobalt, Vladislav Golyanik

Comments 15 pages, 12 figures and 3 tables; project page: https://4dqv.mpiinf.mpg.de/ThinShellSfT; CVPR 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2503.19906 2025-03-27 cs.CV

AvatarArtist: Open-Domain 4D Avatarization

Hongyu Liu, Xuan Wang, Ziyu Wan, Yue Ma, Jingye Chen, Yanbo Fan, Yujun Shen, Yibing Song, Qifeng Chen

Comments Accepted to CVPR 2025. Project page: https://kumapowerliu.github.io/AvatarArtist

详情

展开后加载摘要…

URL PDF HTML 收藏
2503.19846 2025-03-27 cs.CV cs.LG

Attention IoU: Examining Biases in CelebA using Attention Maps

Aaron Serianni, Tyler Zhu, Olga Russakovsky, Vikram V. Ramaswamy

Comments To appear in CVPR 2025. Code and data is available at https://github.com/aaronserianni/attention-iou . 15 pages, 14 figures, including appendix

详情

展开后加载摘要…

URL PDF HTML 收藏