arXivDaily arXiv每日学术速递 周一至周五更新

期刊&会议

Conference on Computer Vision and Pattern Recognition · 会议 · Computer Vision

共收录 11877
2503.21991 2025-03-31 cs.CV cs.AI cs.GR

BOOTPLACE: Bootstrapped Object Placement with Detection Transformers

Hang Zhou, Xinxin Zuo, Rui Ma, Li Cheng

Comments CVPR 2025. Project page: https://ryanhangzhou.github.io/bootplace/ , code: https://github.com/RyanHangZhou/BOOTPLACE

详情

展开后加载摘要…

URL PDF HTML 收藏
2503.21931 2025-03-31 cs.GR cs.CV

Locally Orderless Images for Optimization in Differentiable Rendering

Ishit Mehta, Manmohan Chandraker, Ravi Ramamoorthi

Comments CVPR 2025. Project: https://ishit.github.io/loir/

详情

展开后加载摘要…

URL PDF HTML 收藏
2503.21860 2025-03-31 cs.RO cs.CV

ManipTrans: Efficient Dexterous Bimanual Manipulation Transfer via Residual Learning

Kailin Li, Puhao Li, Tengyu Liu, Yuyang Li, Siyuan Huang

Comments Accepted to CVPR 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2503.21824 2025-03-31 cs.CV cs.CR

Protecting Your Video Content: Disrupting Automated Video-based LLM Annotations

Haitong Liu, Kuofeng Gao, Yang Bai, Jinmin Li, Jinxiao Shan, Tao Dai, Shu-Tao Xia

Comments Accepted by CVPR 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2503.21815 2025-03-31 quant-ph cs.AI cs.LG

ATP: Adaptive Threshold Pruning for Efficient Data Encoding in Quantum Neural Networks

Mohamed Afane, Gabrielle Ebbrecht, Ying Wang, Juntao Chen, Junaid Farooq

Comments Accepted at the IEEE/CVF Conference on Computer Vision and Pattern Recognition (CVPR), 2025.a

详情

展开后加载摘要…

URL PDF HTML 收藏
2503.19458 2025-03-31 cs.CV

GaussianUDF: Inferring Unsigned Distance Functions through 3D Gaussian Splatting

Shujuan Li, Yu-Shen Liu, Zhizhong Han

Comments Accepted by CVPR 2025. Project page: https://lisj575.github.io/GaussianUDF/

详情

展开后加载摘要…

URL PDF HTML 收藏
2503.16707 2025-03-31 cs.CV

Cross-Modal and Uncertainty-Aware Agglomeration for Open-Vocabulary 3D Scene Understanding

Jinlong Li, Cristiano Saltori, Fabio Poiesi, Nicu Sebe

Comments Accepted by CVPR 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2503.07135 2025-03-31 cs.RO cs.CV

VidBot: Learning Generalizable 3D Actions from In-the-Wild 2D Human Videos for Zero-Shot Robotic Manipulation

Hanzhi Chen, Boyang Sun, Anran Zhang, Marc Pollefeys, Stefan Leutenegger

Comments Accepted to CVPR 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2503.00359 2025-03-31 cs.CV

Solving Instance Detection from an Open-World Perspective

Qianqian Shen, Yunhan Zhao, Nahyun Kwon, Jeeeun Kim, Yanan Li, Shu Kong

Comments Accepted at CVPR 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2501.05226 2025-03-31 cs.CV cs.LG

Light Transport-aware Diffusion Posterior Sampling for Single-View Reconstruction of 3D Volumes

Ludwic Leonard, Nils Thuerey, Ruediger Westermann

Comments CVPR 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2501.04293 2025-03-31 cs.CV

TADFormer : Task-Adaptive Dynamic Transformer for Efficient Multi-Task Learning

Seungmin Baek, Soyul Lee, Hayeon Jo, Hyesong Choi, Dongbo Min

Comments CVPR 2025 accepted

详情

展开后加载摘要…

URL PDF HTML 收藏
2412.02700 2025-03-31 cs.CV

Motion Prompting: Controlling Video Generation with Motion Trajectories

Daniel Geng, Charles Herrmann, Junhwa Hur, Forrester Cole, Serena Zhang, Tobias Pfaff, Tatiana Lopez-Guevara, Carl Doersch, Yusuf Aytar, Michael Rubinstein, Chen Sun, Oliver Wang, Andrew Owens, Deqing Sun

Comments CVPR 2025 camera ready. Project page: https://motion-prompting.github.io/

详情

展开后加载摘要…

URL PDF HTML 收藏
2411.11396 2025-03-31 cs.CV

Stacking Brick by Brick: Aligned Feature Isolation for Incremental Face Forgery Detection

Jikang Cheng, Zhiyuan Yan, Ying Zhang, Li Hao, Jiaxin Ai, Qin Zou, Chen Li, Zhongyuan Wang

Journal ref The IEEE/CVF Conference on Computer Vision and Pattern Recognition 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2410.05346 2025-03-31 cs.LG cs.AI

AnyAttack: Towards Large-scale Self-supervised Adversarial Attacks on Vision-language Models

Jiaming Zhang, Junhong Ye, Xingjun Ma, Yige Li, Yunfan Yang, Yunhao Chen, Jitao Sang, Dit-Yan Yeung

Comments CVPR 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2211.09810 2025-03-31 cs.LG cs.AI cs.CR

Tightening Robustness Verification of MaxPool-based Neural Networks via Minimizing the Over-Approximation Zone

Yuan Xiao, Yuchen Chen, Shiqing Ma, Chunrong Fang, Tongtong Bai, Mingzheng Gu, Yuxin Cheng, Yanwei Chen, Zhenyu Chen

Comments Accepted to CVPR 2025. Code Link: https://github.com/xiaoyuanpigo/Ti-Lin-Hybrid-Lin

详情

展开后加载摘要…

URL PDF HTML 收藏
2503.21781 2025-03-28 cs.CV

VideoMage: Multi-Subject and Motion Customization of Text-to-Video Diffusion Models

Chi-Pin Huang, Yen-Siang Wu, Hung-Kai Chung, Kai-Po Chang, Fu-En Yang, Yu-Chiang Frank Wang

Comments CVPR 2025. Project Page: https://jasper0314-huang.github.io/videomage-customization

详情

展开后加载摘要…

URL PDF HTML 收藏
2503.21780 2025-03-28 cs.CV

Semantic Library Adaptation: LoRA Retrieval and Fusion for Open-Vocabulary Semantic Segmentation

Reza Qorbani, Gianluca Villani, Theodoros Panagiotakopoulos, Marc Botet Colomer, Linus Härenstam-Nielsen, Mattia Segu, Pier Luigi Dovesi, Jussi Karlgren, Daniel Cremers, Federico Tombari, Matteo Poggi

Comments CVPR 2025. Project page: https://thegoodailab.org/semla Code: https://github.com/rezaqorbani/SemLA

详情

展开后加载摘要…

URL PDF HTML 收藏
2503.21777 2025-03-28 cs.CV cs.LG

Test-Time Visual In-Context Tuning

Jiahao Xie, Alessio Tonioni, Nathalie Rauschmayr, Federico Tombari, Bernt Schiele

Comments CVPR 2025. Code: https://github.com/Jiahao000/VICT

详情

展开后加载摘要…

URL PDF HTML 收藏
2503.21772 2025-03-28 cs.CV

LOCORE: Image Re-ranking with Long-Context Sequence Modeling

Zilin Xiao, Pavel Suma, Ayush Sachdeva, Hao-Jen Wang, Giorgos Kordopatis-Zilos, Giorgos Tolias, Vicente Ordonez

Comments CVPR 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2503.21766 2025-03-28 cs.CV cs.AI

Stable-SCore: A Stable Registration-based Framework for 3D Shape Correspondence

Haolin Liu, Xiaohang Zhan, Zizheng Yan, Zhongjin Luo, Yuxin Wen, Xiaoguang Han

Comments Accepted by CVPR 2025. Homepage: https://haolinliu97.github.io/Stable-Score/

详情

展开后加载摘要…

URL PDF HTML 收藏
2503.21761 2025-03-28 cs.CV cs.AI cs.LG

Uni4D: Unifying Visual Foundation Models for 4D Modeling from a Single Video

David Yifan Yao, Albert J. Zhai, Shenlong Wang

Comments CVPR 2025. Project page (with code): https://davidyao99.github.io/uni4d

详情

展开后加载摘要…

URL PDF HTML 收藏
2503.21751 2025-03-28 cs.CV

Reconstructing Humans with a Biomechanically Accurate Skeleton

Yan Xia, Xiaowei Zhou, Etienne Vouga, Qixing Huang, Georgios Pavlakos

Comments CVPR 2025. Project Webpage: https://isshikihugh.github.io/HSMR/

详情

展开后加载摘要…

URL PDF HTML 收藏
2503.21747 2025-03-28 cs.CV cs.AI cs.LG

CTRL-O: Language-Controllable Object-Centric Visual Representation Learning

Aniket Didolkar, Andrii Zadaianchuk, Rabiul Awal, Maximilian Seitzer, Efstratios Gavves, Aishwarya Agrawal

Comments Accepted at CVPR 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2503.21694 2025-03-28 cs.GR cs.AI cs.CV

Progressive Rendering Distillation: Adapting Stable Diffusion for Instant Text-to-Mesh Generation without 3D Data

Zhiyuan Ma, Xinyue Liang, Rongyuan Wu, Xiangyu Zhu, Zhen Lei, Lei Zhang

Comments Accepted to CVPR 2025. Code:https://github.com/theEricMa/TriplaneTurbo. Demo:https://huggingface.co/spaces/ZhiyuanthePony/TriplaneTurbo

详情

展开后加载摘要…

URL PDF HTML 收藏
2503.21483 2025-03-28 cs.CV

BOLT: Boost Large Vision-Language Model Without Training for Long-form Video Understanding

Shuming Liu, Chen Zhao, Tianqi Xu, Bernard Ghanem

Comments Accepted to CVPR 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2503.21459 2025-03-28 cs.CV

RoadSocial: A Diverse VideoQA Dataset and Benchmark for Road Event Understanding from Social Video Narratives

Chirag Parikh, Deepti Rawat, Rakshitha R. T., Tathagata Ghosh, Ravi Kiran Sarvadevabhatla

Comments Accepted at CVPR 2025; Project Page: https://roadsocial.github.io/

详情

展开后加载摘要…

URL PDF HTML 收藏
2503.21261 2025-03-28 cs.LG

HOT: Hadamard-based Optimized Training

Seonggon Kim, Juncheol Shin, Seung-taek Woo, Eunhyeok Park

Comments Accepted in CVPR 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2503.20321 2025-03-28 cs.CV

Recovering Dynamic 3D Sketches from Videos

Jaeah Lee, Changwoon Choi, Young Min Kim, Jaesik Park

Comments Accepted to CVPR 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2503.13985 2025-03-28 cs.CV cs.AI cs.LG

DefectFill: Realistic Defect Generation with Inpainting Diffusion Model for Visual Inspection

Jaewoo Song, Daemin Park, Kanghyun Baek, Sangyub Lee, Jooyoung Choi, Eunji Kim, Sungroh Yoon

Comments Accepted to CVPR 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2501.05510 2025-03-28 cs.CV cs.AI

OVO-Bench: How Far is Your Video-LLMs from Real-World Online Video Understanding?

Yifei Li, Junbo Niu, Ziyang Miao, Chunjiang Ge, Yuanhang Zhou, Qihao He, Xiaoyi Dong, Haodong Duan, Shuangrui Ding, Rui Qian, Pan Zhang, Yuhang Zang, Yuhang Cao, Conghui He, Jiaqi Wang

Comments CVPR 2025

详情

展开后加载摘要…

URL PDF HTML 收藏