arXivDaily arXiv每日学术速递 周一至周五更新

期刊&会议

Conference on Computer Vision and Pattern Recognition · 会议 · Computer Vision

共收录 11877
2503.20297 2025-04-04 cs.CV

Traversing Distortion-Perception Tradeoff using a Single Score-Based Generative Model

Yuhan Wang, Suzhi Bi, Ying-Jun Angela Zhang, Xiaojun Yuan

Comments Accepted by IEEE/CVF Conference on Computer Vision and Pattern Recognition 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2503.19901 2025-04-04 cs.CV

TokenHSI: Unified Synthesis of Physical Human-Scene Interactions through Task Tokenization

Liang Pan, Zeshi Yang, Zhiyang Dou, Wenjia Wang, Buzhen Huang, Bo Dai, Taku Komura, Jingbo Wang

Comments CVPR 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2412.17811 2025-04-04 cs.CV

ChatGarment: Garment Estimation, Generation and Editing via Large Language Models

Siyuan Bian, Chenghao Xu, Yuliang Xiu, Artur Grigorev, Zhen Liu, Cewu Lu, Michael J. Black, Yao Feng

Comments CVPR 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2412.09754 2025-04-04 cs.CV

ViCaS: A Dataset for Combining Holistic and Pixel-level Video Understanding using Captions with Grounded Segmentation

Ali Athar, Xueqing Deng, Liang-Chieh Chen

Comments Accepted to CVPR 2025. Project page: https://ali2500.github.io/vicas-project/

详情

展开后加载摘要…

URL PDF HTML 收藏
2412.07237 2025-04-04 cs.CV cs.AI cs.RO

ArtFormer: Controllable Generation of Diverse 3D Articulated Objects

Jiayi Su, Youhe Feng, Zheng Li, Jinhua Song, Yangfan He, Botao Ren, Botian Xu

Comments CVPR 2025. impl. repo: https://github.com/ShuYuMo2003/ArtFormer

详情

展开后加载摘要…

URL PDF HTML 收藏
2411.19041 2025-04-04 cs.CV

TAMT: Temporal-Aware Model Tuning for Cross-Domain Few-Shot Action Recognition

Yilong Wang, Zilin Gao, Qilong Wang, Zhaofeng Chen, Peihua Li, Qinghua Hu

Comments Accepted by CVPR 2025; Project page: https://github.com/TJU-YDragonW/TAMT

详情

展开后加载摘要…

URL PDF HTML 收藏
2405.05256 2025-04-04 cs.CV cs.AI cs.LG

THRONE: An Object-based Hallucination Benchmark for the Free-form Generations of Large Vision-Language Models

Prannay Kaul, Zhizhong Li, Hao Yang, Yonatan Dukler, Ashwin Swaminathan, C. J. Taylor, Stefano Soatto

Comments In CVPR 2024. Code https://github.com/amazon-science/THRONE

详情

展开后加载摘要…

URL PDF HTML 收藏
2504.02244 2025-04-04 cs.CV

SocialGesture: Delving into Multi-person Gesture Understanding

Xu Cao, Pranav Virupaksha, Wenqi Jia, Bolin Lai, Fiona Ryan, Sangmin Lee, James M. Rehg

Comments CVPR 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2504.02199 2025-04-04 cs.CV cs.AI

ESC: Erasing Space Concept for Knowledge Deletion

Tae-Young Lee, Sundong Park, Minwoo Jeon, Hyoseok Hwang, Gyeong-Moon Park

Comments 22 pages, 14 figures, 18 tables, CVPR 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2504.02168 2025-04-04 cs.CV cs.AI cs.LG

MDP: Multidimensional Vision Model Pruning with Latency Constraint

Xinglong Sun, Barath Lakshmanan, Maying Shen, Shiyi Lan, Jingde Chen, Jose M. Alvarez

Comments Accepted at CVPR 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2504.02011 2025-04-04 cs.LG cs.AI

Random Conditioning with Distillation for Data-Efficient Diffusion Model Compression

Dohyun Kim, Sehwan Park, Geonhee Han, Seung Wook Kim, Paul Hongsuck Seo

Comments Accepted to CVPR 2025. 8 pages main paper + 4 pages references + 5 pages supplementary, 9 figures in total

详情

展开后加载摘要…

URL PDF HTML 收藏
2504.02007 2025-04-04 eess.IV

OccludeNeRF: Geometric-aware 3D Scene Inpainting with Collaborative Score Distillation in NeRF

Jingyu Shi, Achleshwar Luthra, Jiazhi Li, Xiang Gao, Xiyun Song, Zongfang Lin, David Gu, Heather Yu

Comments CVPR 2025 CV4Metaverse

详情

展开后加载摘要…

URL PDF HTML 收藏
2503.23241 2025-04-04 cs.GR cs.CV

Geometry in Style: 3D Stylization via Surface Normal Deformation

Nam Anh Dinh, Itai Lang, Hyunwoo Kim, Oded Stein, Rana Hanocka

Comments CVPR 2025. Our project page is at https://threedle.github.io/geometry-in-style

详情

展开后加载摘要…

URL PDF HTML 收藏
2412.15119 2025-04-04 cs.CV

Parallelized Autoregressive Visual Generation

Yuqing Wang, Shuhuai Ren, Zhijie Lin, Yujin Han, Haoyuan Guo, Zhenheng Yang, Difan Zou, Jiashi Feng, Xihui Liu

Comments CVPR 2025 Accepted - Project Page: https://yuqingwang1029.github.io/PAR-project

详情

展开后加载摘要…

URL PDF HTML 收藏
2405.16738 2025-04-04 cs.CV

CARL: A Framework for Equivariant Image Registration

Hastings Greer, Lin Tian, Francois-Xavier Vialard, Roland Kwitt, Raul San Jose Estepar, Marc Niethammer

Comments CVPR 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2311.01479 2025-04-04 cs.LG eess.IV

Detecting Out-of-Distribution Through the Lens of Neural Collapse

Litian Liu, Yao Qin

Comments CVPR 2025 main conference paper

详情

展开后加载摘要…

URL PDF HTML 收藏
2504.01819 2025-04-03 cs.CV cs.AI

Implicit Bias Injection Attacks against Text-to-Image Diffusion Models

Huayang Huang, Xiangye Jin, Jiaxu Miao, Yu Wu

Comments Accept to CVPR 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2504.01786 2025-04-03 cs.GR cs.LG

BlenderGym: Benchmarking Foundational Model Systems for Graphics Editing

Yunqi Gu, Ian Huang, Jihyeon Je, Guandao Yang, Leonidas Guibas

Comments CVPR 2025 Accepted

详情

展开后加载摘要…

URL PDF HTML 收藏
2504.01603 2025-04-03 cs.CV

A$^\text{T}$A: Adaptive Transformation Agent for Text-Guided Subject-Position Variable Background Inpainting

Yizhe Tang, Zhimin Sun, Yuzhen Du, Ran Yi, Guangben Lu, Teng Hu, Luying Li, Lizhuang Ma, Fangyuan Zou

Comments Accepted by CVPR 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2504.01383 2025-04-03 cs.CV

v-CLR: View-Consistent Learning for Open-World Instance Segmentation

Chang-Bin Zhang, Jinhong Ni, Yujie Zhong, Kai Han

Comments Accepted by CVPR 2025, Project page: https://visual-ai.github.io/vclr, Code: https://github.com/Visual-AI/vCLR

详情

展开后加载摘要…

URL PDF HTML 收藏
2504.01348 2025-04-03 cs.CV cs.IR

Prompt-Guided Attention Head Selection for Focus-Oriented Image Retrieval

Yuji Nozawa, Yu-Chieh Lin, Kazumoto Nakamura, Youyang Ng

Comments Accepted to CVPR 2025 PixFoundation Workshop

详情

展开后加载摘要…

URL PDF HTML 收藏
2503.14485 2025-04-03 cs.GR cs.CV

Lux Post Facto: Learning Portrait Performance Relighting with Conditional Video Diffusion and a Hybrid Dataset

Yiqun Mei, Mingming He, Li Ma, Julien Philip, Wenqi Xian, David M George, Xueming Yu, Gabriel Dedic, Ahmet Levent Taşel, Ning Yu, Vishal M. Patel, Paul Debevec

Comments CVPR 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2503.02175 2025-04-03 cs.CV cs.AI cs.LG

DivPrune: Diversity-based Visual Token Pruning for Large Multimodal Models

Saeed Ranjbar Alvar, Gursimran Singh, Mohammad Akbari, Yong Zhang

Comments Accepted to CVPR 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2503.01845 2025-04-03 cs.CV

Denoising Functional Maps: Diffusion Models for Shape Correspondence

Aleksei Zhuravlev, Zorah Lähner, Vladislav Golyanik

Comments CVPR 2025; Project page: https://alekseizhuravlev.github.io/denoising-functional-maps/

详情

展开后加载摘要…

URL PDF HTML 收藏
2502.06682 2025-04-03 cs.CV

Transfer Your Perspective: Controllable 3D Generation from Any Viewpoint in a Driving Scene

Tai-Yu Pan, Sooyoung Jeon, Mengdi Fan, Jinsu Yoo, Zhenyang Feng, Mark Campbell, Kilian Q. Weinberger, Bharath Hariharan, Wei-Lun Chao

Comments Accepted to CVPR 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2412.09612 2025-04-03 cs.CV cs.AI cs.CL

Olympus: A Universal Task Router for Computer Vision Tasks

Yuanze Lin, Yunsheng Li, Dongdong Chen, Weijian Xu, Ronald Clark, Philip H. S. Torr

Comments Accepted to CVPR 2025, Project webpage: http://yuanze-lin.me/Olympus_page/

详情

展开后加载摘要…

URL PDF HTML 收藏
2411.16863 2025-04-03 cs.CV cs.AI cs.CL cs.MM

Augmenting Multimodal LLMs with Self-Reflective Tokens for Knowledge-based Visual Question Answering

Federico Cocchi, Nicholas Moratelli, Marcella Cornia, Lorenzo Baraldi, Rita Cucchiara

Comments CVPR 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2408.16807 2025-04-03 cs.CV

STEREO: A Two-Stage Framework for Adversarially Robust Concept Erasing from Text-to-Image Diffusion Models

Koushik Srivatsan, Fahad Shamshad, Muzammal Naseer, Vishal M. Patel, Karthik Nandakumar

Comments Accepted to CVPR-2025. Code: https://github.com/koushiksrivats/robust-concept-erasing

详情

展开后加载摘要…

URL PDF HTML 收藏
2406.10462 2025-04-03 cs.CV

CoMM: A Coherent Interleaved Image-Text Dataset for Multimodal Understanding and Generation

Wei Chen, Lin Li, Yongqi Yang, Bin Wen, Fan Yang, Tingting Gao, Yu Wu, Long Chen

Comments 22 pages, Accepted by CVPR 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2405.14325 2025-04-03 cs.CV

Dinomaly: The Less Is More Philosophy in Multi-Class Unsupervised Anomaly Detection

Jia Guo, Shuai Lu, Weihang Zhang, Fang Chen, Huiqi Li, Hongen Liao

Comments IEEE/CVF CVPR 2025

详情

展开后加载摘要…

URL PDF HTML 收藏