arXivDaily arXiv每日学术速递 周一至周五更新

期刊&会议

Conference on Computer Vision and Pattern Recognition · 会议 · Computer Vision

共收录 11877
2403.11116 2025-04-15 cs.CV cs.AI

PhD: A ChatGPT-Prompted Visual hallucination Evaluation Dataset

Jiazhen Liu, Yuhan Fu, Ruobing Xie, Runquan Xie, Xingwu Sun, Fengzong Lian, Zhanhui Kang, Xirong Li

Comments Accepted by CVPR 2025, Highlight

详情

展开后加载摘要…

URL PDF HTML 收藏
2504.08966 2025-04-15 cs.CV

PACT: Pruning and Clustering-Based Token Reduction for Faster Visual Language Models

Mohamed Dhouib, Davide Buscaldi, Sonia Vanier, Aymen Shabou

Comments Accepted to CVPR 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2504.08902 2025-04-15 cs.CV cs.LG

LookingGlass: Generative Anamorphoses via Laplacian Pyramid Warping

Pascal Chang, Sergio Sancho, Jingwei Tang, Markus Gross, Vinicius C. Azevedo

Comments Accepted at CVPR 2025 (Oral)

详情

展开后加载摘要…

URL PDF HTML 收藏
2504.07334 2025-04-15 cs.CV cs.AI cs.LG

Objaverse++: Curated 3D Object Dataset with Quality Annotations

Chendi Lin, Heshan Liu, Qunshu Lin, Zachary Bright, Shitao Tang, Yihui He, Minghao Liu, Ling Zhu, Cindy Le

Comments 8 pages, 8 figures. Accepted to CVPR 2025 Workshop on Efficient Large Vision Models (April 2025)

详情

展开后加载摘要…

URL PDF HTML 收藏
2412.10061 2025-04-15 cs.CV cs.GR

Quaffure: Real-Time Quasi-Static Neural Hair Simulation

Tuur Stuyck, Gene Wei-Chin Lin, Egor Larionov, Hsiao-yu Chen, Aljaz Bozic, Nikolaos Sarafianos, Doug Roble

Comments CVPR 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2412.03572 2025-04-15 cs.CV cs.AI cs.LG cs.RO

Navigation World Models

Amir Bar, Gaoyue Zhou, Danny Tran, Trevor Darrell, Yann LeCun

Comments CVPR 2025. Project page: https://www.amirbar.net/nwm/

详情

展开后加载摘要…

URL PDF HTML 收藏
2412.00932 2025-04-15 cs.CV

FIction: 4D Future Interaction Prediction from Video

Kumar Ashutosh, Georgios Pavlakos, Kristen Grauman

Comments CVPR 2025 (Highlight)

详情

展开后加载摘要…

URL PDF HTML 收藏
2411.16219 2025-04-15 cs.CV

Weakly Supervised Panoptic Segmentation for Defect-Based Grading of Fresh Produce

Manuel Knott, Divinefavour Odion, Sameer Sontakke, Anup Karwa, Thijs Defraeye

Comments Accepted as a paper to the 6th International Workshop on Agriculture-Vision: Challenges & Opportunities for Computer Vision in Agriculture in conjunction with IEEE/CVF CVPR 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2408.00672 2025-04-15 cs.CV

ExpertAF: Expert Actionable Feedback from Video

Kumar Ashutosh, Tushar Nagarajan, Georgios Pavlakos, Kris Kitani, Kristen Grauman

Comments CVPR 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2504.08729 2025-04-14 cs.CV cs.AI cs.LG

Steering CLIP's vision transformer with sparse autoencoders

Sonia Joseph, Praneet Suresh, Ethan Goldfarb, Lorenz Hufe, Yossi Gandelsman, Robert Graham, Danilo Bzdok, Wojciech Samek, Blake Aaron Richards

Comments 8 pages, 7 figures. Accepted to the CVPR 2025 Workshop on Mechanistic Interpretability for Vision (MIV)

详情

展开后加载摘要…

URL PDF HTML 收藏
2504.08710 2025-04-14 cs.CV

Hypergraph Vision Transformers: Images are More than Nodes, More than Edges

Joshua Fixelle

Comments Accepted by CVPR 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2504.08348 2025-04-14 cs.CV

Geometric Consistency Refinement for Single Image Novel View Synthesis via Test-Time Adaptation of Diffusion Models

Josef Bengtson, David Nilsson, Fredrik Kahl

Comments Accepted to CVPR 2025 EDGE Workshop. Project page: https://gc-ref.github.io/

详情

展开后加载摘要…

URL PDF HTML 收藏
2504.08296 2025-04-14 cs.CV

Generative AI for Film Creation: A Survey of Recent Advances

Ruihan Zhang, Borou Yu, Jiajian Min, Yetong Xin, Zheng Wei, Juncheng Nemo Shi, Mingzhen Huang, Xianghao Kong, Nix Liu Xin, Shanshan Jiang, Praagya Bahuguna, Mark Chan, Khushi Hora, Lijian Yang, Yongqi Liang, Runhe Bian, Yunlei Liu, Isabela Campillo Valencia, Patricia Morales Tredinick, Ilia Kozlov, Sijia Jiang, Peiwen Huang, Na Chen, Xuanxuan Liu, Anyi Rao

Comments Accepted at CVPR 2025 CVEU workshop: AI for Creative Visual Content Generation Editing and Understanding

详情

展开后加载摘要…

URL PDF HTML 收藏
2504.08125 2025-04-14 cs.CV

Gen3DEval: Using vLLMs for Automatic Evaluation of Generated 3D Objects

Shalini Maiti, Lourdes Agapito, Filippos Kokkinos

Comments CVPR 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2503.13947 2025-04-14 cs.CV

Conformal Prediction and MLLM aided Uncertainty Quantification in Scene Graph Generation

Sayak Nag, Udita Ghosh, Calvin-Khang Ta, Sarosij Bose, Jiachen Li, Amit K Roy Chowdhury

Comments Accepted at CVPR 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2503.12165 2025-04-14 cs.CV

VTON 360: High-Fidelity Virtual Try-On from Any Viewing Direction

Zijian He, Yuwei Ning, Yipeng Qin, Guangrun Wang, Sibei Yang, Liang Lin, Guanbin Li

Comments Accepted to CVPR 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2501.01421 2025-04-14 cs.CV

R-SCoRe: Revisiting Scene Coordinate Regression for Robust Large-Scale Visual Localization

Xudong Jiang, Fangjinhua Wang, Silvano Galliani, Christoph Vogel, Marc Pollefeys

Comments CVPR 2025 camera ready. Code: https://github.com/cvg/scrstudio

详情

展开后加载摘要…

URL PDF HTML 收藏
2412.20651 2025-04-14 cs.CV cs.AI

Latent Drifting in Diffusion Models for Counterfactual Medical Image Synthesis

Yousef Yeganeh, Azade Farshad, Ioannis Charisiadis, Marta Hasny, Martin Hartenberger, Björn Ommer, Nassir Navab, Ehsan Adeli

Comments Accepted to CVPR 2025 (highlight)

详情

展开后加载摘要…

URL PDF HTML 收藏
2411.06019 2025-04-14 cs.CV cs.GR

GaussianSpa: An "Optimizing-Sparsifying" Simplification Framework for Compact and High-Quality 3D Gaussian Splatting

Yangming Zhang, Wenqi Jia, Wei Niu, Miao Yin

Comments CVPR 2025. Project page at https://noodle-lab.github.io/gaussianspa/

详情

展开后加载摘要…

URL PDF HTML 收藏
2312.01508 2025-04-14 cs.CV

CityGen: Infinite and Controllable City Layout Generation

Jie Deng, Wenhao Chai, Jianshu Guo, Qixuan Huang, Junsheng Huang, Wenhao Hu, Shengyu Hao, Jenq-Neng Hwang, Gaoang Wang

Comments Accepted to CVPR 2025 USM3D Workshop

详情

展开后加载摘要…

URL PDF HTML 收藏
2504.07962 2025-04-11 cs.CV

GLUS: Global-Local Reasoning Unified into A Single Large Language Model for Video Segmentation

Lang Lin, Xueyang Yu, Ziqi Pang, Yu-Xiong Wang

Comments CVPR 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2504.07670 2025-04-11 cs.CV

LAPIS: A novel dataset for personalized image aesthetic assessment

Anne-Sofie Maerten, Li-Wei Chen, Stefanie De Winter, Christophe Bossens, Johan Wagemans

Comments accepted at the CVPR 2025 workshop on AI for Creative Visual Content Generation Editing and Understanding (CVEU)

详情

展开后加载摘要…

URL PDF HTML 收藏
2504.04834 2025-04-11 cs.CV

Learning Affine Correspondences by Integrating Geometric Constraints

Pengju Sun, Banglei Guan, Zhenbao Yu, Yang Shang, Qifeng Yu, Daniel Barath

Comments Accepted by IEEE/CVF Conference on Computer Vision and Pattern Recognition (CVPR), 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2402.18206 2025-04-11 cs.CV

Balancing Act: Distribution-Guided Debiasing in Diffusion Models

Rishubh Parihar, Abhijnya Bhat, Abhipsa Basu, Saswat Mallick, Jogendra Nath Kundu, R. Venkatesh Babu

Comments CVPR 2024. Project Page : https://ab-34.github.io/balancing_act/

详情

展开后加载摘要…

URL PDF HTML 收藏
2504.06801 2025-04-11 cs.CV

MonoPlace3D: Learning 3D-Aware Object Placement for 3D Monocular Detection

Rishubh Parihar, Srinjay Sarkar, Sarthak Vora, Jogendra Kundu, R. Venkatesh Babu

Comments CVPR 2025 Camera Ready. Project page - https://rishubhpar.github.io/monoplace3D

详情

展开后加载摘要…

URL PDF HTML 收藏
2504.06752 2025-04-11 cs.CV

Compass Control: Multi Object Orientation Control for Text-to-Image Generation

Rishubh Parihar, Vaibhav Agrawal, Sachidanand VS, R. Venkatesh Babu

Comments CVPR 2025 Camera Ready. Project page: https://rishubhpar.github.io/compasscontrol

详情

展开后加载摘要…

URL PDF HTML 收藏
2504.03970 2025-04-11 cs.CV cs.AI cs.CL cs.IR

VideoComp: Advancing Fine-Grained Compositional and Temporal Alignment in Video-Text Models

Dahun Kim, AJ Piergiovanni, Ganesh Mallya, Anelia Angelova

Comments CVPR 2025, project page at https://github.com/google-deepmind/video_comp

详情

展开后加载摘要…

URL PDF HTML 收藏
2503.18445 2025-04-11 cs.CV

Benchmarking Multi-modal Semantic Segmentation under Sensor Failures: Missing and Noisy Modality Robustness

Chenfei Liao, Kaiyu Lei, Xu Zheng, Junha Moon, Zhixiong Wang, Yixuan Wang, Danda Pani Paudel, Luc Van Gool, Xuming Hu

Comments This paper has been accepted by the CVPR 2025 Workshop: TMM-OpenWorld as an oral presentation paper

详情

展开后加载摘要…

URL PDF HTML 收藏
2411.15139 2025-04-11 cs.CV cs.RO

DiffusionDrive: Truncated Diffusion Model for End-to-End Autonomous Driving

Bencheng Liao, Shaoyu Chen, Haoran Yin, Bo Jiang, Cheng Wang, Sixu Yan, Xinbang Zhang, Xiangyu Li, Ying Zhang, Qian Zhang, Xinggang Wang

Comments Accepted to CVPR 2025 as Highlight. Code & demo & model are available at https://github.com/hustvl/DiffusionDrive

详情

展开后加载摘要…

URL PDF HTML 收藏
2411.08753 2025-04-11 cs.CV

Which Viewpoint Shows it Best? Language for Weakly Supervising View Selection in Multi-view Instructional Videos

Sagnik Majumder, Tushar Nagarajan, Ziad Al-Halah, Reina Pradhan, Kristen Grauman

Comments Accepted to CVPR 2025 (Highlight)

详情

展开后加载摘要…

URL PDF HTML 收藏