arXivDaily arXiv每日学术速递 周一至周五更新

期刊&会议

Conference on Computer Vision and Pattern Recognition · 会议 · Computer Vision

共收录 11877
2501.05069 2025-03-26 cs.CV cs.AI

Commonsense Video Question Answering through Video-Grounded Entailment Tree Reasoning

Huabin Liu, Filip Ilievski, Cees G. M. Snoek

Comments Accepted by CVPR 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2412.14963 2025-03-26 cs.CV cs.GR cs.LG

IDOL: Instant Photorealistic 3D Human Creation from a Single Image

Yiyu Zhuang, Jiaxi Lv, Hao Wen, Qing Shuai, Ailing Zeng, Hao Zhu, Shifeng Chen, Yujiu Yang, Xun Cao, Wei Liu

Comments 22 pages, 16 figures, includes main content, supplementary materials, and references

Journal ref CVPR 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2412.13652 2025-03-26 cs.CV

RelationField: Relate Anything in Radiance Fields

Sebastian Koch, Johanna Wald, Mirco Colosi, Narunas Vaskevicius, Pedro Hermosilla, Federico Tombari, Timo Ropinski

Comments CVPR 2025. Project page: https://relationfield.github.io

详情

展开后加载摘要…

URL PDF HTML 收藏
2412.12507 2025-03-26 cs.GR cs.CV

3DGUT: Enabling Distorted Cameras and Secondary Rays in Gaussian Splatting

Qi Wu, Janick Martinez Esturo, Ashkan Mirzaei, Nicolas Moenne-Loccoz, Zan Gojcic

Comments Our paper has been accepted by CVPR 2025. For more details and updates, please visit our project website: https://research.nvidia.com/labs/toronto-ai/3DGUT

详情

展开后加载摘要…

URL PDF HTML 收藏
2412.11102 2025-03-26 cs.CV

Empowering LLMs to Understand and Generate Complex Vector Graphics

Ximing Xing, Juncheng Hu, Guotao Liang, Jing Zhang, Dong Xu, Qian Yu

Comments Accepted by CVPR 2025. Project Page: https://ximinng.github.io/LLM4SVGProject/

详情

展开后加载摘要…

URL PDF HTML 收藏
2412.06359 2025-03-26 cs.RO cs.CV

On-Device Self-Supervised Learning of Low-Latency Monocular Depth from Only Events

Jesse Hagenaars, Yilun Wu, Federico Paredes-Vallés, Stein Stroobants, Guido de Croon

Comments Accepted at CVPR 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2412.06011 2025-03-26 eess.IV cs.CV

TopoCellGen: Generating Histopathology Cell Topology with a Diffusion Model

Meilong Xu, Saumya Gupta, Xiaoling Hu, Chen Li, Shahira Abousamra, Dimitris Samaras, Prateek Prasanna, Chao Chen

Comments Accepted by CVPR 2025. 15 pages, 8 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2412.05818 2025-03-26 cs.CV cs.AI cs.CL cs.LG cs.MM

SILMM: Self-Improving Large Multimodal Models for Compositional Text-to-Image Generation

Leigang Qu, Haochuan Li, Wenjie Wang, Xiang Liu, Juncheng Li, Liqiang Nie, Tat-Seng Chua

Comments CVPR 2025 Camera-ready. Project page: https://silmm.github.io/

详情

展开后加载摘要…

URL PDF HTML 收藏
2412.04456 2025-03-26 cs.CV

HeatFormer: A Neural Optimizer for Multiview Human Mesh Recovery

Yuto Matsubara, Ko Nishino

Comments To be published in CVPR 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2412.04120 2025-03-26 cs.CV

CrossSDF: 3D Reconstruction of Thin Structures From Cross-Sections

Thomas Walker, Salvatore Esposito, Daniel Rebain, Amir Vaxman, Arno Onken, Changjian Li, Oisin Mac Aodha

Journal ref 2025 IEEE/CVF Conference on Computer Vision and Pattern Recognition (CVPR)

详情

展开后加载摘要…

URL PDF HTML 收藏
2412.02168 2025-03-26 cs.CV

Generative Photography: Scene-Consistent Camera Control for Realistic Text-to-Image Synthesis

Yu Yuan, Xijun Wang, Yichen Sheng, Prateek Chennuri, Xingguang Zhang, Stanley Chan

Comments Accepted by CVPR 2025. Project page: https://generative-photography.github.io/project/

详情

展开后加载摘要…

URL PDF HTML 收藏
2412.01987 2025-03-26 cs.CV

ShowHowTo: Generating Scene-Conditioned Step-by-Step Visual Instructions

Tomáš Souček, Prajwal Gatti, Michael Wray, Ivan Laptev, Dima Damen, Josef Sivic

Comments CVPR 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2412.00719 2025-03-26 cs.CV

Synergizing Motion and Appearance: Multi-Scale Compensatory Codebooks for Talking Head Video Generation

Shuling Zhao, Fa-Ting Hong, Xiaoshui Huang, Dan Xu

Comments Accepted by CVPR 2025. Project page: https://shaelynz.github.io/synergize-motion-appearance/

详情

展开后加载摘要…

URL PDF HTML 收藏
2411.18936 2025-03-26 cs.CV

Self-Cross Diffusion Guidance for Text-to-Image Synthesis of Similar Subjects

Weimin Qiu, Jieke Wang, Meng Tang

Comments Conference on Computer Vision and Pattern Recognition (CVPR), 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2411.18335 2025-03-26 cs.CV cs.AI cs.RO

Helvipad: A Real-World Dataset for Omnidirectional Stereo Depth Estimation

Mehdi Zayene, Jannik Endres, Albias Havolli, Charles Corbière, Salim Cherkaoui, Alexandre Kontouli, Alexandre Alahi

Comments Accepted to CVPR 2025. Project page: https://vita-epfl.github.io/Helvipad

详情

展开后加载摘要…

URL PDF HTML 收藏
2411.15843 2025-03-26 cs.CV cs.LG

Unveil Inversion and Invariance in Flow Transformer for Versatile Image Editing

Pengcheng Xu, Boyuan Jiang, Xiaobin Hu, Donghao Luo, Qingdong He, Jiangning Zhang, Chengjie Wang, Yunsheng Wu, Charles Ling, Boyu Wang

Comments CVPR 2025 Page: https://pengchengpcx.github.io/EditFT/

详情

展开后加载摘要…

URL PDF HTML 收藏
2411.15553 2025-03-26 cs.CV

Improving Transferable Targeted Attacks with Feature Tuning Mixup

Kaisheng Liang, Xuelong Dai, Yanjie Li, Dong Wang, Bin Xiao

Comments CVPR 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2411.13059 2025-03-26 cs.CV

Towards Unbiased and Robust Spatio-Temporal Scene Graph Generation and Anticipation

Rohith Peddi, Saurabh, Ayush Abhay Shrivastava, Parag Singla, Vibhav Gogate

Comments CVPR 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2411.12355 2025-03-26 cs.CV

DynFocus: Dynamic Cooperative Network Empowers LLMs with Video Understanding

Yudong Han, Qingpei Guo, Liyuan Pan, Liu Liu, Yu Guan, Ming Yang

Comments Accepted by CVPR 25

详情

展开后加载摘要…

URL PDF HTML 收藏
2411.10364 2025-03-26 cs.AI cs.LG

Forming Auxiliary High-confident Instance-level Loss to Promote Learning from Label Proportions

Tianhao Ma, Han Chen, Juncheng Hu, Yungang Zhu, Ximing Li

Comments Accepted as a conference paper at CVPR 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2410.17249 2025-03-26 cs.CV

SpectroMotion: Dynamic 3D Reconstruction of Specular Scenes

Cheng-De Fan, Chen-Wei Chang, Yi-Ruei Liu, Jie-Ying Lee, Jiun-Long Huang, Yu-Chee Tseng, Yu-Lun Liu

Comments Paper accepted to CVPR 2025. Project page: https://cdfan0627.github.io/spectromotion/

详情

展开后加载摘要…

URL PDF HTML 收藏
2410.13862 2025-03-26 cs.CV

DepthSplat: Connecting Gaussian Splatting and Depth

Haofei Xu, Songyou Peng, Fangjinhua Wang, Hermann Blum, Daniel Barath, Andreas Geiger, Marc Pollefeys

Comments CVPR 2025, Project page: https://haofeixu.github.io/depthsplat/, Code: https://github.com/cvg/depthsplat

详情

展开后加载摘要…

URL PDF HTML 收藏
2409.16434 2025-03-26 cs.LG cs.AI cs.CV

Lessons and Insights from a Unifying Study of Parameter-Efficient Fine-Tuning (PEFT) in Visual Recognition

Zheda Mai, Ping Zhang, Cheng-Hao Tu, Hong-You Chen, Li Zhang, Wei-Lun Chao

Comments CVPR 2025. The code is available at https://github.com/OSU-MLB/ViT_PEFT_Vision

详情

展开后加载摘要…

URL PDF HTML 收藏
2406.09394 2025-03-26 cs.CV cs.GR

WonderWorld: Interactive 3D Scene Generation from a Single Image

Hong-Xing Yu, Haoyi Duan, Charles Herrmann, William T. Freeman, Jiajun Wu

Comments CVPR 2025. Project website: https://kovenyu.com/WonderWorld/. The first two authors contributed equally

详情

展开后加载摘要…

URL PDF HTML 收藏
2406.01591 2025-03-26 cs.CV

DeNVeR: Deformable Neural Vessel Representations for Unsupervised Video Vessel Segmentation

Chun-Hung Wu, Shih-Hong Chen, Chih-Yao Hu, Hsin-Yu Wu, Kai-Hsin Chen, Yu-You Chen, Chih-Hai Su, Chih-Kuo Lee, Yu-Lun Liu

Comments Paper accepted to CVPR 2025. Project page: https://kirito878.github.io/DeNVeR/

详情

展开后加载摘要…

URL PDF HTML 收藏
2404.17152 2025-03-26 cs.CV

CSCO: Connectivity Search of Convolutional Operators

Tunhou Zhang, Shiyu Li, Hsin-Pai Cheng, Feng Yan, Hai Li, Yiran Chen

Comments Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition (CVPR) Workshops (2024)

详情

展开后加载摘要…

URL PDF HTML 收藏
2306.11339 2025-03-26 cs.CV cs.LG

Masking meets Supervision: A Strong Learning Alliance

Byeongho Heo, Taekyung Kim, Sangdoo Yun, Dongyoon Han

Comments Accepted to CVPR 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2503.18880 2025-03-25 cs.CV cs.SD eess.AS

Seeing Speech and Sound: Distinguishing and Locating Audios in Visual Scenes

Hyeonggon Ryu, Seongyu Kim, Joon Son Chung, Arda Senocak

Comments CVPR 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2503.18817 2025-03-25 cs.CV cs.AI

Enhanced OoD Detection through Cross-Modal Alignment of Multi-Modal Representations

Jeonghyeon Kim, Sangheum Hwang

Comments CVPR 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2503.18803 2025-03-25 cs.CV

Change3D: Revisiting Change Detection and Captioning from A Video Modeling Perspective

Duowang Zhu, Xiaohu Huang, Haiyan Huang, Hao Zhou, Zhenfeng Shao

Comments conference paper, accepted by CVPR 2025

详情

展开后加载摘要…

URL PDF HTML 收藏