arXivDaily arXiv每日学术速递 周一至周五更新
arXiv周末暂无论文更新,休息一下吧,周末愉快~~

期刊&会议

Conference on Computer Vision and Pattern Recognition · 会议 · Computer Vision

共收录 11877
2404.10193 2024-04-17 cs.CV

Consistency and Uncertainty: Identifying Unreliable Responses From Black-Box Vision-Language Models for Selective Visual Question Answering

Zaid Khan, Yun Fu

Comments CVPR 2024

详情

展开后加载摘要…

URL PDF HTML 收藏
2404.10157 2024-04-17 cs.CV cs.LG

Salient Object-Aware Background Generation using Text-Guided Diffusion Models

Amir Erfan Eshratifar, Joao V. B. Soares, Kapil Thadani, Shaunak Mishra, Mikhail Kuznetsov, Yueh-Ning Ku, Paloma de Juan

Comments Accepted for publication at CVPR 2024's Generative Models for Computer Vision workshop

详情

展开后加载摘要…

URL PDF HTML 收藏
2404.10146 2024-04-17 cs.CV

Cross-Modal Self-Training: Aligning Images and Pointclouds to Learn Classification without Labels

Amaya Dharmasiri, Muzammal Naseer, Salman Khan, Fahad Shahbaz Khan

Comments To be published in Workshop for Learning 3D with Multi-View Supervision (3DMV) at CVPR 2024

详情

展开后加载摘要…

URL PDF HTML 收藏
2404.10124 2024-04-17 cs.LG cs.CV

Epistemic Uncertainty Quantification For Pre-trained Neural Network

Hanjing Wang, Qiang Ji

Comments Published at CVPR 2024

详情

展开后加载摘要…

URL PDF HTML 收藏
2404.10054 2024-04-17 cs.CV cs.AI cs.CL cs.RO

AIGeN: An Adversarial Approach for Instruction Generation in VLN

Niyati Rawal, Roberto Bigazzi, Lorenzo Baraldi, Rita Cucchiara

Comments Accepted to 7th Multimodal Learning and Applications Workshop (MULA 2024) at the IEEE/CVF Conference on Computer Vision and Pattern Recognition 2024

详情

展开后加载摘要…

URL PDF HTML 收藏
2404.08504 2024-04-17 cs.CV

3D Human Scan With A Moving Event Camera

Kai Kohyama, Shintaro Shiba, Yoshimitsu Aoki

Journal ref IEEE Conference on Computer Vision and Pattern Recognition (CVPR) Workshop On Computer Vision For Mixed Reality (CV4MR), Seattle, 2024

详情

展开后加载摘要…

URL PDF HTML 收藏
2402.18975 2024-04-17 cs.CV cs.AI

Theoretically Achieving Continuous Representation of Oriented Bounding Boxes

Zi-Kai Xiao, Guo-Ye Yang, Xue Yang, Tai-Jiang Mu, Junchi Yan, Shi-min Hu

Comments 17 pages, 12 tables, 8 figures. Accepted by CVPR'24. Code: https://github.com/514flowey/JDet-COBB

详情

展开后加载摘要…

URL PDF HTML 收藏
2402.16846 2024-04-17 cs.CV cs.AI cs.CL

GROUNDHOG: Grounding Large Language Models to Holistic Segmentation

Yichi Zhang, Ziqiao Ma, Xiaofeng Gao, Suhaila Shakiah, Qiaozi Gao, Joyce Chai

Comments Accepted to CVPR 2024. Website: https://groundhog-mllm.github.io/

详情

展开后加载摘要…

URL PDF HTML 收藏
2401.06129 2024-04-17 cs.CV

Distilling Vision-Language Models on Millions of Videos

Yue Zhao, Long Zhao, Xingyi Zhou, Jialin Wu, Chun-Te Chu, Hui Miao, Florian Schroff, Hartwig Adam, Ting Liu, Boqing Gong, Philipp Krähenbühl, Liangzhe Yuan

Comments CVPR 2024. Project page: https://zhaoyue-zephyrus.github.io/video-instruction-tuning

详情

展开后加载摘要…

URL PDF HTML 收藏
2312.13150 2024-04-17 cs.CV

Splatter Image: Ultra-Fast Single-View 3D Reconstruction

Stanislaw Szymanowicz, Christian Rupprecht, Andrea Vedaldi

Comments CVPR 2024. Project page: https://szymanowiczs.github.io/splatter-image.html . Code: https://github.com/szymanowiczs/splatter-image , Demo: https://huggingface.co/spaces/szymanowiczs/splatter_image

详情

展开后加载摘要…

URL PDF HTML 收藏
2312.04016 2024-04-17 cs.CV

PartDistill: 3D Shape Part Segmentation by Vision-Language Model Distillation

Ardian Umam, Cheng-Kun Yang, Min-Hung Chen, Jen-Hui Chuang, Yen-Yu Lin

Comments CVPR 2024 Accepted

详情

展开后加载摘要…

URL PDF HTML 收藏
2312.02155 2024-04-17 cs.CV

GPS-Gaussian: Generalizable Pixel-wise 3D Gaussian Splatting for Real-time Human Novel View Synthesis

Shunyuan Zheng, Boyao Zhou, Ruizhi Shao, Boning Liu, Shengping Zhang, Liqiang Nie, Yebin Liu

Comments Accepted by CVPR 2024 (Highlight). Project page: https://shunyuanzheng.github.io/GPS-Gaussian

详情

展开后加载摘要…

URL PDF HTML 收藏
2312.02126 2024-04-17 cs.CV cs.AI cs.RO

SplaTAM: Splat, Track & Map 3D Gaussians for Dense RGB-D SLAM

Nikhil Keetha, Jay Karhade, Krishna Murthy Jatavallabhula, Gengshan Yang, Sebastian Scherer, Deva Ramanan, Jonathon Luiten

Comments CVPR 2024. Website: https://spla-tam.github.io/

详情

展开后加载摘要…

URL PDF HTML 收藏
2311.13602 2024-04-17 cs.CV

Retrieval-Augmented Layout Transformer for Content-Aware Layout Generation

Daichi Horita, Naoto Inoue, Kotaro Kikuchi, Kota Yamaguchi, Kiyoharu Aizawa

Comments Accepted to CVPR 2024 (Oral), Project website: https://udonda.github.io/RALF/ , GitHub: https://github.com/CyberAgentAILab/RALF

详情

展开后加载摘要…

URL PDF HTML 收藏
2305.11443 2024-04-17 cs.CV

Equivariant Multi-Modality Image Fusion

Zixiang Zhao, Haowen Bai, Jiangshe Zhang, Yulun Zhang, Kai Zhang, Shuang Xu, Dongdong Chen, Radu Timofte, Luc Van Gool

Comments Accepted by CVPR 2024

详情

展开后加载摘要…

URL PDF HTML 收藏
2303.15368 2024-04-17 cs.CV

2S-UDF: A Novel Two-stage UDF Learning Method for Robust Non-watertight Model Reconstruction from Multi-view Images

Junkai Deng, Fei Hou, Xuhui Chen, Wencheng Wang, Ying He

Comments accepted to CVPR 2024

详情

展开后加载摘要…

URL PDF HTML 收藏
2404.09993 2024-04-16 cs.CV

No More Ambiguity in 360° Room Layout via Bi-Layout Estimation

Yu-Ju Tsai, Jin-Cheng Jhang, Jingjing Zheng, Wei Wang, Albert Y. C. Chen, Min Sun, Cheng-Hao Kuo, Ming-Hsuan Yang

Comments CVPR 2024, Project page: https://liagm.github.io/Bi_Layout/

详情

展开后加载摘要…

URL PDF HTML 收藏
2404.09979 2024-04-16 cs.CV eess.IV

One-Click Upgrade from 2D to 3D: Sandwiched RGB-D Video Compression for Stereoscopic Teleconferencing

Yueyu Hu, Onur G. Guleryuz, Philip A. Chou, Danhang Tang, Jonathan Taylor, Rus Maxham, Yao Wang

Comments Accepted by CVPR 2024 Workshop (AIS: Vision, Graphics and AI for Streaming https://ai4streaming-workshop.github.io )

详情

展开后加载摘要…

URL PDF HTML 收藏
2404.09918 2024-04-16 cs.CV

EdgeRelight360: Text-Conditioned 360-Degree HDR Image Generation for Real-Time On-Device Video Portrait Relighting

Min-Hui Lin, Mahesh Reddy, Guillaume Berger, Michel Sarkis, Fatih Porikli, Ning Bi

Comments Camera-ready version (CVPR workshop - EDGE'24)

详情

展开后加载摘要…

URL PDF HTML 收藏
2404.09884 2024-04-16 cs.CV cs.LG

Map-Relative Pose Regression for Visual Re-Localization

Shuai Chen, Tommaso Cavallari, Victor Adrian Prisacariu, Eric Brachmann

Comments IEEE / CVF Computer Vision and Pattern Recognition Conference (CVPR) 2024, Highlight Paper

详情

展开后加载摘要…

URL PDF HTML 收藏
2404.09842 2024-04-16 cs.CV

STMixer: A One-Stage Sparse Action Detector

Tao Wu, Mengqi Cao, Ziteng Gao, Gangshan Wu, Limin Wang

Comments Extended version of the paper arXiv:2303.15879 presented at CVPR 2023. Accepted by TPAMI 2024

详情

展开后加载摘要…

URL PDF HTML 收藏
2404.09833 2024-04-16 cs.CV cs.AI

Video2Game: Real-time, Interactive, Realistic and Browser-Compatible Environment from a Single Video

Hongchi Xia, Zhi-Hao Lin, Wei-Chiu Ma, Shenlong Wang

Comments CVPR 2024. Project page (with code): https://video2game.github.io/

详情

展开后加载摘要…

URL PDF HTML 收藏
2404.09692 2024-04-16 cs.CV

XoFTR: Cross-modal Feature Matching Transformer

Önder Tuzcuoğlu, Aybora Köksal, Buğra Sofu, Sinan Kalkan, A. Aydın Alatan

Comments CVPR Image Matching Workshop, 2024. 12 pages, 7 figures, 5 tables. Codes and dataset are available at https://github.com/OnderT/XoFTR

详情

展开后加载摘要…

URL PDF HTML 收藏
2404.00368 2024-04-16 cs.CV

Towards Variable and Coordinated Holistic Co-Speech Motion Generation

Yifei Liu, Qiong Cao, Yandong Wen, Huaiguang Jiang, Changxing Ding

Comments CVPR 2024

详情

展开后加载摘要…

URL PDF HTML 收藏
2403.17192 2024-04-16 cs.CV

Strategies to Improve Real-World Applicability of Laparoscopic Anatomy Segmentation Models

Fiona R. Kolbinger, Jiangpeng He, Jinge Ma, Fengqing Zhu

Comments 14 pages, 5 figures, 4 tables; accepted for the workshop "Data Curation and Augmentation in Medical Imaging" at CVPR 2024 (archival track)

详情

展开后加载摘要…

URL PDF HTML 收藏
2403.16092 2024-04-16 cs.CV cs.RO

Are NeRFs ready for autonomous driving? Towards closing the real-to-simulation gap

Carl Lindström, Georg Hess, Adam Lilja, Maryam Fatemi, Lars Hammarstrand, Christoffer Petersson, Lennart Svensson

Comments Accepted at Workshop on Autonomous Driving, CVPR 2024

详情

展开后加载摘要…

URL PDF HTML 收藏
2312.02244 2024-04-16 cs.CV

Geometrically-driven Aggregation for Zero-shot 3D Point Cloud Understanding

Guofeng Mei, Luigi Riz, Yiming Wang, Fabio Poiesi

Comments CVPR 2024

详情

展开后加载摘要…

URL PDF HTML 收藏
2312.00362 2024-04-16 cs.CV cs.LG

Dancing with Still Images: Video Distillation via Static-Dynamic Disentanglement

Ziyu Wang, Yue Xu, Cewu Lu, Yong-Lu Li

Comments CVPR 2024, project page: https://mvig-rhos.com/video-distill

详情

展开后加载摘要…

URL PDF HTML 收藏
2309.11711 2024-04-16 cs.CV

MoDA: Leveraging Motion Priors from Videos for Advancing Unsupervised Domain Adaptation in Semantic Segmentation

Fei Pan, Xu Yin, Seokju Lee, Axi Niu, Sungeui Yoon, In So Kweon

Comments CVPR 2024 Workshop on Learning with Limited Labelled Data for Image and Video Understanding. Best Paper Award

详情

展开后加载摘要…

URL PDF HTML 收藏
2309.05418 2024-04-16 cs.CV

FlowIBR: Leveraging Pre-Training for Efficient Neural Image-Based Rendering of Dynamic Scenes

Marcel Büsching, Josef Bengtson, David Nilsson, Mårten Björkman

Comments Accepted to CVPR 2024 Workshop on Efficient Deep Learning for Computer Vision. Project page: https://flowibr.github.io

详情

展开后加载摘要…

URL PDF HTML 收藏