arXivDaily arXiv每日学术速递 周一至周五更新

期刊&会议

Conference on Computer Vision and Pattern Recognition · 会议 · Computer Vision

共收录 11877
2312.02051 2024-03-29 cs.CV cs.AI cs.CL

TimeChat: A Time-sensitive Multimodal Large Language Model for Long Video Understanding

Shuhuai Ren, Linli Yao, Shicheng Li, Xu Sun, Lu Hou

Comments CVPR 2024 camera-ready version, code is available at https://github.com/RenShuhuai-Andy/TimeChat

详情

展开后加载摘要…

URL PDF HTML 收藏
2311.18331 2024-03-29 cs.CV cs.AI

MRFP: Learning Generalizable Semantic Segmentation from Sim-2-Real with Multi-Resolution Feature Perturbation

Sumanth Udupa, Prajwal Gurunath, Aniruddh Sikdar, Suresh Sundaram

Comments Accepted to CVPR 2024

详情

展开后加载摘要…

URL PDF HTML 收藏
2311.17216 2024-03-29 cs.CV

Self-Discovering Interpretable Diffusion Latent Directions for Responsible Text-to-Image Generation

Hang Li, Chengzhi Shen, Philip Torr, Volker Tresp, Jindong Gu

Comments Accepted to CVPR 2024

详情

展开后加载摘要…

URL PDF HTML 收藏
2311.17113 2024-03-29 cs.CV cs.GR

Human Gaussian Splatting: Real-time Rendering of Animatable Avatars

Arthur Moreau, Jifei Song, Helisa Dhamo, Richard Shaw, Yiren Zhou, Eduardo Pérez-Pellitero

Comments Accepted to CVPR 2024

详情

展开后加载摘要…

URL PDF HTML 收藏
2311.15977 2024-03-29 cs.CV

Text2Loc: 3D Point Cloud Localization from Natural Language

Yan Xia, Letian Shi, Zifeng Ding, João F. Henriques, Daniel Cremers

Comments Accepted by CVPR 2024

详情

展开后加载摘要…

URL PDF HTML 收藏
2311.14097 2024-03-29 cs.CV

ACT-Diffusion: Efficient Adversarial Consistency Training for One-step Diffusion Models

Fei Kong, Jinhao Duan, Lichao Sun, Hao Cheng, Renjing Xu, Hengtao Shen, Xiaofeng Zhu, Xiaoshuang Shi, Kaidi Xu

Comments To appear in CVPR 2024

详情

展开后加载摘要…

URL PDF HTML 收藏
2311.11016 2024-03-29 cs.RO

SNI-SLAM: Semantic Neural Implicit SLAM

Siting Zhu, Guangming Wang, Hermann Blum, Jiuming Liu, Liang Song, Marc Pollefeys, Hesheng Wang

Comments Accepted to CVPR 2024

详情

展开后加载摘要…

URL PDF HTML 收藏
2310.08471 2024-03-29 cs.CV cs.GR

WinSyn: A High Resolution Testbed for Synthetic Data

Tom Kelly, John Femiani, Peter Wonka

Comments cvpr version

详情

展开后加载摘要…

URL PDF HTML 收藏
2308.12532 2024-03-29 cs.LG cs.AI cs.CV

FedSOL: Stabilized Orthogonal Learning with Proximal Restrictions in Federated Learning

Gihun Lee, Minchan Jeong, Sangmook Kim, Jaehoon Oh, Se-Young Yun

Comments The IEEE/CVF Conference on Computer Vision and Pattern Recognition 2024 (CVPR 2024)

详情

展开后加载摘要…

URL PDF HTML 收藏
2306.02240 2024-03-29 cs.CV

ProTeCt: Prompt Tuning for Taxonomic Open Set Classification

Tz-Ying Wu, Chih-Hui Ho, Nuno Vasconcelos

Comments Accepted to CVPR 2024

详情

展开后加载摘要…

URL PDF HTML 收藏
2211.01579 2024-03-29 cs.LG cs.CR cs.CV

Data-free Defense of Black Box Models Against Adversarial Attacks

Gaurav Kumar Nayak, Inder Khatri, Ruchit Rawal, Anirban Chakraborty

Comments CVPR Workshop (Under Review)

详情

展开后加载摘要…

URL PDF HTML 收藏
2202.04291 2024-03-29 cs.CV

L2B: Learning to Bootstrap Robust Models for Combating Label Noise

Yuyin Zhou, Xianhang Li, Fengze Liu, Qingyue Wei, Xuxi Chen, Lequan Yu, Cihang Xie, Matthew P. Lungren, Lei Xing

Comments CVPR 2024; code is available at https://github.com/yuyinzhou/l2b

详情

展开后加载摘要…

URL PDF HTML 收藏
2403.18821 2024-03-28 cs.SD cs.CV cs.MM eess.AS

Real Acoustic Fields: An Audio-Visual Room Acoustics Dataset and Benchmark

Ziyang Chen, Israel D. Gebru, Christian Richardt, Anurag Kumar, William Laney, Andrew Owens, Alexander Richard

Comments Accepted to CVPR 2024. Project site: https://facebookresearch.github.io/real-acoustic-fields/

详情

展开后加载摘要…

URL PDF HTML 收藏
2403.18775 2024-03-28 cs.CV cs.AI cs.LG

ImageNet-D: Benchmarking Neural Network Robustness on Diffusion Synthetic Object

Chenshuang Zhang, Fei Pan, Junmo Kim, In So Kweon, Chengzhi Mao

Comments Accepted at CVPR 2024

详情

展开后加载摘要…

URL PDF HTML 收藏
2403.18708 2024-03-28 cs.CV

Dense Vision Transformer Compression with Few Samples

Hanxiao Zhang, Yifan Zhou, Guo-Hua Wang, Jianxin Wu

Comments Accepted to CVPR 2024. Note: Jianxin Wu is a contributing author for the arXiv version of this paper but is not listed as an author in the CVPR version due to his role as Program Chair

详情

展开后加载摘要…

URL PDF HTML 收藏
2403.18452 2024-03-28 cs.CV cs.LG cs.RO

SingularTrajectory: Universal Trajectory Predictor Using Diffusion Model

Inhwan Bae, Young-Jae Park, Hae-Gon Jeon

Comments Accepted at CVPR 2024

详情

展开后加载摘要…

URL PDF HTML 收藏
2403.18447 2024-03-28 cs.CL cs.CV cs.LG cs.RO

Can Language Beat Numerical Regression? Language-Based Multimodal Trajectory Prediction

Inhwan Bae, Junoh Lee, Hae-Gon Jeon

Comments Accepted at CVPR 2024

详情

展开后加载摘要…

URL PDF HTML 收藏
2403.17935 2024-03-28 cs.CV

OmniVid: A Generative Framework for Universal Video Understanding

Junke Wang, Dongdong Chen, Chong Luo, Bo He, Lu Yuan, Zuxuan Wu, Yu-Gang Jiang

Comments Accepted by CVPR 2024

详情

展开后加载摘要…

URL PDF HTML 收藏
2403.17638 2024-03-28 cs.CV

Learning with Unreliability: Fast Few-shot Voxel Radiance Fields with Relative Geometric Consistency

Yingjie Xu, Bangzhen Liu, Hao Tang, Bailin Deng, Shengfeng He

Comments CVPR 2024 final version

详情

展开后加载摘要…

URL PDF HTML 收藏
2312.01220 2024-03-28 cs.CV

Boosting Object Detection with Zero-Shot Day-Night Domain Adaptation

Zhipeng Du, Miaojing Shi, Jiankang Deng

Comments Accepted to CVPR 2024

详情

展开后加载摘要…

URL PDF HTML 收藏
2311.18113 2024-03-28 cs.CV cs.GR

Back to 3D: Few-Shot 3D Keypoint Detection with Back-Projected 2D Features

Thomas Wimmer, Peter Wonka, Maks Ovsjanikov

Comments Accepted to CVPR 2024, Project page: https://wimmerth.github.io/back-to-3d.html

详情

展开后加载摘要…

URL PDF HTML 收藏
2311.17532 2024-03-28 cs.CV

Weakly-Supervised Emotion Transition Learning for Diverse 3D Co-speech Gesture Generation

Xingqun Qi, Jiahao Pan, Peng Li, Ruibin Yuan, Xiaowei Chi, Mengfei Li, Wenhan Luo, Wei Xue, Shanghang Zhang, Qifeng Liu, Yike Guo

Comments Accepted by CVPR 2024

详情

展开后加载摘要…

URL PDF HTML 收藏
2311.15803 2024-03-28 cs.CV cs.RO

SOAC: Spatio-Temporal Overlap-Aware Multi-Sensor Calibration using Neural Radiance Fields

Quentin Herau, Nathan Piasco, Moussab Bennehar, Luis Roldão, Dzmitry Tsishkou, Cyrille Migniot, Pascal Vasseur, Cédric Demonceaux

Comments Accepted at CVPR 2024. Project page: https://qherau.github.io/SOAC/

详情

展开后加载摘要…

URL PDF HTML 收藏
2311.12386 2024-03-28 cs.CV

Point, Segment and Count: A Generalized Framework for Object Counting

Zhizhong Huang, Mingliang Dai, Yi Zhang, Junping Zhang, Hongming Shan

Comments Accepted by CVPR 2024. Camera ready

详情

展开后加载摘要…

URL PDF HTML 收藏
2311.12028 2024-03-28 cs.CV cs.AI cs.LG

Hourglass Tokenizer for Efficient Transformer-Based 3D Human Pose Estimation

Wenhao Li, Mengyuan Liu, Hong Liu, Pichao Wang, Jialun Cai, Nicu Sebe

Comments Accepted by CVPR 2024, Open Sourced

详情

展开后加载摘要…

URL PDF HTML 收藏
2303.10365 2024-03-28 cs.LG cs.AI

CroSel: Cross Selection of Confident Pseudo Labels for Partial-Label Learning

Shiyu Tian, Hongxin Wei, Yiqun Wang, Lei Feng

Comments Accepted by CVPR 2024

详情

展开后加载摘要…

URL PDF HTML 收藏
2403.18383 2024-03-28 cs.CV cs.AI cs.LG

Generative Multi-modal Models are Good Class-Incremental Learners

Xusheng Cao, Haori Lu, Linlan Huang, Xialei Liu, Ming-Ming Cheng

Comments Accepted at CVPR 2024

详情

展开后加载摘要…

URL PDF HTML 收藏
2403.18356 2024-03-28 cs.CV

MonoHair: High-Fidelity Hair Modeling from a Monocular Video

Keyu Wu, Lingchen Yang, Zhiyi Kuang, Yao Feng, Xutao Han, Yuefan Shen, Hongbo Fu, Kun Zhou, Youyi Zheng

Comments Accepted by IEEE CVPR 2024

详情

展开后加载摘要…

URL PDF HTML 收藏
2403.18342 2024-03-28 cs.CV

Learning Inclusion Matching for Animation Paint Bucket Colorization

Yuekun Dai, Shangchen Zhou, Qinyue Li, Chongyi Li, Chen Change Loy

Comments accepted to CVPR 2024. Project Page: https://ykdai.github.io/projects/InclusionMatching

详情

展开后加载摘要…

URL PDF HTML 收藏
2403.18293 2024-03-28 cs.CV

Efficient Test-Time Adaptation of Vision-Language Models

Adilbek Karmanov, Dayan Guan, Shijian Lu, Abdulmotaleb El Saddik, Eric Xing

Comments Accepted to CVPR 2024. The code has been released in \url{https://kdiaaa.github.io/tda/}

详情

展开后加载摘要…

URL PDF HTML 收藏