arXivDaily arXiv每日学术速递 周一至周五更新

期刊&会议

Conference on Computer Vision and Pattern Recognition · 会议 · Computer Vision

共收录 11883
2303.04995 2023-10-05 cs.CV cs.AI

Text-Visual Prompting for Efficient 2D Temporal Video Grounding

Yimeng Zhang, Xin Chen, Jinghan Jia, Sijia Liu, Ke Ding

Comments Accepted to the CVPR 2023 and code released (https://github.com/intel/TVP)

详情

展开后加载摘要…

URL PDF HTML 收藏
2305.17102 2023-10-03 cs.CV

GeoVLN: Learning Geometry-Enhanced Visual Representation with Slot Attention for Vision-and-Language Navigation

Jingyang Huo, Qiang Sun, Boyan Jiang, Haitao Lin, Yanwei Fu

Comments Accepted by CVPR 2023

详情

展开后加载摘要…

URL PDF HTML 收藏
2309.14203 2023-09-26 cs.CV

Detecting and Grounding Multi-Modal Media Manipulation and Beyond

Rui Shao, Tianxing Wu, Jianlong Wu, Liqiang Nie, Ziwei Liu

Comments Extension of our CVPR 2023 paper: arXiv:2304.02556 Code: https://github.com/rshaojimmy/MultiModal-DeepFake

详情

展开后加载摘要…

URL PDF HTML 收藏
2309.13942 2023-09-26 cs.CV cs.MM cs.SD eess.AS

Speed Co-Augmentation for Unsupervised Audio-Visual Pre-training

Jiangliu Wang, Jianbo Jiao, Yibing Song, Stephen James, Zhan Tong, Chongjian Ge, Pieter Abbeel, Yun-hui Liu

Comments Published at the CVPR 2023 Sight and Sound workshop

详情

展开后加载摘要…

URL PDF HTML 收藏
2306.09172 2023-09-26 cs.CV

Action Sensitivity Learning for the Ego4D Episodic Memory Challenge 2023

Jiayi Shao, Xiaohan Wang, Ruijie Quan, Yi Yang

Comments Accepted to CVPR 2023 Ego4D Workshop; 1st in Ego4D Moment Queries Challenge; 2nd in Ego4D Natural Language Queries Challenge

详情

展开后加载摘要…

URL PDF HTML 收藏
2303.16342 2023-09-26 cs.CV cs.AI cs.CL

Language-Guided Audio-Visual Source Separation via Trimodal Consistency

Reuben Tan, Arijit Ray, Andrea Burns, Bryan A. Plummer, Justin Salamon, Oriol Nieto, Bryan Russell, Kate Saenko

Comments Accepted at CVPR 2023

详情

展开后加载摘要…

URL PDF HTML 收藏
2301.06052 2023-09-26 cs.CV

T2M-GPT: Generating Human Motion from Textual Descriptions with Discrete Representations

Jianrong Zhang, Yangsong Zhang, Xiaodong Cun, Shaoli Huang, Yong Zhang, Hongwei Zhao, Hongtao Lu, Xi Shen

Comments Accepted to CVPR 2023. Project page: https://mael-zys.github.io/T2M-GPT/

详情

展开后加载摘要…

URL PDF HTML 收藏
1712.00796 2023-09-26 cs.CV

Multimodal Visual Concept Learning with Weakly Supervised Techniques

Giorgos Bouritsas, Petros Koutras, Athanasia Zlatintsi, Petros Maragos

Comments CVPR 2018

Journal ref Proc. IEEE/CVF Conf. Comp. Vis. Patt. Rec. (CVPR) pp. 4914 - 4923 (2018)

详情

展开后加载摘要…

URL PDF HTML 收藏
2303.10149 2023-09-22 cs.RO

CoVIO: Online Continual Learning for Visual-Inertial Odometry

Niclas Vödisch, Daniele Cattaneo, Wolfram Burgard, Abhinav Valada

Journal ref 2023 IEEE/CVF Conference on Computer Vision and Pattern Recognition Workshops (CVPRW)

详情

展开后加载摘要…

URL PDF HTML 收藏
2111.01082 2023-09-19 cs.CV cs.GR

FaceScape: 3D Facial Dataset and Benchmark for Single-View 3D Face Reconstruction

Hao Zhu, Haotian Yang, Longwei Guo, Yidi Zhang, Yanru Wang, Mingkai Huang, Menghua Wu, Qiu Shen, Ruigang Yang, Xun Cao

Comments Accepted to T-PAMI 2023; Extension of FaceScape(CVPR 2020); Code & data are available at https://github.com/zhuhao-nju/facescape

详情

展开后加载摘要…

URL PDF HTML 收藏
2210.04643 2023-09-18 cs.LG cs.AI cs.CV q-bio.NC

Critical Learning Periods for Multisensory Integration in Deep Networks

Michael Kleinman, Alessandro Achille, Stefano Soatto

Comments CVPR 2023 (Highlighted Paper)

详情

展开后加载摘要…

URL PDF HTML 收藏
2303.16617 2023-09-15 cs.CV

NeFII: Inverse Rendering for Reflectance Decomposition with Near-Field Indirect Illumination

Haoqian Wu, Zhipeng Hu, Lincheng Li, Yongqiang Zhang, Changjie Fan, Xin Yu

Comments Accepted in CVPR 2023

详情

展开后加载摘要…

URL PDF HTML 收藏
2303.14727 2023-09-13 cs.CV

You Only Need One Thing One Click: Self-Training for Weakly Supervised 3D Scene Understanding

Zhengzhe Liu, Xiaojuan Qi, Chi-Wing Fu

Comments Extension of One Thing One Click (CVPR'2021) arXiv:2104.02246

详情

展开后加载摘要…

URL PDF HTML 收藏
2309.05590 2023-09-12 cs.CV cs.AI cs.MM

Temporal Action Localization with Enhanced Instant Discriminability

Dingfeng Shi, Qiong Cao, Yujie Zhong, Shan An, Jian Cheng, Haogang Zhu, Dacheng Tao

Comments An extended version of the CVPR paper arXiv:2303.07347, submitted to IJCV

详情

展开后加载摘要…

URL PDF HTML 收藏
2212.09993 2023-09-12 cs.AI cs.CV cs.LG

Are Deep Neural Networks SMARTer than Second Graders?

Anoop Cherian, Kuan-Chuan Peng, Suhas Lohit, Kevin A. Smith, Joshua B. Tenenbaum

Comments Extended version of CVPR 2023 paper. For the SMART-101 dataset, see http://smartdataset.github.io/smart101

详情

展开后加载摘要…

URL PDF HTML 收藏
2309.04961 2023-09-12 cs.IR cs.CV

Multi-modal Extreme Classification

Anshul Mittal, Kunal Dahiya, Shreya Malani, Janani Ramaswamy, Seba Kuruvilla, Jitendra Ajmera, Keng-hao Chang, Sumeet Agarwal, Purushottam Kar, Manik Varma

Journal ref Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition (CVPR) 2022

详情

展开后加载摘要…

URL PDF HTML 收藏
2306.06622 2023-09-12 cs.CV

Weakly Supervised Visual Question Answer Generation

Charani Alampalle, Shamanthak Hegde, Soumya Jahagirdar, Shankar Gangisetty

Journal ref Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition, Pages: 5588-5596, 2023

详情

展开后加载摘要…

URL PDF HTML 收藏
2303.00601 2023-09-08 cs.CV

Multimodal Industrial Anomaly Detection via Hybrid Fusion

Yue Wang, Jinlong Peng, Jiangning Zhang, Ran Yi, Yabiao Wang, Chengjie Wang

Comments Accepted by CVPR 2023

详情

展开后加载摘要…

URL PDF HTML 收藏
2303.08356 2023-09-07 cs.CV

Leveraging TCN and Transformer for effective visual-audio fusion in continuous emotion recognition

Weiwei Zhou, Jiada Lu, Zhaolong Xiong, Weifeng Wang

Comments 2023 IEEE/CVF Conference on Computer Vision and Pattern Recognition Workshops (CVPRW)

详情

展开后加载摘要…

URL PDF HTML 收藏
2309.01017 2023-09-06 cs.CV

Contrastive Grouping with Transformer for Referring Image Segmentation

Jiajin Tang, Ge Zheng, Cheng Shi, Sibei Yang

Comments Accepted by CVPR 2023

详情

展开后加载摘要…

URL PDF HTML 收藏
2306.09067 2023-09-06 cs.CV

2nd Place Winning Solution for the CVPR2023 Visual Anomaly and Novelty Detection Challenge: Multimodal Prompting for Data-centric Anomaly Detection

Yunkang Cao, Xiaohao Xu, Chen Sun, Yuqi Cheng, Liang Gao, Weiming Shen

Comments The first two author contribute equally. CVPR workshop challenge report. arXiv admin note: substantial text overlap with arXiv:2305.10724

详情

展开后加载摘要…

URL PDF HTML 收藏
2304.08295 2023-09-06 cs.CV cs.CL cs.LG

Interactive and Explainable Region-guided Radiology Report Generation

Tim Tanida, Philip Müller, Georgios Kaissis, Daniel Rueckert

Comments Accepted at CVPR 2023

Journal ref 2023 IEEE/CVF Conference on Computer Vision and Pattern Recognition (CVPR), pp. 7433-7442

详情

展开后加载摘要…

URL PDF HTML 收藏
2211.14305 2023-09-06 cs.CV cs.GR cs.LG

SpaText: Spatio-Textual Representation for Controllable Image Generation

Omri Avrahami, Thomas Hayes, Oran Gafni, Sonal Gupta, Yaniv Taigman, Devi Parikh, Dani Lischinski, Ohad Fried, Xi Yin

Comments CVPR 2023. Project page available at: https://omriavrahami.com/spatext

详情

展开后加载摘要…

URL PDF HTML 收藏
2211.15398 2023-09-06 cs.CV cs.LG

Leveraging per Image-Token Consistency for Vision-Language Pre-training

Yunhao Gou, Tom Ko, Hansi Yang, James Kwok, Yu Zhang, Mingxuan Wang

Comments Accepted by CVPR 2023

详情

展开后加载摘要…

URL PDF HTML 收藏
2203.03996 2023-09-06 cs.CV cs.LG

DeltaCNN: End-to-End CNN Inference of Sparse Frame Differences in Videos

Mathias Parger, Chengcheng Tang, Christopher D. Twigg, Cem Keskin, Robert Wang, Markus Steinberger

Comments CVPR 2022

详情

展开后加载摘要…

URL PDF HTML 收藏
2207.00026 2023-09-04 cs.CV cs.LG cs.RO

LaserMix for Semi-Supervised LiDAR Semantic Segmentation

Lingdong Kong, Jiawei Ren, Liang Pan, Ziwei Liu

Comments CVPR 2023 (Highlight); 27 pages, 11 figures, 12 tables; Code at https://github.com/ldkong1205/LaserMix

详情

展开后加载摘要…

URL PDF HTML 收藏
2308.16880 2023-09-01 cs.CV

Text2Scene: Text-driven Indoor Scene Stylization with Part-aware Details

Inwoo Hwang, Hyeonwoo Kim, Young Min Kim

Comments Accepted to CVPR 2023

详情

展开后加载摘要…

URL PDF HTML 收藏
2211.11827 2023-09-01 eess.IV cs.CV

High-Perceptual Quality JPEG Decoding via Posterior Sampling

Sean Man, Guy Ohayon, Theo Adrai, Michael Elad

Comments Presented in NTIRE workshop as part of CVPR 2023

详情

展开后加载摘要…

URL PDF HTML 收藏
2308.16187 2023-08-31 cs.CV

Boosting Detection in Crowd Analysis via Underutilized Output Features

Shaokai Wu, Fengyu Yang

Comments project page: https://fredfyyang.github.io/Crowd-Hat/

Journal ref CVPR 2023

详情

展开后加载摘要…

URL PDF HTML 收藏
2112.04764 2023-08-31 cs.CV cs.LG cs.RO

3D-VField: Adversarial Augmentation of Point Clouds for Domain Generalization in 3D Object Detection

Alexander Lehner, Stefano Gasperini, Alvaro Marcos-Ramiro, Michael Schmidt, Mohammad-Ali Nikouei Mahani, Nassir Navab, Benjamin Busam, Federico Tombari

Comments CVPR 2022. Project page: https://3d-vfield.github.io

详情

展开后加载摘要…

URL PDF HTML 收藏