arXivDaily arXiv每日学术速递 周一至周五更新

期刊&会议

Conference on Computer Vision and Pattern Recognition · 会议 · Computer Vision

共收录 11883
2303.18246 2023-07-25 cs.CV cs.AI cs.GR

3D Human Pose Estimation via Intuitive Physics

Shashank Tripathi, Lea Müller, Chun-Hao P. Huang, Omid Taheri, Michael J. Black, Dimitrios Tzionas

Comments Accepted in CVPR'23. Project page: https://ipman.is.tue.mpg.de

详情

展开后加载摘要…

URL PDF HTML 收藏
2211.16761 2023-07-25 cs.CV

Improving Cross-Modal Retrieval with Set of Diverse Embeddings

Dongwon Kim, Namyup Kim, Suha Kwak

Comments Accepted to CVPR 2023 (Highlight)

详情

展开后加载摘要…

URL PDF HTML 收藏
2011.09094 2023-07-25 cs.CV

UP-DETR: Unsupervised Pre-training for Object Detection with Transformers

Zhigang Dai, Bolun Cai, Yugeng Lin, Junying Chen

Comments Accepted by TPAMI 2022 and CVPR 2021

Journal ref IEEE.Trans.Pattern.Anal.Mach.Intell. Oct. 21 (2002)

详情

展开后加载摘要…

URL PDF HTML 收藏
2307.11558 2023-07-24 cs.CV cs.CL

Advancing Visual Grounding with Scene Knowledge: Benchmark and Method

Zhihong Chen, Ruifei Zhang, Yibing Song, Xiang Wan, Guanbin Li

Comments Computer Vision and Natural Language Processing. 21 pages, 14 figures. CVPR-2023

详情

展开后加载摘要…

URL PDF HTML 收藏
2307.11085 2023-07-21 cs.LG cs.CV

Representation Learning in Anomaly Detection: Successes, Limits and a Grand Challenge

Yedid Hoshen

Comments Keynote talk at the Visual Anomaly and Novelty Detection Workshop, CVPR'23

详情

展开后加载摘要…

URL PDF HTML 收藏
2307.10934 2023-07-21 cs.CV

OCTraN: 3D Occupancy Convolutional Transformer Network in Unstructured Traffic Scenarios

Aditya Nalgunda Ganesh, Dhruval Pobbathi Badrinath, Harshith Mohan Kumar, Priya SS, Surabhi Narayan

Comments This work was accepted as a spotlight presentation at the Transformers for Vision Workshop @CVPR 2023

详情

展开后加载摘要…

URL PDF HTML 收藏
2307.10790 2023-07-21 cs.CV cs.RO

Behavioral Analysis of Vision-and-Language Navigation Agents

Zijiao Yang, Arjun Majumdar, Stefan Lee

Comments accepted to CVPR2023

Journal ref In Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition, pp. 2574-2582. 2023

详情

展开后加载摘要…

URL PDF HTML 收藏
2303.12112 2023-07-21 cs.CV cs.AI cs.CL cs.MM

Positive-Augmented Contrastive Learning for Image and Video Captioning Evaluation

Sara Sarto, Manuele Barraco, Marcella Cornia, Lorenzo Baraldi, Rita Cucchiara

Comments CVPR 2023 (highlight paper)

详情

展开后加载摘要…

URL PDF HTML 收藏
2210.05335 2023-07-21 cs.CV cs.CL cs.MM

MAP: Multimodal Uncertainty-Aware Vision-Language Pre-training Model

Yatai Ji, Junjie Wang, Yuan Gong, Lin Zhang, Yanru Zhu, Hongfa Wang, Jiaxing Zhang, Tetsuya Sakai, Yujiu Yang

Comments CVPR 2023 Main Track Long Paper

详情

展开后加载摘要…

URL PDF HTML 收藏
2304.06403 2023-07-20 cs.CV cs.AI

Leveraging triplet loss for unsupervised action segmentation

E. Bueno-Benito, B. Tura, M. Dimiccoli

Comments Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition (CVPR) Workshops, 2023, pp. 4921-4929

详情

展开后加载摘要…

URL PDF HTML 收藏
2304.05417 2023-07-20 cs.CV

The MONET dataset: Multimodal drone thermal dataset recorded in rural scenarios

Luigi Riz, Andrea Caraffa, Matteo Bortolon, Mohamed Lamine Mekhalfi, Davide Boscaini, André Moura, José Antunes, André Dias, Hugo Silva, Andreas Leonidou, Christos Constantinides, Christos Keleshis, Dante Abate, Fabio Poiesi

Comments Published in Computer Vision and Pattern Recognition (CVPR) Workshops 2023 - 6th Multimodal Learning and Applications Workshop

详情

展开后加载摘要…

URL PDF HTML 收藏
2307.09316 2023-07-19 cs.CV

MarS3D: A Plug-and-Play Motion-Aware Model for Semantic Segmentation on Multi-Scan 3D Point Clouds

Jiahui Liu, Chirui Chang, Jianhui Liu, Xiaoyang Wu, Lan Ma, Xiaojuan Qi

Journal ref The IEEE/CVF Conference on Computer Vision and Pattern Recognition 2023

详情

展开后加载摘要…

URL PDF HTML 收藏
2307.08317 2023-07-18 cs.CV

AltFreezing for More General Video Face Forgery Detection

Zhendong Wang, Jianmin Bao, Wengang Zhou, Weilun Wang, Houqiang Li

Comments Accepted by CVPR 2023 Highlight, code and models are available at https: //github.com/ZhendongWang6/AltFreezing

详情

展开后加载摘要…

URL PDF HTML 收藏
2307.08071 2023-07-18 cs.CV cs.HC

Dense Multitask Learning to Reconfigure Comics

Deblina Bhattacharjee, Sabine Süsstrunk, Mathieu Salzmann

Comments CVPR 2023 Workshop. arXiv admin note: text overlap with arXiv:2205.08303

详情

展开后加载摘要…

URL PDF HTML 收藏
2302.14267 2023-07-18 cs.CV

Adversarial Attack with Raindrops

Jiyuan Liu, Bingyi Lu, Mingkang Xiong, Tao Zhang, Huilin Xiong

Comments 10 pages, 7 figures, This manuscript was submitted to CVPR 2023

详情

展开后加载摘要…

URL PDF HTML 收藏
2202.06688 2023-07-18 cs.CV

Geometric Transformer for Fast and Robust Point Cloud Registration

Zheng Qin, Hao Yu, Changjian Wang, Yulan Guo, Yuxing Peng, Kai Xu

Comments Accepted by CVPR 2022. Code and models are available at https://github.com/qinzheng93/GeoTransformer

详情

展开后加载摘要…

URL PDF HTML 收藏
2303.11165 2023-07-18 cs.LG cs.CV

Computationally Budgeted Continual Learning: What Does Matter?

Ameya Prabhu, Hasan Abed Al Kader Hammoud, Puneet Dokania, Philip H. S. Torr, Ser-Nam Lim, Bernard Ghanem, Adel Bibi

Comments CVPR 2023

详情

展开后加载摘要…

URL PDF HTML 收藏
2301.08739 2023-07-18 cs.CV

FlatFormer: Flattened Window Attention for Efficient Point Cloud Transformer

Zhijian Liu, Xinyu Yang, Haotian Tang, Shang Yang, Song Han

Comments CVPR 2023. The first two authors contributed equally to this work. Project page: https://flatformer.mit.edu

详情

展开后加载摘要…

URL PDF HTML 收藏
2307.07147 2023-07-17 cs.CV

Linking vision and motion for self-supervised object-centric perception

Kaylene C. Stocking, Zak Murez, Vijay Badrinarayanan, Jamie Shotton, Alex Kendall, Claire Tomlin, Christopher P. Burgess

Comments Presented at the CVPR 2023 Vision-Centric Autonomous Driving workshop

详情

展开后加载摘要…

URL PDF HTML 收藏
2307.06569 2023-07-14 cs.CV

A Study on Differentiable Logic and LLMs for EPIC-KITCHENS-100 Unsupervised Domain Adaptation Challenge for Action Recognition 2023

Yi Cheng, Ziwei Xu, Fen Fang, Dongyun Lin, Hehe Fan, Yongkang Wong, Ying Sun, Mohan Kankanhalli

Comments Technical report submitted to CVPR 2023 EPIC-Kitchens challenges

详情

展开后加载摘要…

URL PDF HTML 收藏
2303.05156 2023-07-14 cs.CV

Local Implicit Normalizing Flow for Arbitrary-Scale Image Super-Resolution

Jie-En Yao, Li-Yuan Tsao, Yi-Chen Lo, Roy Tseng, Chia-Che Chang, Chun-Yi Lee

Comments Accepted by CVPR 2023. Code: https://github.com/JNNNNYao/LINF

详情

展开后加载摘要…

URL PDF HTML 收藏
2307.05541 2023-07-13 cs.CV

High Fidelity 3D Hand Shape Reconstruction via Scalable Graph Frequency Decomposition

Tianyu Luan, Yuanhao Zhai, Jingjing Meng, Zhong Li, Zhang Chen, Yi Xu, Junsong Yuan

Comments CVPR 2023

Journal ref In Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition, pp. 16795-16804. 2023

详情

展开后加载摘要…

URL PDF HTML 收藏
2210.09452 2023-07-13 cs.CV cs.AI cs.LG

Multiple Instance Learning via Iterative Self-Paced Supervised Contrastive Learning

Kangning Liu, Weicheng Zhu, Yiqiu Shen, Sheng Liu, Narges Razavian, Krzysztof J. Geras, Carlos Fernandez-Granda

Comments CVPR 2023 camera-ready version. The first two authors contribute equally. The last two authors are joint last authors

详情

展开后加载摘要…

URL PDF HTML 收藏
2307.05158 2023-07-12 cs.CV

A Modular Multimodal Architecture for Gaze Target Prediction: Application to Privacy-Sensitive Settings

Anshul Gupta, Samy Tafasca, Jean-Marc Odobez

Comments In the proceedings of the GAZE workshop at CVPR 2022

详情

展开后加载摘要…

URL PDF HTML 收藏
2303.13223 2023-07-12 cs.CV

Exploring Structured Semantic Prior for Multi Label Recognition with Incomplete Labels

Zixuan Ding, Ao Wang, Hui Chen, Qiang Zhang, Pengzhang Liu, Yongjun Bao, Weipeng Yan, Jungong Han

Journal ref IEEE/CVF Conference on Computer Vision and Pattern Recognition 2023

详情

展开后加载摘要…

URL PDF HTML 收藏
2307.04916 2023-07-12 cs.CV eess.IV

Rapid Deforestation and Burned Area Detection using Deep Multimodal Learning on Satellite Imagery

Gabor Fodor, Marcos V. Conde

Comments CVPR 2023 Workshop on Multimodal Learning for Earth and Environment (MultiEarth)

详情

展开后加载摘要…

URL PDF HTML 收藏
2106.08596 2023-07-12 cs.CV cs.HC

Temporal Convolution Networks with Positional Encoding for Evoked Expression Estimation

VanThong Huynh, Guee-Sang Lee, Hyung-Jeong Yang, Soo-Huyng Kim

Comments Oral presentation at AUVi Workshop - CVPR 2021 (https://sites.google.com/view/auvi-cvpr2021/program). Source code available at https://github.com/th2l/EvokedExpression-tcnpe

详情

展开后加载摘要…

URL PDF HTML 收藏
2307.04392 2023-07-11 cs.CV

FODVid: Flow-guided Object Discovery in Videos

Silky Singh, Shripad Deshmukh, Mausoom Sarkar, Rishabh Jain, Mayur Hemani, Balaji Krishnamurthy

Comments CVPR 2023 (L3D-IVU workshop)

详情

展开后加载摘要…

URL PDF HTML 收藏
2307.04189 2023-07-11 cs.CV

Histopathology Whole Slide Image Analysis with Heterogeneous Graph Representation Learning

Tsai Hor Chan, Fernando Julio Cendra, Lan Ma, Guosheng Yin, Lequan Yu

Comments Accepted by CVPR 2023

详情

展开后加载摘要…

URL PDF HTML 收藏
2303.16897 2023-07-11 cs.CV cs.LG cs.SD eess.AS

Physics-Driven Diffusion Models for Impact Sound Synthesis from Videos

Kun Su, Kaizhi Qian, Eli Shlizerman, Antonio Torralba, Chuang Gan

Comments CVPR 2023. Project page: https://sukun1045.github.io/video-physics-sound-diffusion/

详情

展开后加载摘要…

URL PDF HTML 收藏