arXivDaily arXiv每日学术速递 周一至周五更新

期刊&会议

Conference on Computer Vision and Pattern Recognition · 会议 · Computer Vision

共收录 11883
2203.01601 2022-06-08 cs.CV

Syntax-Aware Network for Handwritten Mathematical Expression Recognition

Ye Yuan, Xiao Liu, Wondimu Dikubab, Hui Liu, Zhilong Ji, Zhongqin Wu, Xiang Bai

Comments CVPR 2022

详情

展开后加载摘要…

URL PDF HTML 收藏
2206.02647 2022-06-07 cs.CV

Scaling Vision Transformers to Gigapixel Images via Hierarchical Self-Supervised Learning

Richard J. Chen, Chengkuan Chen, Yicong Li, Tiffany Y. Chen, Andrew D. Trister, Rahul G. Krishnan, Faisal Mahmood

Comments Accepted to CVPR 2022 (Oral)

详情

展开后加载摘要…

URL PDF HTML 收藏
2206.02187 2022-06-07 cs.CV cs.SD eess.AS

M2FNet: Multi-modal Fusion Network for Emotion Recognition in Conversation

Vishal Chudasama, Purbayan Kar, Ashish Gudmalwar, Nirmesh Shah, Pankaj Wasnik, Naoyuki Onoe

Comments Accepted for publication in the 5th Multimodal Learning and Applications (MULA) Workshop at CVPR 2022

详情

展开后加载摘要…

URL PDF HTML 收藏
2206.02163 2022-06-07 cs.CV

MotionCNN: A Strong Baseline for Motion Prediction in Autonomous Driving

Stepan Konev, Kirill Brodt, Artsiom Sanakoyeu

Comments CVPR Workshop on Autonomous Driving 2021. Waymo Motion Prediction Challenge 2021

详情

展开后加载摘要…

URL PDF HTML 收藏
2206.02116 2022-06-07 cs.CV

Cannot See the Forest for the Trees: Aggregating Multiple Viewpoints to Better Classify Objects in Videos

Sukjun Hwang, Miran Heo, Seoung Wug Oh, Seon Joo Kim

Comments Accepted to CVPR 2022

详情

展开后加载摘要…

URL PDF HTML 收藏
2206.02099 2022-06-07 cs.CV

Point-to-Voxel Knowledge Distillation for LiDAR Semantic Segmentation

Yuenan Hou, Xinge Zhu, Yuexin Ma, Chen Change Loy, Yikang Li

Comments CVPR 2022; Our model ranks 1st on Waymo and SemanticKITTI (single-scan) challenges, and ranks 3rd on SemanticKITTI (multi-scan) challenge; Code: https://github.com/cardwing/Codes-for-PVKD

详情

展开后加载摘要…

URL PDF HTML 收藏
2206.01988 2022-06-07 cs.CV

Cross-modal Clinical Graph Transformer for Ophthalmic Report Generation

Mingjie Li, Wenjia Cai, Karin Verspoor, Shirui Pan, Xiaodan Liang, Xiaojun Chang

Comments CVPR 2022 (Poster)

详情

展开后加载摘要…

URL PDF HTML 收藏
2206.01813 2022-06-07 cs.CV eess.IV

Learning sRGB-to-Raw-RGB De-rendering with Content-Aware Metadata

Seonghyeon Nam, Abhijith Punnappurath, Marcus A. Brubaker, Michael S. Brown

Comments CVPR 2022 (GitHub: https://github.com/SamsungLabs/content-aware-metadata)

详情

展开后加载摘要…

URL PDF HTML 收藏
2206.01738 2022-06-07 eess.IV cs.CV

RIDDLE: Lidar Data Compression with Range Image Deep Delta Encoding

Xuanyu Zhou, Charles R. Qi, Yin Zhou, Dragomir Anguelov

Comments 14 pages, 10 figures; CVPR 2022

详情

展开后加载摘要…

URL PDF HTML 收藏
2203.15118 2022-06-07 cs.CV cs.LG

LiDAR Snowfall Simulation for Robust 3D Object Detection

Martin Hahner, Christos Sakaridis, Mario Bijelic, Felix Heide, Fisher Yu, Dengxin Dai, Luc Van Gool

Comments Oral at CVPR 2022

详情

展开后加载摘要…

URL PDF HTML 收藏
2203.01502 2022-06-07 cs.CV

NeW CRFs: Neural Window Fully-connected CRFs for Monocular Depth Estimation

Weihao Yuan, Xiaodong Gu, Zuozhuo Dai, Siyu Zhu, Ping Tan

Comments Accepted by CVPR 2022

详情

展开后加载摘要…

URL PDF HTML 收藏
2111.14819 2022-06-07 cs.CV cs.AI cs.LG

Point-BERT: Pre-training 3D Point Cloud Transformers with Masked Point Modeling

Xumin Yu, Lulu Tang, Yongming Rao, Tiejun Huang, Jie Zhou, Jiwen Lu

Comments Accepted to CVPR 2022, Project page: https://point-bert.ivg-research.xyz

详情

展开后加载摘要…

URL PDF HTML 收藏
2206.01720 2022-06-06 cs.CV cs.AI cs.CL cs.LG

Revisiting the "Video" in Video-Language Understanding

Shyamal Buch, Cristóbal Eyzaguirre, Adrien Gaidon, Jiajun Wu, Li Fei-Fei, Juan Carlos Niebles

Comments CVPR 2022 (Oral)

详情

展开后加载摘要…

URL PDF HTML 收藏
2206.01690 2022-06-06 cs.LG cs.CV

Dynamic Kernel Selection for Improved Generalization and Memory Efficiency in Meta-learning

Arnav Chavan, Rishabh Tiwari, Udbhav Bamba, Deepak K. Gupta

Comments Published at CVPR 2022

详情

展开后加载摘要…

URL PDF HTML 收藏
1906.11713 2022-06-06 cs.CV cs.LG

GASP, a generalized framework for agglomerative clustering of signed graphs and its application to Instance Segmentation

Alberto Bailoni, Constantin Pape, Nathan Hütsch, Steffen Wolf, Thorsten Beier, Anna Kreshuk, Fred A. Hamprecht

Comments Published in CVPR 2022

详情

展开后加载摘要…

URL PDF HTML 收藏
2203.15143 2022-06-06 cs.CV

Towards End-to-End Unified Scene Text Detection and Layout Analysis

Shangbang Long, Siyang Qin, Dmitry Panteleev, Alessandro Bissacco, Yasuhisa Fujii, Michalis Raptis

Comments To appear at CVPR 2022. Code Available: https://github.com/tensorflow/models/tree/master/official/projects/unified_detector

详情

展开后加载摘要…

URL PDF HTML 收藏
2206.01103 2022-06-03 eess.IV cs.CV

Noise2NoiseFlow: Realistic Camera Noise Modeling without Clean Images

Ali Maleky, Shayan Kousha, Michael S. Brown, Marcus A. Brubaker

Comments CVPR 2022

详情

展开后加载摘要…

URL PDF HTML 收藏
2206.00923 2022-06-03 cs.CV

Modeling Image Composition for Complex Scene Generation

Zuopeng Yang, Daqing Liu, Chaoyue Wang, Jie Yang, Dacheng Tao

Comments CVPR 2022

详情

展开后加载摘要…

URL PDF HTML 收藏
2206.00845 2022-06-03 cs.LG cs.AI cs.CV

Hyperspherical Consistency Regularization

Cheng Tan, Zhangyang Gao, Lirong Wu, Siyuan Li, Stan Z. Li

Comments Accepted by CVPR 2022

详情

展开后加载摘要…

URL PDF HTML 收藏
2206.00812 2022-06-03 cs.CV eess.IV

Modeling sRGB Camera Noise with Normalizing Flows

Shayan Kousha, Ali Maleky, Michael S. Brown, Marcus A. Brubaker

Comments CVPR 2022

详情

展开后加载摘要…

URL PDF HTML 收藏
2206.00252 2022-06-03 cs.CV cs.AI cs.LG

Interpretable Deep Learning Classifier by Detection of Prototypical Parts on Kidney Stones Images

Daniel Flores-Araiza, Francisco Lopez-Tiro, Elias Villalvazo-Avila, Jonathan El-Beze, Jacques Hubert, Gilberto Ochoa-Ruiz, Christian Daul

Comments Extended abstract accepted at LatinX in Computer Vision Research Workshop, at CVPR 2022

详情

展开后加载摘要…

URL PDF HTML 收藏
2110.11524 2022-06-03 cs.CV cs.AI cs.LG

Sequential Voting with Relational Box Fields for Active Object Detection

Qichen Fu, Xingyu Liu, Kris M. Kitani

Comments In CVPR 2022. Project: https://fuqichen1998.github.io/SequentialVotingDet/

详情

展开后加载摘要…

URL PDF HTML 收藏
2108.05713 2022-06-03 cs.RO cs.AI cs.CV cs.LG

Towards real-world navigation with deep differentiable planners

Shu Ishida, João F. Henriques

Comments Published in CVPR 2022 (Conference on Computer Vision and Pattern Recognition)

详情

展开后加载摘要…

URL PDF HTML 收藏
2106.05682 2022-06-03 cs.CV cs.AI cs.LG

DASO: Distribution-Aware Semantics-Oriented Pseudo-label for Imbalanced Semi-Supervised Learning

Youngtaek Oh, Dong-Jin Kim, In So Kweon

Comments CVPR 2022; Project page: https://ytaek-oh.github.io/daso

详情

展开后加载摘要…

URL PDF HTML 收藏
2206.00608 2022-06-02 cs.CV cs.LG cs.RO

On the Choice of Data for Efficient Training and Validation of End-to-End Driving Models

Marvin Klingner, Konstantin Müller, Mona Mirzaie, Jasmin Breitenstein, Jan-Aike Termöhlen, Tim Fingscheidt

Comments Accepted at CVPR VDU Workshop 2022

详情

展开后加载摘要…

URL PDF HTML 收藏
2206.00393 2022-06-02 cs.SD cs.CV cs.LG cs.RO eess.AS

Towards Generalisable Audio Representations for Audio-Visual Navigation

Shunqi Mao, Chaoyi Zhang, Heng Wang, Weidong Cai

Comments CVPR 2022 Embodied AI Workshop

详情

展开后加载摘要…

URL PDF HTML 收藏
2206.00389 2022-06-02 eess.IV cs.CV cs.LG

A comparative study between vision transformers and CNNs in digital pathology

Luca Deininger, Bernhard Stimpel, Anil Yuce, Samaneh Abbasi-Sureshjani, Simon Schönenberger, Paolo Ocampo, Konstanty Korski, Fabien Gaire

Comments 8 pages, 2 figures, accepted for workshop T4Vision (CVPR 2022)

详情

展开后加载摘要…

URL PDF HTML 收藏
2206.00291 2022-06-02 cs.CV eess.IV

Efficient Multi-Purpose Cross-Attention Based Image Alignment Block for Edge Devices

Bahri Batuhan Bilecen, Alparslan Fisne, Mustafa Ayazoglu

Comments Accepted into Embedded Vision Workshop 2022 of CVPR 2022

详情

展开后加载摘要…

URL PDF HTML 收藏
2206.00100 2022-06-02 cs.CV cs.CL

VALHALLA: Visual Hallucination for Machine Translation

Yi Li, Rameswar Panda, Yoon Kim, Chun-Fu Chen, Rogerio Feris, David Cox, Nuno Vasconcelos

Comments CVPR 2022

详情

展开后加载摘要…

URL PDF HTML 收藏
2205.15749 2022-06-02 cs.LG cs.CV

Non-Iterative Recovery from Nonlinear Observations using Generative Models

Jiulong Liu, Zhaoqiang Liu

Comments CVPR 2022

详情

展开后加载摘要…

URL PDF HTML 收藏