arXivDaily arXiv每日学术速递 周一至周五更新
arXiv周末暂无论文更新,休息一下吧,周末愉快~~

期刊&会议

Conference on Computer Vision and Pattern Recognition · 会议 · Computer Vision

共收录 11877
2406.08292 2024-06-13 cs.CV

Outdoor Scene Extrapolation with Hierarchical Generative Cellular Automata

Dongsu Zhang, Francis Williams, Zan Gojcic, Karsten Kreis, Sanja Fidler, Young Min Kim, Amlan Kar

Comments Accepted to CVPR 2024 as highlight

详情

展开后加载摘要…

URL PDF HTML 收藏
2406.08192 2024-06-13 cs.CV

2nd Place Solution for MOSE Track in CVPR 2024 PVUW workshop: Complex Video Object Segmentation

Zhensong Xu, Jiangtao Yao, Chengjing Wu, Ting Liu, Luoqi Liu

Comments 5pages, 4 figures, technique report for MOSE Track in CVPR 2024 PVUW workshop: Complex Video Object Segmentation

详情

展开后加载摘要…

URL PDF HTML 收藏
2406.08089 2024-06-13 cs.CV

Identification of Conversation Partners from Egocentric Video

Tobias Dorszewski, Søren A. Fuglsang, Jens Hjortkjær

Comments First Joint Egocentric Vision (EgoVis) Workshop at CVPR 2024

详情

展开后加载摘要…

URL PDF HTML 收藏
2406.07876 2024-06-13 cs.CV cs.AI cs.LG

Small Scale Data-Free Knowledge Distillation

He Liu, Yikai Wang, Huaping Liu, Fuchun Sun, Anbang Yao

Comments This work is accepted to CVPR 2024. The project page: https://github.com/OSVAI/SSD-KD

详情

展开后加载摘要…

URL PDF HTML 收藏
2406.07792 2024-06-13 cs.CV

Hierarchical Patch Diffusion Models for High-Resolution Video Generation

Ivan Skorokhodov, Willi Menapace, Aliaksandr Siarohin, Sergey Tulyakov

Comments CVPR 2024

详情

展开后加载摘要…

URL PDF HTML 收藏
2406.07785 2024-06-13 cs.CV cs.LG

From Variance to Veracity: Unbundling and Mitigating Gradient Variance in Differentiable Bundle Adjustment Layers

Swaminathan Gurumurthy, Karnik Ram, Bingqing Chen, Zachary Manchester, Zico Kolter

Comments Accepted at CVPR 2024

详情

展开后加载摘要…

URL PDF HTML 收藏
2406.07763 2024-06-13 eess.IV cs.CV

Gene-Level Representation Learning via Interventional Style Transfer in Optical Pooled Screening

Mahtab Bigverdi, Burkhard Hockendorf, Heming Yao, Phil Hanslovsky, Romain Lopez, David Richmond

Comments 11 pages, 5 figures, CVPR workshop paper

详情

展开后加载摘要…

URL PDF HTML 收藏
2406.07738 2024-06-13 cs.CV

On the Application of Egocentric Computer Vision to Industrial Scenarios

Vivek Chavan, Oliver Heimann, Jörg Krüger

Comments To be presented at the First Joint Egocentric Vision (EgoVis) Workshop, held in conjunction with CVPR 2024

详情

展开后加载摘要…

URL PDF HTML 收藏
2406.00123 2024-06-13 eess.IV cs.CV

Correlation-aware Coarse-to-fine MLPs for Deformable Medical Image Registration

Mingyuan Meng, Dagan Feng, Lei Bi, Jinman Kim

Comments Accepted at CVPR2024 as Oral Presentation && Best Paper Candidate

Journal ref Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition (CVPR), 2024, pp. 9645-9654

详情

展开后加载摘要…

URL PDF HTML 收藏
2404.14027 2024-06-13 cs.CV cs.LG

OccFeat: Self-supervised Occupancy Feature Prediction for Pretraining BEV Segmentation Networks

Sophia Sirko-Galouchenko, Alexandre Boulch, Spyros Gidaris, Andrei Bursuc, Antonin Vobecky, Patrick Pérez, Renaud Marlet

Comments Accepted to CVPR 2024, Workshop on Autonomous Driving

详情

展开后加载摘要…

URL PDF HTML 收藏
2406.07551 2024-06-12 cs.CV

Blur-aware Spatio-temporal Sparse Transformer for Video Deblurring

Huicong Zhang, Haozhe Xie, Hongxun Yao

Comments CVPR 2024

详情

展开后加载摘要…

URL PDF HTML 收藏
2406.07532 2024-06-12 cs.SD cs.CV cs.LG eess.AS

Hearing Anything Anywhere

Mason Wang, Ryosuke Sawata, Samuel Clarke, Ruohan Gao, Shangzhe Wu, Jiajun Wu

Comments CVPR 2024. The first two authors contributed equally. Project page: https://masonlwang.com/hearinganythinganywhere/

详情

展开后加载摘要…

URL PDF HTML 收藏
2406.07061 2024-06-12 eess.IV cs.CV

Triage of 3D pathology data via 2.5D multiple-instance learning to guide pathologist assessments

Gan Gao, Andrew H. Song, Fiona Wang, David Brenes, Rui Wang, Sarah S. L. Chow, Kevin W. Bishop, Lawrence D. True, Faisal Mahmood, Jonathan T. C. Liu

Comments CVPR CVMI 2024

Journal ref Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition (CVPR) Workshops, 2024, pp. 6955-6965

详情

展开后加载摘要…

URL PDF HTML 收藏
2406.07006 2024-06-12 cs.CV

MIPI 2024 Challenge on Few-shot RAW Image Denoising: Methods and Results

Xin Jin, Chunle Guo, Xiaoming Li, Zongsheng Yue, Chongyi Li, Shangchen Zhou, Ruicheng Feng, Yuekun Dai, Peiqing Yang, Chen Change Loy, Ruoqi Li, Chang Liu, Ziyi Wang, Yao Du, Jingjing Yang, Long Bao, Heng Sun, Xiangyu Kong, Xiaoxia Xing, Jinlong Wu, Yuanyang Xue, Hyunhee Park, Sejun Song, Changho Kim, Jingfan Tan, Wenhan Luo, Zikun Liu, Mingde Qiao, Junjun Jiang, Kui Jiang, Yao Xiao, Chuyang Sun, Jinhui Hu, Weijian Ruan, Yubo Dong, Kai Chen, Hyejeong Jo, Jiahao Qin, Bingjie Han, Pinle Qin, Rui Chai, Pengyuan Wang

Comments CVPR 2024 Mobile Intelligent Photography and Imaging (MIPI) Workshop--Few-shot RAWImage Denoising Challenge Report. Website: https://mipi-challenge.org/MIPI2024/

详情

展开后加载摘要…

URL PDF HTML 收藏
2406.06820 2024-06-12 cs.CV cs.LG

Adapters Strike Back

Jan-Martin O. Steitz, Stefan Roth

Comments To appear at CVPR 2024. Code: https://github.com/visinf/adapter_plus

详情

展开后加载摘要…

URL PDF HTML 收藏
2406.06813 2024-06-12 cs.CV

Stable Neighbor Denoising for Source-free Domain Adaptive Segmentation

Dong Zhao, Shuang Wang, Qi Zang, Licheng Jiao, Nicu Sebe, Zhun Zhong

Comments 2024 Conference on Computer Vision and Pattern Recognition

Journal ref (2024 Conference on Computer Vision and Pattern Recognition)

详情

展开后加载摘要…

URL PDF HTML 收藏
2406.06730 2024-06-12 cs.CV cs.AI

TRINS: Towards Multimodal Language Models that Can Read

Ruiyi Zhang, Yanzhe Zhang, Jian Chen, Yufan Zhou, Jiuxiang Gu, Changyou Chen, Tong Sun

Comments CVPR 2024

详情

展开后加载摘要…

URL PDF HTML 收藏
2406.03461 2024-06-12 cs.CV eess.IV

Polarization Wavefront Lidar: Learning Large Scene Reconstruction from Polarized Wavefronts

Dominik Scheuble, Chenyang Lei, Seung-Hwan Baek, Mario Bijelic, Felix Heide

Comments Accepted at CVPR 2024; Project Website: https://light.princeton.edu/publication/pollidar

详情

展开后加载摘要…

URL PDF HTML 收藏
2404.01331 2024-06-12 cs.CL cs.AI

LLaVA-Gemma: Accelerating Multimodal Foundation Models with a Compact Language Model

Musashi Hinck, Matthew L. Olson, David Cobbley, Shao-Yen Tseng, Vasudev Lal

Comments CVPR 2024, MMFM workshop. Authors 1 and 2 contributed equally. Models available at https://huggingface.co/intel/llava-gemma-2b/ and https://huggingface.co/intel/llava-gemma-7b/ Training code at https://github.com/IntelLabs/multimodal_cognitive_ai/tree/main/LLaVA-Gemma

详情

展开后加载摘要…

URL PDF HTML 收藏
2403.01444 2024-06-12 cs.CV

3DGStream: On-the-Fly Training of 3D Gaussians for Efficient Streaming of Photo-Realistic Free-Viewpoint Videos

Jiakai Sun, Han Jiao, Guangyuan Li, Zhanjie Zhang, Lei Zhao, Wei Xing

Comments CVPR 2024 Accepted (Highlight). Project Page: https://sjojok.github.io/3dgstream

详情

展开后加载摘要…

URL PDF HTML 收藏
2312.00093 2024-06-12 cs.CV cs.GR cs.LG

GraphDreamer: Compositional 3D Scene Synthesis from Scene Graphs

Gege Gao, Weiyang Liu, Anpei Chen, Andreas Geiger, Bernhard Schölkopf

Comments CVPR 2024 (18 pages, 11 figures, https://graphdreamer.github.io/)

详情

展开后加载摘要…

URL PDF HTML 收藏
2406.06386 2024-06-11 cs.CV

FPN-IAIA-BL: A Multi-Scale Interpretable Deep Learning Model for Classification of Mass Margins in Digital Mammography

Julia Yang, Alina Jade Barnett, Jon Donnelly, Satvik Kishore, Jerry Fang, Fides Regina Schwartz, Chaofan Chen, Joseph Y. Lo, Cynthia Rudin

Comments 8 pages, 6 figures, Accepted for oral presentation at the 2024 CVPR Workshop on Domain adaptation, Explainability, Fairness in AI for Medical Image Analysis (DEF-AI-MIA)

详情

展开后加载摘要…

URL PDF HTML 收藏
2406.06352 2024-06-11 cs.CV

Latent Directions: A Simple Pathway to Bias Mitigation in Generative AI

Carolina Lopez Olmos, Alexandros Neophytou, Sunando Sengupta, Dim P. Papadopoulos

Comments Accepted at CVPR workshop 2024, proceedings of ReGenAI: First Workshop on Responsible Generative AI

详情

展开后加载摘要…

URL PDF HTML 收藏
2312.09558 2024-06-11 cs.CV

Towards Transferable Targeted 3D Adversarial Attack in the Physical World

Yao Huang, Yinpeng Dong, Shouwei Ruan, Xiao Yang, Hang Su, Xingxing Wei

Comments Accepted by CVPR 2024

详情

展开后加载摘要…

URL PDF HTML 收藏
2406.06264 2024-06-11 cs.CV

DualAD: Disentangling the Dynamic and Static World for End-to-End Driving

Simon Doll, Niklas Hanselmann, Lukas Schneider, Richard Schulz, Marius Cordts, Markus Enzweiler, Hendrik P. A. Lensch

Comments Accepted at CVPR 2024; Copyright 2024 IEEE; Project Website: https://simondoll.github.io/publications/dualad

详情

展开后加载摘要…

URL PDF HTML 收藏
2406.06134 2024-06-11 cs.CV cs.AI cs.LG

DiffInject: Revisiting Debias via Synthetic Data Generation using Diffusion-based Style Injection

Donggeun Ko, Sangwoo Jo, Dongjun Lee, Namjun Park, Jaekwang Kim

Comments 10 pages (including supplementary), 3 figures, SynData4CV@CVPR 24 (Workshop)

详情

展开后加载摘要…

URL PDF HTML 收藏
2406.05871 2024-06-11 cs.CV cs.LG

OmniControlNet: Dual-stage Integration for Conditional Image Generation

Yilin Wang, Haiyang Xu, Xiang Zhang, Zeyuan Chen, Zhizhou Sha, Zirui Wang, Zhuowen Tu

Comments Accepted to CVPR 2024 Workshop: Generative Models for Computer Vision

详情

展开后加载摘要…

URL PDF HTML 收藏
2406.05850 2024-06-11 cs.CV cs.LG

Scaling Graph Convolutions for Mobile Vision

William Avery, Mustafa Munir, Radu Marculescu

Comments Proceedings of the 2024 IEEE/CVF Conference on Computer Vision and Pattern Recognition (CVPR) Workshops

详情

展开后加载摘要…

URL PDF HTML 收藏
2406.05837 2024-06-11 cs.CV cs.AI

Solution for CVPR 2024 UG2+ Challenge Track on All Weather Semantic Segmentation

Jun Yu, Yunxiang Zhang, Fengzhao Sun, Leilei Wang, Renjie Lu

Comments Solution for CVPR 2024 UG2+ Challenge Track on All Weather Semantic Segmentation

详情

展开后加载摘要…

URL PDF HTML 收藏
2406.05722 2024-06-11 cs.CV

ALGO: Object-Grounded Visual Commonsense Reasoning for Open-World Egocentric Action Recognition

Sanjoy Kundu, Shubham Trehan, Sathyanarayanan N. Aakur

Comments Extended abstract of arXiv:2305.16602 for CVPR EgoVis Workshop

详情

展开后加载摘要…

URL PDF HTML 收藏