arXivDaily arXiv每日学术速递 周一至周五更新

期刊&会议

Conference on Computer Vision and Pattern Recognition · 会议 · Computer Vision

共收录 11877
2503.18794 2025-03-25 cs.CV

NexusGS: Sparse View Synthesis with Epipolar Depth Priors in 3D Gaussian Splatting

Yulong Zheng, Zicheng Jiang, Shengfeng He, Yandu Sun, Junyu Dong, Huaidong Zhang, Yong Du

Comments This paper is accepted by CVPR 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2503.18746 2025-03-25 cs.CV

Linguistics-aware Masked Image Modeling for Self-supervised Scene Text Recognition

Yifei Zhang, Chang Liu, Jin Wei, Xiaomeng Yang, Yu Zhou, Can Ma, Xiangyang Ji

Comments CVPR 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2503.18695 2025-03-25 cs.CV cs.LG

OCRT: Boosting Foundation Models in the Open World with Object-Concept-Relation Triad

Luyao Tang, Yuxuan Yuan, Chaoqi Chen, Zeyu Zhang, Yue Huang, Kun Zhang

Comments Accepted by CVPR 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2503.18637 2025-03-25 cs.CV

Unbiasing through Textual Descriptions: Mitigating Representation Bias in Video Benchmarks

Nina Shvetsova, Arsha Nagrani, Bernt Schiele, Hilde Kuehne, Christian Rupprecht

Comments To be published at CVPR 2025, project webpage https://utd-project.github.io/

详情

展开后加载摘要…

URL PDF HTML 收藏
2503.18512 2025-03-25 cs.CV eess.IV

Uncertainty-guided Perturbation for Image Super-Resolution Diffusion Model

Leheng Zhang, Weiyi You, Kexuan Shi, Shuhang Gu

Comments Accepted to CVPR 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2503.18507 2025-03-25 cs.CV

Can Text-to-Video Generation help Video-Language Alignment?

Luca Zanella, Massimiliano Mancini, Willi Menapace, Sergey Tulyakov, Yiming Wang, Elisa Ricci

Comments CVPR 2025. Project website at https://lucazanella.github.io/synvita/

详情

展开后加载摘要…

URL PDF HTML 收藏
2503.18503 2025-03-25 cs.LG cs.CR

Deterministic Certification of Graph Neural Networks against Graph Poisoning Attacks with Arbitrary Perturbations

Jiate Li, Meng Pang, Yun Dong, Binghui Wang

Comments Accepted at CVPR 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2503.15910 2025-03-25 cs.CV cs.AI

No Thing, Nothing: Highlighting Safety-Critical Classes for Robust LiDAR Semantic Segmentation in Adverse Weather

Junsung Park, Hwijeong Lee, Inha Kang, Hyunjung Shim

Comments 18 pages, accepted in CVPR 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2503.14405 2025-03-25 cs.CV cs.LG

DUNE: Distilling a Universal Encoder from Heterogeneous 2D and 3D Teachers

Mert Bulent Sariyildiz, Philippe Weinzaepfel, Thomas Lucas, Pau de Jorge, Diane Larlus, Yannis Kalantidis

Comments Accepted to CVPR-2025. Project page: https://europe.naverlabs.com/dune

详情

展开后加载摘要…

URL PDF HTML 收藏
2502.19908 2025-03-25 cs.RO cs.CV cs.LG

CarPlanner: Consistent Auto-regressive Trajectory Planning for Large-scale Reinforcement Learning in Autonomous Driving

Dongkun Zhang, Jiaming Liang, Ke Guo, Sha Lu, Qi Wang, Rong Xiong, Zhenwei Miao, Yue Wang

Comments CVPR 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2412.18883 2025-03-25 cs.CV eess.IV

MotionMap: Representing Multimodality in Human Pose Forecasting

Reyhaneh Hosseininejad, Megh Shukla, Saeed Saadatnejad, Mathieu Salzmann, Alexandre Alahi

Comments CVPR 2025. We propose a new representation for learning multimodality in human pose forecasting which does not depend on generative models

详情

展开后加载摘要…

URL PDF HTML 收藏
2412.17808 2025-03-25 cs.CV

Dora: Sampling and Benchmarking for 3D Shape Variational Auto-Encoders

Rui Chen, Jianfeng Zhang, Yixun Liang, Guan Luo, Weiyu Li, Jiarui Liu, Xiu Li, Xiaoxiao Long, Jiashi Feng, Ping Tan

Comments Accepted by CVPR 2025. Project page: https://aruichen.github.io/Dora/

详情

展开后加载摘要…

URL PDF HTML 收藏
2412.13193 2025-03-25 cs.CV

GaussTR: Foundation Model-Aligned Gaussian Transformer for Self-Supervised 3D Spatial Understanding

Haoyi Jiang, Liu Liu, Tianheng Cheng, Xinjie Wang, Tianwei Lin, Zhizhong Su, Wenyu Liu, Xinggang Wang

Comments CVPR 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2412.12725 2025-03-25 cs.CV

RaCFormer: Towards High-Quality 3D Object Detection via Query-based Radar-Camera Fusion

Xiaomeng Chu, Jiajun Deng, Guoliang You, Yifan Duan, Houqiang Li, Yanyong Zhang

Comments Accepted to CVPR 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2412.03240 2025-03-25 cs.CV

Task-driven Image Fusion with Learnable Fusion Loss

Haowen Bai, Jiangshe Zhang, Zixiang Zhao, Yichen Wu, Lilun Deng, Yukun Cui, Tao Feng, Shuang Xu

Comments Accepted to CVPR 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2412.01820 2025-03-25 cs.CV

Towards Universal Soccer Video Understanding

Jiayuan Rao, Haoning Wu, Hao Jiang, Ya Zhang, Yanfeng Wang, Weidi Xie

Comments CVPR 2025; Project Page: https://jyrao.github.io/UniSoccer/

详情

展开后加载摘要…

URL PDF HTML 收藏
2412.00133 2025-03-25 cs.CV cs.LG cs.RO

ETAP: Event-based Tracking of Any Point

Friedhelm Hamann, Daniel Gehrig, Filbert Febryanto, Kostas Daniilidis, Guillermo Gallego

Comments 17 pages, 15 figures, 8 tables. Project page: https://github.com/tub-rip/ETAP

Journal ref IEEE Conference on Computer Vision and Pattern Recognition (CVPR), Nashville, 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2411.16683 2025-03-25 cs.CV

Generative Omnimatte: Learning to Decompose Video into Layers

Yao-Chih Lee, Erika Lu, Sarah Rumbley, Michal Geyer, Jia-Bin Huang, Tali Dekel, Forrester Cole

Comments CVPR 2025. Project page: https://gen-omnimatte.github.io/

详情

展开后加载摘要…

URL PDF HTML 收藏
2405.17811 2025-03-25 cs.GR cs.CV

Mani-GS: Gaussian Splatting Manipulation with Triangular Mesh

Xiangjun Gao, Xiaoyu Li, Yiyu Zhuang, Qi Zhang, Wenbo Hu, Chaopeng Zhang, Yao Yao, Ying Shan, Long Quan

Comments CVPR 2025. Project page here: https://gaoxiangjun.github.io/mani_gs/

详情

展开后加载摘要…

URL PDF HTML 收藏
2404.06842 2025-03-25 cs.CV

MoCha-Stereo: Motif Channel Attention Network for Stereo Matching

Ziyang Chen, Wei Long, He Yao, Yongjun Zhang, Bingshu Wang, Yongbin Qin, Jia Wu

Comments Accepted to CVPR 2024

Journal ref The IEEE/CVF Conference on Computer Vision and Pattern Recognition 2024

详情

展开后加载摘要…

URL PDF HTML 收藏
2403.16848 2025-03-25 cs.CV

Multiple Object Tracking as ID Prediction

Ruopeng Gao, Ji Qi, Limin Wang

Comments Accepted by CVPR 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2312.07352 2025-03-25 cs.CV cs.AI

CholecTrack20: A Multi-Perspective Tracking Dataset for Surgical Tools

Chinedu Innocent Nwoye, Kareem Elgohary, Anvita Srinivas, Fauzan Zaid, Joël L. Lavanchy, Nicolas Padoy

Comments Surgical tool tracking dataset paper, 11 pages, 10 figures, 3 tables, CVPR 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2503.18434 2025-03-25 cs.CV

A Simple yet Effective Layout Token in Large Language Models for Document Understanding

Zhaoqing Zhu, Chuwei Luo, Zirui Shao, Feiyu Gao, Hangdi Xing, Qi Zheng, Ji Zhang

Comments CVPR 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2503.18429 2025-03-25 cs.CV

Teller: Real-Time Streaming Audio-Driven Portrait Animation with Autoregressive Motion Generation

Dingcheng Zhen, Shunshun Yin, Shiyang Qin, Hou Yi, Ziwei Zhang, Siyuan Liu, Gan Qi, Ming Tao

Comments Accept in CVPR 2025 Conference Submission

详情

展开后加载摘要…

URL PDF HTML 收藏
2503.18371 2025-03-25 cs.CV cs.LG

Do Your Best and Get Enough Rest for Continual Learning

Hankyul Kang, Gregor Seifer, Donghyun Lee, Jongbin Ryu

Comments Accepted by IEEE/CVF Conference on Computer Vision and Pattern Recognition (CVPR) 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2503.18359 2025-03-25 cs.CV

Context-Enhanced Memory-Refined Transformer for Online Action Detection

Zhanzhong Pang, Fadime Sener, Angela Yao

Comments Accepted by CVPR 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2503.18328 2025-03-25 cs.CV

TensoFlow: Tensorial Flow-based Sampler for Inverse Rendering

Chun Gu, Xiaofei Wei, Li Zhang, Xiatian Zhu

Comments CVPR 2025. Code: https://github.com/fudan-zvg/tensoflow

详情

展开后加载摘要…

URL PDF HTML 收藏
2503.18325 2025-03-25 cs.CV

Towards Training-free Anomaly Detection with Vision and Language Foundation Models

Jinjin Zhang, Guodong Wang, Yizhou Jin, Di Huang

Comments Accepted to CVPR 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2503.18286 2025-03-25 cs.CV

CO-SPY: Combining Semantic and Pixel Features to Detect Synthetic Images by AI

Siyuan Cheng, Lingjuan Lyu, Zhenting Wang, Xiangyu Zhang, Vikash Sehwag

Journal ref The IEEE/CVF Conference on Computer Vision and Pattern Recognition 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2503.18267 2025-03-25 cs.CV

Enhancing Dataset Distillation via Non-Critical Region Refinement

Minh-Tuan Tran, Trung Le, Xuan-May Le, Thanh-Toan Do, Dinh Phung

Comments Accepted at CVPR 2025

详情

展开后加载摘要…

URL PDF HTML 收藏