arXivDaily arXiv每日学术速递 周一至周五更新

期刊&会议

Conference on Computer Vision and Pattern Recognition · 会议 · Computer Vision

共收录 11877
2412.05507 2025-05-15 cs.RO cs.CV

AutoURDF: Unsupervised Robot Modeling from Point Cloud Frames Using Cluster Registration

Jiong Lin, Lechen Zhang, Kwansoo Lee, Jialong Ning, Judah Goldfeder, Hod Lipson

机构 * Columbia University(哥伦比亚大学) Creative Machines Lab(创意机器实验室)

Comments 16 pages

Journal ref CVPR 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2403.03370 2025-05-15 cs.CV cs.RO

F$^3$Loc: Fusion and Filtering for Floorplan Localization

Changan Chen, Rui Wang, Christoph Vogel, Marc Pollefeys

机构 * ETH Zürich(苏黎世联邦理工学院) Microsoft Mixed Reality & AI Lab(微软混合现实与人工智能实验室)

Comments 10 pages, 11 figure, accepted to CVPR 2024 (fixed typo eq.8: s_x,s_y, s_phi -> x, y, phi)

详情

展开后加载摘要…

URL PDF HTML 收藏
2504.02870 2025-05-14 cs.CL cs.AI

AI Hiring with LLMs: A Context-Aware and Explainable Multi-Agent Framework for Resume Screening

Frank P. -W. Lo, Jianing Qiu, Zeyu Wang, Haibao Yu, Yeming Chen, Gao Zhang, Benny Lo

机构 * Imperial College London(伦敦帝国理工学院) The Chinese University of Hong Kong(香港中文大学) The University of Hong Kong(香港大学) Wedon Education Technologies(韦登教育科技) Brest Business School(布雷斯特商学院)

Comments Accepted by CVPR 2025 Workshop

详情

展开后加载摘要…

URL PDF HTML 收藏
2505.08255 2025-05-14 cs.CR cs.CV

Where the Devil Hides: Deepfake Detectors Can No Longer Be Trusted

Shuaiwei Yuan, Junyu Dong, Yuezun Li

机构 * School of Computer Science and Technology, Ocean University of China(中国海洋大学计算机科学与技术学院)

Comments CVPR 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2504.21435 2025-05-14 cs.CV cs.AI cs.CL

SeriesBench: A Benchmark for Narrative-Driven Drama Series Understanding

Chenkai Zhang, Yiming Lei, Zeming Liu, Haitao Leng, Shaoguo Liu, Tingting Gao, Qingjie Liu, Yunhong Wang

机构 * State Key Laboratory of Virtual Reality Technology and Systems, Beihang University(虚拟现实技术与系统国家重点实验室,北京航空航天大学) School of Computer Science and Engineering, Beihang University(北京航空航天大学计算机科学与工程学院) Hangzhou Innovation Institute, Beihang University(杭州创新研究院) Kuaishou Technology(快手科技)

Comments 29 pages, 15 figures, CVPR 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2504.02697 2025-05-14 cs.CV eess.IV

Learning Phase Distortion with Selective State Space Models for Video Turbulence Mitigation

Xingguang Zhang, Nicholas Chimitt, Xijun Wang, Yu Yuan, Stanley H. Chan

机构 * School of Electrical and Computer Engineering, Purdue University(电子与计算机工程学院,普渡大学)

Comments CVPR 2025 Highlight (extended), project page: https://xg416.github.io/MambaTM/

详情

展开后加载摘要…

URL PDF HTML 收藏
2501.09425 2025-05-14 cs.CV cs.CL

Vision-Language Models Do Not Understand Negation

Kumail Alhamoud, Shaden Alshammari, Yonglong Tian, Guohao Li, Philip Torr, Yoon Kim, Marzyeh Ghassemi

机构 * Institution1(机构1) Institution2(机构2)

Comments CVPR 2025; project page: https://negbench.github.io

详情

展开后加载摘要…

URL PDF HTML 收藏
2505.07652 2025-05-13 cs.CV

ShotAdapter: Text-to-Multi-Shot Video Generation with Diffusion Models

Ozgur Kara, Krishna Kumar Singh, Feng Liu, Duygu Ceylan, James M. Rehg, Tobias Hinz

机构 * UIUC(伊利诺伊大学香槟分校) Adobe(Adobe公司)

Comments CVPR 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2505.07333 2025-05-13 cs.CV

Link to the Past: Temporal Propagation for Fast 3D Human Reconstruction from Monocular Video

Matthew Marchellus, Nadhira Noor, In Kyu Park

机构 * Department of Electrical and Computer Engineering, Inha University(电子与计算机工程系,釜山大学)

Comments Accepted in CVPR 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2505.07300 2025-05-13 cs.CV

L-SWAG: Layer-Sample Wise Activation with Gradients information for Zero-Shot NAS on Vision Transformers

Sofia Casarin, Sergio Escalera, Oswald Lanz

机构 * Free University of Bozen-Bolzano(博兹纳-博尔扎诺自由大学) Computer Vision Center(计算机视觉中心) Universitat de Barcelona(巴塞罗那大学)

Comments accepted at CVPR 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2310.14356 2025-05-13 cs.CV cs.CL cs.CY cs.HC

Semantic and Expressive Variation in Image Captions Across Languages

Andre Ye, Sebastin Santy, Jena D. Hwang, Amy X. Zhang, Ranjay Krishna

机构 * University of Washington(华盛顿大学) Allen Institute for AI(人工智能研究院)

Comments CVPR 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2505.07209 2025-05-13 cs.CV

Discovering Fine-Grained Visual-Concept Relations by Disentangled Optimal Transport Concept Bottleneck Models

Yan Xie, Zequn Zeng, Hao Zhang, Yucheng Ding, Yi Wang, Zhengjue Wang, Bo Chen, Hongwei Liu

机构 * National Key Laboratory of Radar Signal Processing, Xidian University, Xi’an, 710071, China(雷达信号处理国家级重点实验室,西安电子科技大学) State Key Laboratory of Integrated Service Networks, Xidian University, Xi’an, 710071, China(集成服务网络国家重点实验室,西安电子科技大学)

Comments CVPR 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2505.07164 2025-05-13 cs.MM

EmoVLM-KD: Fusing Distilled Expertise with Vision-Language Models for Visual Emotion Analysis

SangEun Lee, Yubeen Lee, Eunil Park

Comments Accepted at Workshop and Competition on Affective & Behavior Analysis in-the-wild (ABAW), CVPR 2025, 10 pages, 4 figures, 4 tables

详情

展开后加载摘要…

URL PDF HTML 收藏
2505.06580 2025-05-13 cs.AI stat.ML

TAROT: Towards Essentially Domain-Invariant Robustness with Theoretical Justification

Dongyoon Yang, Jihu Lee, Yongdai Kim

机构 * AI Advanced Technology, SK Hynix(SK Hynix人工智能高级技术) Department of Statistics, Seoul National University(首尔国立大学统计系)

Comments Accepted in CVPR 2025 (19 pages, 7 figures)

详情

展开后加载摘要…

URL PDF HTML 收藏
2505.06276 2025-05-13 cs.RO

SynSHRP2: A Synthetic Multimodal Benchmark for Driving Safety-critical Events Derived from Real-world Driving Data

Liang Shi, Boyu Jiang, Zhenyuan Yuan, Miguel A. Perez, Feng Guo

机构 * Virginia Tech Transportation Institute(弗吉尼亚理工运输研究所) Department of Statistics, Virginia Tech(弗吉尼亚理工统计系) Department of Biomedical Engineering and Mechanics, Virginia Tech(弗吉尼亚理工生物医学工程与力学系)

Comments Accepted as a poster in CVPR 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2504.15118 2025-05-13 cs.CV cs.SD

Improving Sound Source Localization with Joint Slot Attention on Image and Audio

Inho Kim, Youngkil Song, Jicheol Park, Won Hwa Kim, Suha Kwak

机构 * Dept. of CSE, POSTECH(计算机科学与工程系,POSTECH)

Comments Accepted to CVPR 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2503.15816 2025-05-13 cs.CV

A Vision Centric Remote Sensing Benchmark

Abduljaleel Adejumo, Faegheh Yeganli, Clifford Broni-bediako, Aoran Xiao, Naoto Yokoya, Mennatullah Siam

机构 * AMMI/AIMS Senegal(AMMI/AIMS塞内加尔) University of British Columbia(不列颠哥伦比亚大学) RIKEN AIP(理化学研究所AIP) the University of Tokyo(东京大学)

Comments Eval-FoMo2 Workshop in CVPR 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2505.06218 2025-05-12 cs.RO cs.AI cs.CV

Let Humanoids Hike! Integrative Skill Development on Complex Trails

Kwan-Yee Lin, Stella X. Yu

机构 * University of Michigan(密歇根大学)

Comments CVPR 2025. Project page: https://lego-h-humanoidrobothiking.github.io/

详情

展开后加载摘要…

URL PDF HTML 收藏
2505.06166 2025-05-12 cs.CV

DiffLocks: Generating 3D Hair from a Single Image using Diffusion Models

Radu Alexandru Rosu, Keyu Wu, Yao Feng, Youyi Zheng, Michael J. Black

机构 * Meshcapade Zhejiang University(浙江大学) Stanford University(斯坦福大学) Max Planck Institute for Intelligent Systems(马克斯·普朗克智能系统研究所)

Comments Accepted to CVPR 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2505.05868 2025-05-12 cs.LG

Open Set Label Shift with Test Time Out-of-Distribution Reference

Changkun Ye, Russell Tsuchida, Lars Petersson, Nick Barnes

机构 * Australian National University(澳大利亚国立大学) Data61 CSIRO Data Science and AI Group, Monash University(墨尔本大学数据科学与人工智能小组)

Comments Accepted at CVPR 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2505.05853 2025-05-12 cs.CV

PICD: Versatile Perceptual Image Compression with Diffusion Rendering

Tongda Xu, Jiahao Li, Bin Li, Yan Wang, Ya-Qin Zhang, Yan Lu

机构 * AIR, Tsinghua University(清华大学人工智能研究院) Microsoft Research Asia(微软亚洲研究院)

Comments CVPR 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2505.05722 2025-05-12 cs.CV

You Are Your Best Teacher: Semi-Supervised Surgical Point Tracking with Cycle-Consistent Self-Distillation

Valay Bundele, Mehran Hosseinzadeh, Hendrik Lensch

机构 * University of Tübingen(图宾根大学)

Comments Accepted at CVPR 2025 SynData4CV Workshop

详情

展开后加载摘要…

URL PDF HTML 收藏
2505.05711 2025-05-12 cs.CV

DiGIT: Multi-Dilated Gated Encoder and Central-Adjacent Region Integrated Decoder for Temporal Action Detection Transformer

Ho-Joong Kim, Yearang Lee, Jung-Ho Hong, Seong-Whan Lee

机构 * Dept. of Artificial Intelligence, Korea University(人工智能系,韩国大学)

Comments CVPR 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2505.05587 2025-05-12 cs.CV

Steepest Descent Density Control for Compact 3D Gaussian Splatting

Peihao Wang, Yuehao Wang, Dilin Wang, Sreyas Mohan, Zhiwen Fan, Lemeng Wu, Ruisi Cai, Yu-Ying Yeh, Zhangyang Wang, Qiang Liu, Rakesh Ranjan

机构 * The University of Texas at Austin(德克萨斯大学奥斯汀分校) Meta Reality Labs(Meta现实实验室)

Comments CVPR 2025, Project page: https://vita-group.github.io/SteepGS/

详情

展开后加载摘要…

URL PDF HTML 收藏
2503.11341 2025-05-12 cs.CV

Self-Supervised Pretraining for Fine-Grained Plankton Recognition

Joona Kareinen, Tuomas Eerola, Kaisa Kraft, Lasse Lensu, Sanna Suikkanen, Heikki Kälviäinen

机构 * LUT University, Computer Vision and Pattern Recognition Laboratory(卢霍斯大学) Finnish Environment Institute(芬兰环境研究所) Brno University of Technology, Faculty of Information Technology(布拉格技术大学)

Comments CVPR 2025, FGVC12 workshop paper

详情

展开后加载摘要…

URL PDF HTML 收藏
2412.04378 2025-05-12 cs.CV cs.AI

VladVA: Discriminative Fine-tuning of LVLMs

Yassine Ouali, Adrian Bulat, Alexandros Xenos, Anestis Zaganidis, Ioannis Maniadis Metaxas, Brais Martinez, Georgios Tzimiropoulos

机构 * Samsung AI Cambridge(三星AI剑桥) Technical University of Iasi(伊阿西技术大学) Queen Mary University of London(伦敦女王学院)

Comments Published at CVPR 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2405.13637 2025-05-12 cs.CV cs.AI cs.LG

Curriculum Direct Preference Optimization for Diffusion and Consistency Models

Florinel-Alin Croitoru, Vlad Hondru, Radu Tudor Ionescu, Nicu Sebe, Mubarak Shah

机构 * University of Bucharest(布加勒斯特大学) University of Trento(特伦托大学) University of Central Florida(中央佛罗里达大学)

Comments Accepted at CVPR 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2505.05475 2025-05-09 cs.CV

SVAD: From Single Image to 3D Avatar via Synthetic Data Generation with Video Diffusion and Data Augmentation

Yonwoo Choi

机构 * SECERN AI

Comments Accepted by CVPR 2025 SyntaGen Workshop, Project Page: https://yc4ny.github.io/SVAD/

详情

展开后加载摘要…

URL PDF HTML 收藏
2505.05473 2025-05-09 cs.CV

DiffusionSfM: Predicting Structure and Motion via Ray Origin and Endpoint Diffusion

Qitao Zhao, Amy Lin, Jeff Tan, Jason Y. Zhang, Deva Ramanan, Shubham Tulsiani

机构 * Carnegie Mellon University(卡内基梅隆大学)

Comments CVPR 2025. Project website: https://qitaozhao.github.io/DiffusionSfM

详情

展开后加载摘要…

URL PDF HTML 收藏
2505.05309 2025-05-09 eess.IV cs.CV

Augmented Deep Contexts for Spatially Embedded Video Coding

Yifan Bian, Chuanbo Tang, Li Li, Dong Liu

机构 * MOE Key Laboratory of Brain-Inspired Intelligent Perception and Cognition(脑启发式智能感知与认知实验室) University of Science and Technology of China(中国科学技术大学)

Comments 15 pages,CVPR

详情

展开后加载摘要…

URL PDF HTML 收藏