arXivDaily arXiv每日学术速递 周一至周五更新

期刊&会议

Conference on Computer Vision and Pattern Recognition · 会议 · Computer Vision

共收录 11877
2505.05091 2025-05-09 cs.CV cs.LG

DispBench: Benchmarking Disparity Estimation to Synthetic Corruptions

Shashank Agnihotri, Amaan Ansari, Annika Dackermann, Fabian Rösch, Margret Keuper

机构 * Data and Web Science Group, University of Mannheim(曼海姆大学数据与网络科学小组) Max-Planck-Institute for Informatics, Saarland Informatics Campus(马克斯·普朗克研究所信息学研究所萨尔兰信息学校区)

Comments Accepted at CVPR 2025 Workshop on Synthetic Data for Computer Vision

详情

展开后加载摘要…

URL PDF HTML 收藏
2505.04917 2025-05-09 cs.CV

A Simple Detector with Frame Dynamics is a Strong Tracker

Chenxu Peng, Chenxu Wang, Minrui Zou, Danyang Li, Zhengpeng Yang, Yimian Dai, Ming-Ming Cheng, Xiang Li

机构 * VCIP, CS, Nankai University(南开大学计算机科学与技术学院)

Comments 2025 CVPR Anti-UAV Workshop

详情

展开后加载摘要…

URL PDF HTML 收藏
2505.04915 2025-05-09 cs.CV

GlyphMastero: A Glyph Encoder for High-Fidelity Scene Text Editing

Tong Wang, Ting Liu, Xiaochao Qu, Chengjing Wu, Luoqi Liu, Xiaolin Hu

机构 * MT Lab, Meitu Inc.(美图公司MT实验室) Department of Computer Science and Technology, BNRist, IDG/McGovern Institute for Brain Research, Tsinghua University(清华大学计算机科学与技术系)

Comments CVPR 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2505.04835 2025-05-09 cs.CV

Are Synthetic Corruptions A Reliable Proxy For Real-World Corruptions?

Shashank Agnihotri, David Schader, Nico Sharei, Mehmet Ege Kaçar, Margret Keuper

机构 * Data and Web Science Group, University of Mannheim(曼海姆大学数据与网络科学小组) Max-Planck-Institute for Informatics, Saarland Informatics Campus(马克斯·普朗克研究所信息学研究所萨尔兰信息学校区)

Comments Accepted at CVPR 2025 Workshop on Synthetic Data for Computer Vision

详情

展开后加载摘要…

URL PDF HTML 收藏
2505.04668 2025-05-09 cs.GR

SGCR: Spherical Gaussians for Efficient 3D Curve Reconstruction

Xinran Yang, Donghao Ji, Yuanqi Li, Jie Guo, Yanwen Guo, Junyuan Xie

Comments The IEEE/CVF Conference on Computer Vision and Pattern Recognition 2025, 8 pages

详情

展开后加载摘要…

URL PDF HTML 收藏
2505.04657 2025-05-09 eess.IV cs.MM

EvEnhancer: Empowering Effectiveness, Efficiency and Generalizability for Continuous Space-Time Video Super-Resolution with Events

Shuoyan Wei, Feng Li, Shengeng Tang, Yao Zhao, Huihui Bai

Comments 19 pages, 11 figures, 11 tables. Accepted to CVPR 2025 (Highlight)

详情

展开后加载摘要…

URL PDF HTML 收藏
2505.04656 2025-05-09 cs.GR

MeshGen: Generating PBR Textured Mesh with Render-Enhanced Auto-Encoder and Generative Data Augmentation

Zilong Chen, Yikai Wang, Wenqiang Sun, Feng Wang, Yiwen Chen, Huaping Liu

Comments To appear at CVPR 2025 with highlight

详情

展开后加载摘要…

URL PDF HTML 收藏
2411.14423 2025-05-09 cs.CV

PhysFlow: Unleashing the Potential of Multi-modal Foundation Models and Video Diffusion for 4D Dynamic Physical Scene Simulation

Zhuoman Liu, Weicai Ye, Yan Luximon, Pengfei Wan, Di Zhang

机构 * The Hong Kong Polytechnic University(香港理工大学) Kuaishou Technology(快手科技)

Comments CVPR 2025. Homepage: https://zhuomanliu.github.io/PhysFlow/

详情

展开后加载摘要…

URL PDF HTML 收藏
2501.04666 2025-05-08 cs.CV

Enhancing Virtual Try-On with Synthetic Pairs and Error-Aware Noise Scheduling

Nannan Li, Kevin J. Shih, Bryan A. Plummer

机构 * Boston University(波士顿大学) NVIDIA(英伟达)

Comments Accepted in CVPR 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2505.04270 2025-05-08 cs.CV cs.AI

Object-Shot Enhanced Grounding Network for Egocentric Video

Yisen Feng, Haoyu Zhang, Meng Liu, Weili Guan, Liqiang Nie

机构 * Harbin Institute of Technology (Shenzhen)(哈尔滨工业大学(深圳)) Pengcheng Laboratory(鹏城实验室) Shandong Jianzhu University(山东建筑大学)

Comments Accepted by CVPR 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2505.04109 2025-05-08 cs.CV

One2Any: One-Reference 6D Pose Estimation for Any Object

Mengya Liu, Siyuan Li, Ajad Chhatkuli, Prune Truong, Luc Van Gool, Federico Tombari

机构 * ETH Zurich(苏黎世联邦理工学院) Google(谷歌) TUM(慕尼黑工业大学)

Comments accepted by CVPR 2025

Journal ref CVPR 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2505.04055 2025-05-08 cs.CV

FoodTrack: Estimating Handheld Food Portions with Egocentric Video

Ervin Wang, Yuhao Chen

机构 * University of Waterloo(滑铁卢大学)

Comments Accepted as extended abstract at CVPR 2025 Metafood workshop

详情

展开后加载摘要…

URL PDF HTML 收藏
2504.20468 2025-05-08 cs.CV

Antidote: A Unified Framework for Mitigating LVLM Hallucinations in Counterfactual Presupposition and Object Perception

Yuanchen Wu, Lu Zhang, Hang Yao, Junlong Du, Ke Yan, Shouhong Ding, Yunsheng Wu, Xiaoqiang Li

机构 * School of Computer Engineering & Science, Shanghai University(上海大学计算机工程与科学学院) Tencent YouTu Lab(腾讯优图实验室)

Comments Accepted to CVPR 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2412.04472 2025-05-08 cs.CV

Stereo Anywhere: Robust Zero-Shot Deep Stereo Matching Even Where Either Stereo or Mono Fail

Luca Bartolomei, Fabio Tosi, Matteo Poggi, Stefano Mattoccia

机构 * Advanced Research Center on Electronic System (ARCES)(电子系统高级研究中心) Department of Computer Science and Engineering (DISI)(计算机科学与工程系) University of Bologna(博洛尼亚大学)

Comments CVPR 2025. Code: https://github.com/bartn8/stereoanywhere - Project page: https://stereoanywhere.github.io/

详情

展开后加载摘要…

URL PDF HTML 收藏
2505.03394 2025-05-07 cs.CV

EOPose : Exemplar-based object reposing using Generalized Pose Correspondences

Sarthak Mehrotra, Rishabh Jain, Mayur Hemani, Balaji Krishnamurthy, Mausoom Sarkar

机构 * Indian Institute of Technology, Bombay(印度理工学院班加罗尔分校) MDSR Lab, Adobe(Adobe MDSR实验室)

Comments Accepted in CVPR 2025 AI4CC workshop

详情

展开后加载摘要…

URL PDF HTML 收藏
2505.03299 2025-05-07 cs.CV cs.AI

Towards Efficient Benchmarking of Foundation Models in Remote Sensing: A Capabilities Encoding Approach

Pierre Adorni, Minh-Tan Pham, Stéphane May, Sébastien Lefèvre

机构 * IRISA, Université Bretagne Sud, UMR 6074(IRISA,布列塔尼大学,UMR 6074) Centre National d’Études Spatiales (CNES)(国家空间科学研究中心(CNES)) UiT The Arctic University of Norway(挪威北极大学)

Comments Accepted at the MORSE workshop of CVPR 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2412.04280 2025-05-07 cs.CV cs.GR

HumanEdit: A High-Quality Human-Rewarded Dataset for Instruction-based Image Editing

Jinbin Bai, Wei Chow, Ling Yang, Xiangtai Li, Juncheng Li, Hanwang Zhang, Shuicheng Yan

机构 * National University of Singapore(新加坡国立大学) Skywork AI Peking University(北京大学) Nanyang Technological University(南洋理工大学)

Comments Accepted to CVPR 2025 AI for Content Creation (AI4CC) Workshop. Codes and Supplementary Material: https://github.com/viiika/HumanEdit

详情

展开后加载摘要…

URL PDF HTML 收藏
2411.18673 2025-05-07 cs.CV

AC3D: Analyzing and Improving 3D Camera Control in Video Diffusion Transformers

Sherwin Bahmani, Ivan Skorokhodov, Guocheng Qian, Aliaksandr Siarohin, Willi Menapace, Andrea Tagliasacchi, David B. Lindell, Sergey Tulyakov

机构 * University of Toronto(多伦多大学) Vector Institute(向量研究所) Snap Inc.(Snap公司)

Comments CVPR 2025; Project Page: https://snap-research.github.io/ac3d/

详情

展开后加载摘要…

URL PDF HTML 收藏
2411.15279 2025-05-07 cs.LG cs.GR

Don't Mesh with Me: Generating Constructive Solid Geometry Instead of Meshes by Fine-Tuning a Code-Generation LLM

Maximilian Mews, Ansar Aynetdinov, Vivian Schiller, Peter Eisert, Alan Akbik

机构 * Humboldt-Universität zu Berlin(柏林洪堡大学)

Comments Accepted to the AI for Content Creation Workshop at CVPR 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2411.03239 2025-05-07 cs.CV

Decoupling Fine Detail and Global Geometry for Compressed Depth Map Super-Resolution

Huan Zheng, Wencheng Han, Jianbing Shen

机构 * SKL-IOTSC, CIS, University of Macau(物联网系统与智能计算研究院,计算机科学与信息工程学院,澳门大学)

Comments Accepted by CVPR 2025 & The 1st place award for the ECCV 2024 AIM Compressed Depth Upsampling Challenge

详情

展开后加载摘要…

URL PDF HTML 收藏
2406.13896 2025-05-07 cs.CV

SMORE: Simultaneous Map and Object REconstruction

Nathaniel Chodosh, Anish Madan, Simon Lucey, Deva Ramanan

机构 * Villanova University(维拉诺瓦大学) Carnegie Mellon University(卡内基梅隆大学) University of Adelaide(阿德莱德大学)

Comments 3DV 2025,CVPR 2025 4D Vision Workshop

详情

展开后加载摘要…

URL PDF HTML 收藏
2303.12675 2025-05-07 cs.CV

VecFontSDF: Learning to Reconstruct and Synthesize High-quality Vector Fonts via Signed Distance Functions

Zeqing Xia, Bojun Xiong, Zhouhui Lian

机构 * Peking University(北京大学) Center For Chinese Font Design and Research(中文字体设计与研究中心) Key Laboratory of Science, Technology and Standard in Press Industry(印刷工业科学、技术与标准重点实验室) Wangxuan Institute of Computer Technology(王轩计算机技术研究所)

Comments Accepted to CVPR 2023. Project Page: https://xiazeqing.github.io/VecFontSDF

详情

展开后加载摘要…

URL PDF HTML 收藏
2505.03116 2025-05-07 cs.CV

TimeTracker: Event-based Continuous Point Tracking for Video Frame Interpolation with Non-linear Motion

Haoyue Liu, Jinghan Xu, Yi Chang, Hanyu Zhou, Haozhi Zhao, Lin Wang, Luxin Yan

机构 * National Key Lab of Multispectral Information Intelligent Processing Technology(multispectral information intelligent processing technology key laboratory) School of Artificial Intelligence and Automation, Huazhong University of Science and Technology(人工智能与自动化学院,华中科技大学) School of Electrical and Electronic Engineering, Nanyang Technological University(电子与电气工程学院,南洋理工大学)

Comments Accepted by CVPR 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2505.03097 2025-05-07 cs.CV

Not All Parameters Matter: Masking Diffusion Models for Enhancing Generation Ability

Lei Wang, Senmao Li, Fei Yang, Jianye Wang, Ziheng Zhang, Yuhan Liu, Yaxing Wang, Jian Yang

机构 * PCA Lab, VCIP, College of Computer Science, Nankai University(PCA实验室、VCIP、计算机科学学院、南开大学)

Comments Accepted to CVPR 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2505.03020 2025-05-07 cs.AI

The Multimodal Paradox: How Added and Missing Modalities Shape Bias and Performance in Multimodal AI

Kishore Sampath, Pratheesh, Ayaazuddin Mohammad, Resmi Ramachandranpillai

机构 * Northeastern University(东北大学) Institute for Experiential AI(体验人工智能研究所)

Comments CVPR 2025 Second Workshop on Responsible Generative AI

详情

展开后加载摘要…

URL PDF HTML 收藏
2505.03012 2025-05-07 cs.CV

GIF: Generative Inspiration for Face Recognition at Scale

Saeed Ebrahimi, Sahar Rahimi, Ali Dabouei, Srinjoy Das, Jeremy M. Dawson, Nasser M. Nasrabadi

机构 * West Virginia University(西弗吉尼亚大学)

Journal ref CVPR 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2503.03307 2025-05-07 cs.CV

Full-DoF Egomotion Estimation for Event Cameras Using Geometric Solvers

Ji Zhao, Banglei Guan, Zibin Liu, Laurent Kneip

机构 * Independent Researcher(独立研究者) College of Aerospace Science and Engineering, National University of Defense Technology(国防科技大学航空航天学院) Mobile Perception Lab, ShanghaiTech University(上海科技大学移动感知实验室)

Comments Accepted by IEEE/CVF Conference on Computer Vision and Pattern Recognition (CVPR), 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2505.02753 2025-05-06 cs.CV

Advancing Generalizable Tumor Segmentation with Anomaly-Aware Open-Vocabulary Attention Maps and Frozen Foundation Diffusion Models

Yankai Jiang, Peng Zhang, Donglin Yang, Yuan Tian, Hai Lin, Xiaosong Wang

机构 * Shanghai AI Laboratory(上海人工智能实验室) Zhejiang University(浙江大学) The University of British Columbia(不列颠哥伦比亚大学)

Comments This paper is accepted to CVPR 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2505.02626 2025-05-06 cs.CV

Detect, Classify, Act: Categorizing Industrial Anomalies with Multi-Modal Large Language Models

Sassan Mokhtar, Arian Mousakhan, Silvio Galesso, Jawad Tayyub, Thomas Brox

机构 * University of Bonn(波恩大学) University of Freiburg(弗赖堡大学) Endress + Hauser(Endress+Hauser公司)

Comments Accepted as a spotlight presentation paper at the VAND Workshop, CVPR 2025. 10 pages, 6 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2505.02593 2025-05-06 cs.CV

DELTA: Dense Depth from Events and LiDAR using Transformer's Attention

Vincent Brebion, Julien Moreau, Franck Davoine

Comments Accepted for the CVPR 2025 Workshop on Event-based Vision. For the project page, see https://vbrebion.github.io/DELTA/

详情

展开后加载摘要…

URL PDF HTML 收藏