arXivDaily arXiv每日学术速递 周一至周五更新

期刊&会议

Conference on Computer Vision and Pattern Recognition · 会议 · Computer Vision

共收录 11876
2411.04125 2025-06-11 cs.CV

Community Forensics: Using Thousands of Generators to Train Fake Image Detectors

Jeongsoo Park, Andrew Owens

机构 * University of Michigan(密歇根大学)

Comments 16 pages; CVPR 2025; Project page: https://jespark.net/projects/2024/community_forensics

Journal ref In Proceedings of the Computer Vision and Pattern Recognition Conference (CVPR), pp. 8245-8257. 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2506.07996 2025-06-10 cs.CV cs.RO

UA-Pose: Uncertainty-Aware 6D Object Pose Estimation and Online Object Completion with Partial References

Ming-Feng Li, Xin Yang, Fu-En Wang, Hritam Basak, Yuyin Sun, Shreekant Gayaka, Min Sun, Cheng-Hao Kuo

机构 * Carnegie Mellon University(卡内基梅隆大学) Stony Brook University(石溪大学) National Tsing Hua University(国立清华大学) Amazon(亚马逊)

Comments CVPR 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2506.07897 2025-06-10 cs.GR cs.AI cs.CV cs.LG

GaussianVAE: Adaptive Learning Dynamics of 3D Gaussians for High-Fidelity Super-Resolution

Shuja Khalid, Mohamed Ibrahim, Yang Liu

机构 * Huawei Canada(华为加拿大)

Journal ref The Conference on Computer Vision and Pattern Recognition (CVPR) 2025 - Second Workshop on Visual Concepts

详情

展开后加载摘要…

URL PDF HTML 收藏
2506.07878 2025-06-10 cs.CV

Spatio-Temporal State Space Model For Efficient Event-Based Optical Flow

Muhammad Ahmed Humais, Xiaoqian Huang, Hussain Sajwani, Sajid Javed, Yahya Zweiri

机构 * Advanced Research and Innovation Center (ARIC), Khalifa University, Abu Dhabi, UAE(高级研究与创新中心(ARIC),哈利法大学,阿布扎比,阿联酋) Department of Computer Science, Khalifa University of Science and Technology, Abu Dhabi, UAE(计算机科学系,哈利法科学技术大学,阿布扎比,阿联酋)

Journal ref IEEE Conference on Computer Vision and Pattern Recognition (CVPR) Workshops, Nashville, 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2506.07865 2025-06-10 cs.CV cs.AI cs.CE cs.LG cs.RO

FreeGave: 3D Physics Learning from Dynamic Videos by Gaussian Velocity

Jinxi Li, Ziyang Song, Siyuan Zhou, Bo Yang

机构 * vLAR Group, The Hong Kong Polytechnic University(vLAR小组,香港理工大学)

Comments CVPR 2025. Code and data are available at: https://github.com/vLAR-group/FreeGave

详情

展开后加载摘要…

URL PDF HTML 收藏
2506.07860 2025-06-10 cs.CV

Egocentric Event-Based Vision for Ping Pong Ball Trajectory Prediction

Ivan Alberico, Marco Cannici, Giovanni Cioffi, Davide Scaramuzza

机构 * Robotics and Perception Group, University of Zurich(苏黎世大学机器人与感知组)

Comments IEEE/CVF Conference on Computer Vision and Pattern Recognition Workshops (CVPRW), Nashville (TN), USA, 2025; 5th International Workshop on Event-Based Vision

详情

展开后加载摘要…

URL PDF HTML 收藏
2506.07857 2025-06-10 cs.CV cs.AI cs.LG cs.RO

LogoSP: Local-global Grouping of Superpoints for Unsupervised Semantic Segmentation of 3D Point Clouds

Zihui Zhang, Weisheng Dai, Hongtao Wen, Bo Yang

机构 * Shenzhen Research Institute, The Hong Kong Polytechnic University(深圳研究院,香港理工大学) vLAR Group, The Hong Kong Polytechnic University(vLAR集团,香港理工大学)

Comments CVPR 2025. Code and data are available at: https://github.com/vLAR-group/LogoSP

详情

展开后加载摘要…

URL PDF HTML 收藏
2506.07750 2025-06-10 cs.CV

Difference Inversion: Interpolate and Isolate the Difference with Token Consistency for Image Analogy Generation

Hyunsoo Kim, Donghyun Kim, Suhyun Kim

机构 * Korea University(韩国大学) Korea Institute of Science and Technology(韩国科学技术院) Kyung Hee University(庆熙大学)

Comments Published at CVPR 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2506.07643 2025-06-10 cs.CV

Synthetic Visual Genome

Jae Sung Park, Zixian Ma, Linjie Li, Chenhao Zheng, Cheng-Yu Hsieh, Ximing Lu, Khyathi Chandu, Quan Kong, Norimasa Kobori, Ali Farhadi, Yejin Choi, Ranjay Krishna

机构 * University of Washington(华盛顿大学) Allen Institute for Artificial Intelligence(人工智能艾伦研究所) Stanford University(斯坦福大学) Woven by Toyota(丰田编织)

Comments CVPR 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2506.07540 2025-06-10 cs.RO cs.SY eess.SY

Fractional Collisions: A Framework for Risk Estimation of Counterfactual Conflicts using Autonomous Driving Behavior Simulations

Sreeja Roy-Singh, Sarvesh Kolekar, Daniel P. Bonny, Kyle Foss

机构 * Nuro AI

Journal ref CVPR 2025 - Workshop on Data-Driven Autonomous Driving Simulation (DDADS)

详情

展开后加载摘要…

URL PDF HTML 收藏
2506.07286 2025-06-10 cs.CV cs.LG cs.RO

Multi-Step Guided Diffusion for Image Restoration on Edge Devices: Toward Lightweight Perception in Embodied AI

Aditya Chakravarty

Comments Accepted in CVPR 2025 Embodied AI Workshop

详情

展开后加载摘要…

URL PDF HTML 收藏
2506.07087 2025-06-10 cs.CV

UCOD-DPL: Unsupervised Camouflaged Object Detection via Dynamic Pseudo-label Learning

Weiqi Yan, Lvhai Chen, Huaijia Kou, Shengchuan Zhang, Yan Zhang, Liujuan Cao

机构 * Key Laboratory of Multimedia Trusted Perception and Efficient Computing, Ministry of Education of China, Xiamen University(多媒体可信感知与高效计算重点实验室、教育部、厦门大学)

Comments Accepted by CVPR 2025 (Hightlight)

详情

展开后加载摘要…

URL PDF HTML 收藏
2503.20220 2025-06-10 cs.CV

DINeMo: Learning Neural Mesh Models with no 3D Annotations

Weijie Guo, Guofeng Zhang, Wufei Ma, Alan Yuille

机构 * Johns Hopkins University(约翰霍普金斯大学) Peking University(北京大学)

Comments Accepted to 3rd Workshop on Compositional 3D Vision at CVPR 2025 (C3DV)

详情

展开后加载摘要…

URL PDF HTML 收藏
2503.19824 2025-06-10 cs.CV cs.GR cs.MM

AudCast: Audio-Driven Human Video Generation by Cascaded Diffusion Transformers

Jiazhi Guan, Kaisiyuan Wang, Zhiliang Xu, Quanwei Yang, Yasheng Sun, Shengyi He, Borong Liang, Yukang Cao, Yingying Li, Haocheng Feng, Errui Ding, Jingdong Wang, Youjian Zhao, Hang Zhou, Ziwei Liu

机构 * DCST, Tsinghua University(清华大学智能驾驶实验室) Baidu Inc.(百度公司) S-Lab, Nanyang Technological University(南洋理工大学S实验室) Zhongguancun Laboratory(中关村实验室) University of Science and Technology of China(中国科学技术大学)

Comments Accepted to IEEE/CVF Conference on Computer Vision and Pattern Recognition (CVPR), 2025. Project page: https://guanjz20.github.io/projects/AudCast

详情

展开后加载摘要…

URL PDF HTML 收藏
2503.09402 2025-06-10 cs.CV

VLog: Video-Language Models by Generative Retrieval of Narration Vocabulary

Kevin Qinghong Lin, Mike Zheng Shou

机构 * Show Lab, National University of Singapore(新加坡国立大学展示实验室)

Comments Accepted by CVPR 2025. Github: https://github.com/showlab/VLog

详情

展开后加载摘要…

URL PDF HTML 收藏
2412.20047 2025-06-10 cs.CV cs.LG

SimLTD: Simple Supervised and Semi-Supervised Long-Tailed Object Detection

Phi Vu Tran

机构 * LexisNexis Risk Solutions(LexisNexis 风险解决方案)

Comments CVPR 2025. The reference code is available at https://github.com/lexisnexis-risk-open-source/simltd

详情

展开后加载摘要…

URL PDF HTML 收藏
2411.08466 2025-06-10 cs.CV

Weakly Supervised Temporal Action Localization via Dual-Prior Collaborative Learning Guided by Multimodal Large Language Models

Quan Zhang, Jinwei Fang, Rui Yuan, Xi Tang, Yuxin Qi, Ke Zhang, Chun Yuan

机构 * Tsinghua University(清华大学) University of Science and Technology of China(中国科学技术大学) Shanghai Jiao Tong University(上海交通大学)

Comments Accepted to CVPR

详情

展开后加载摘要…

URL PDF HTML 收藏
2408.12483 2025-06-10 cs.CV cs.AI

Not All Samples Should Be Utilized Equally: Towards Understanding and Improving Dataset Distillation

Shaobo Wang, Yantai Yang, Qilong Wang, Kaixin Li, Linfeng Zhang, Junchi Yan

机构 * School of Artificial Intelligence, Shanghai Jiao Tong University(上海交通大学人工智能学院) EPIC Lab, Shanghai Jiao Tong University(上海交通大学EPIC实验室) National University of Singapore(新加坡国立大学)

Comments Accepted by Synthetic Data for Computer Vision Workshop at CVPR, 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2408.07116 2025-06-10 cs.CV cs.GR

Generative Photomontage

Sean J. Liu, Nupur Kumari, Ariel Shamir, Jun-Yan Zhu

机构 * Carnegie Mellon University(卡内基梅隆大学) Reichman University(里赫曼大学)

Comments CVPR 2025. Project webpage: https://lseancs.github.io/generativephotomontage/

Journal ref In IEEE Conference on Computer Vision and Pattern Recognition (CVPR), 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2403.14539 2025-06-10 cs.CV cs.AI cs.LG

Robust 3D Shape Reconstruction in Zero-Shot from a Single Image in the Wild

Junhyeong Cho, Kim Youwang, Hunmin Yang, Tae-Hyun Oh

机构 * POSTECH Department of Electrical Engineering, POSTECH(电气工程系,POSTECH) KAIST(韩国科学技术院) Department of Mechanical Engineering, KAIST(机械工程系,KAIST) School of Computing, KAIST(计算学院,KAIST)

Comments Accepted to CVPR 2025, Project Page: https://ZeroShape-W.github.io

详情

展开后加载摘要…

URL PDF HTML 收藏
2403.05369 2025-06-10 cs.CV

Frequency-Adaptive Dilated Convolution for Semantic Segmentation

Linwei Chen, Lin Gu, Ying Fu

机构 * Beijing Institute of Technology(北京理工大学) RIKEN(日本理化学研究所) The University of Tokyo(东京大学)

Comments CVPR 2024 highlighted

详情

展开后加载摘要…

URL PDF HTML 收藏
2506.06918 2025-06-10 cs.CV cs.RO

Reading in the Dark with Foveated Event Vision

Carl Brander, Giovanni Cioffi, Nico Messikommer, Davide Scaramuzza

机构 * Robotics and Perception Group, University of Zurich, Switzerland(苏黎世大学机器人与感知组)

Comments CVPR 2025 Workshop on Event-based Vision

详情

展开后加载摘要…

URL PDF HTML 收藏
2506.06898 2025-06-10 cs.CV cs.LG eess.IV q-bio.NC

NSD-Imagery: A benchmark dataset for extending fMRI vision decoding methods to mental imagery

Reese Kneeland, Paul S. Scotti, Ghislain St-Yves, Jesse Breedlove, Kendrick Kay, Thomas Naselaris

机构 * University of Minnesota(明尼苏达大学) Alljoined Princeton Neuroscience Institute(普林斯顿神经科学研究所) Stability AI/Medical AI Research Center (MedARC)(Stability AI/医疗AI研究中心(MedARC))

Comments Published at CVPR 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2506.06780 2025-06-10 cs.CV cs.LG

Continuous-Time SO(3) Forecasting with Savitzky--Golay Neural Controlled Differential Equations

Lennart Bastian, Mohammad Rashed, Nassir Navab, Tolga Birdal

机构 * Technical University of Munich(慕尼黑技术大学) Munich Center of Machine Learning(慕尼黑机器学习中心) Imperial College London(伦敦帝国学院)

Comments Extended abstract, presented at the CVPR Workshop on 4D Vision

详情

展开后加载摘要…

URL PDF HTML 收藏
2506.06664 2025-06-10 cs.RO cs.CV

Generalized Trajectory Scoring for End-to-end Multimodal Planning

Zhenxin Li, Wenhao Yao, Zi Wang, Xinglong Sun, Joshua Chen, Nadine Chang, Maying Shen, Zuxuan Wu, Shiyi Lan, Jose M. Alvarez

机构 * NVIDIA Fudan University(复旦大学)

Comments The 1st place solution of the End-to-end Driving Track at the CVPR 2025 Autonomous Grand Challenge

详情

展开后加载摘要…

URL PDF HTML 收藏
2506.06596 2025-06-10 cs.CV

EV-LayerSegNet: Self-supervised Motion Segmentation using Event Cameras

Youssef Farah, Federico Paredes-Vallés, Guido De Croon, Muhammad Ahmed Humais, Hussain Sajwani, Yahya Zweiri

机构 * Advanced Research and Innovation Center, Khalifa University(哈利法大学先进研究与创新中心) MAVLab, TU Delft(代尔夫特理工大学MAVLab)

Comments This paper has been accepted for publication at the IEEE Conference on Computer Vision and Pattern Recognition (CVPR) Workshops, Nashville, 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2506.06440 2025-06-10 cs.GR cs.CV

Vid2Sim: Generalizable, Video-based Reconstruction of Appearance, Geometry and Physics for Mesh-free Simulation

Chuhao Chen, Zhiyang Dou, Chen Wang, Yiming Huang, Anjun Chen, Qiao Feng, Jiatao Gu, Lingjie Liu

机构 * University of Pennsylvania(宾夕法尼亚大学) The University of Hong Kong(香港大学) Zhejiang University(浙江大学)

Comments Accepted by CVPR 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2505.16652 2025-06-10 cs.CV cs.LG

Seeing Far and Clearly: Mitigating Hallucinations in MLLMs with Attention Causal Decoding

Feilong Tang, Chengzhi Liu, Zhongxing Xu, Ming Hu, Zelin Peng, Zhiwei Yang, Jionglong Su, Minquan Lin, Yifan Peng, Xuelian Cheng, Imran Razzak, Zongyuan Ge

机构 * Monash University(蒙纳士大学) MBZUAI XJTLU(西安交通大学) Shanghai Jiaotong University(上海交通大学) Fudan University(复旦大学) University of Minnesota(明尼苏达大学) Cornell University(康奈尔大学)

Comments Clarification note for the CVPR 2025 paper (FarSight). Prepared by a subset of the original authors; remaining co-authors are acknowledged in the text

详情

展开后加载摘要…

URL PDF HTML 收藏
2501.09277 2025-06-10 cs.CV

Bias for Action: Video Implicit Neural Representations with Bias Modulation

Alper Kayabasi, Anil Kumar Vadathya, Guha Balakrishnan, Vishwanath Saragadam

机构 * University of California Riverside(加州大学河滨分校) Rice University(Rice大学) Neal Cancer Center, Houston Methodist Hospital(休斯顿 Methodist医院 Neal癌症中心)

Comments Accepted to CVPR 2025. Project webpage: https://alpoler.github.io/actioner

详情

展开后加载摘要…

URL PDF HTML 收藏
2412.03283 2025-06-10 cs.CR cs.AI cs.CV

Black-Box Forgery Attacks on Semantic Watermarks for Diffusion Models

Andreas Müller, Denis Lukovnikov, Jonas Thietke, Asja Fischer, Erwin Quiring

机构 * Ruhr University Bochum(博尔塔尔大学博德姆)

Comments CVPR 2025

Journal ref Proc. IEEE/CVF Conf. on Computer Vision and Pattern Recognition (CVPR), 2025, pp. 20937-20946

详情

展开后加载摘要…

URL PDF HTML 收藏