arXivDaily arXiv每日学术速递 周一至周五更新

期刊&会议

Conference on Computer Vision and Pattern Recognition · 会议 · Computer Vision

共收录 11876
2411.19946 2025-06-10 cs.CV cs.AI cs.LG

DELT: A Simple Diversity-driven EarlyLate Training for Dataset Distillation

Zhiqiang Shen, Ammar Sherif, Zeyuan Yin, Shitong Shao

机构 * VILA Lab, MBZUAI(VILA实验室,MBZUAI)

Comments CVPR 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2407.17929 2025-06-10 cs.CV cs.LG

GLASS: Guided Latent Slot Diffusion for Object-Centric Learning

Krishnakant Singh, Simone Schaub-Meyer, Stefan Roth

机构 * Department of Computer Science, TU Darmstadt(图宾根大学计算机科学系)

Comments CVPR 2025. Project Page: http://visinf.github.io/glass/

详情

展开后加载摘要…

URL PDF HTML 收藏
2403.13778 2025-06-10 cs.CV cs.RO

Certified Human Trajectory Prediction

Mohammadhossein Bahari, Saeed Saadatnejad, Amirhossein Askari Farsangi, Seyed-Mohsen Moosavi-Dezfooli, Alexandre Alahi

机构 * EPFL(苏黎世联邦理工学院) Apple(苹果公司)

Comments CVPR 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2506.05934 2025-06-09 cs.CV cs.AI

FADE: Frequency-Aware Diffusion Model Factorization for Video Editing

Yixuan Zhu, Haolin Wang, Shilin Ma, Wenliang Zhao, Yansong Tang, Lei Chen, Jie Zhou

机构 * Department of Automation, Tsinghua University(自动化系,清华大学) Tsinghua Shenzhen International Graduate School, Tsinghua University(清华大学深圳国际研究生院,清华大学)

Comments Accepted by IEEE/CVF Conference on Computer Vision and Pattern Recognition (CVPR) 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2506.05890 2025-06-09 cs.CV

Unleashing the Potential of Consistency Learning for Detecting and Grounding Multi-Modal Media Manipulation

Yiheng Li, Yang Yang, Zichang Tan, Huan Liu, Weihua Chen, Xu Zhou, Zhen Lei

机构 * MAIS, Institute of Automation, Chinese Academy of Sciences(中国科学院自动化研究所MAIS) School of Artificial Intelligence, University of Chinese Academy of Sciences(中国科学院大学人工智能学院) CAIR, HKISI, Chinese Academy of Sciences(中国科学院计算机辅助研究部) School of Computer Science and Engineering, the Faculty of Innovation Engineering, M.U.S.T(慕斯科技大学计算机科学与工程学院) Sangfor Technologies Inc.(Sangfor技术有限公司) Beijing Jiaotong University(北京交通大学) Alibaba Group(阿里巴巴集团)

Comments Accepted by CVPR 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2506.05820 2025-06-09 cs.CV

DeformCL: Learning Deformable Centerline Representation for Vessel Extraction in 3D Medical Image

Ziwei Zhao, Zhixing Zhang, Yuhang Liu, Zhao Zhang, Haojun Yu, Dong Wang, Liwei Wang

机构 * Yizhun Medical AI Co., Ltd(亿智臻医疗人工智能有限公司) Center for Data Science, Peking University(北京大学数据科学中心) Center for Machine Learning Research, Peking University(北京大学机器学习研究中心) State Key Laboratory of General Artificial Intelligence, School of Intelligence Science and Technology, Peking University(北京大学人工智能国家重点实验室) Pazhou Laboratory (Huangpu), Guangzhou, Guangdong, China(琶洲实验室(黄埔),广州,广东,中国)

Comments Accepted by CVPR 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2506.05815 2025-06-09 cs.CV

NTIRE 2025 Challenge on HR Depth from Images of Specular and Transparent Surfaces

Pierluigi Zama Ramirez, Fabio Tosi, Luigi Di Stefano, Radu Timofte, Alex Costanzino, Matteo Poggi, Samuele Salti, Stefano Mattoccia, Zhe Zhang, Yang Yang, Wu Chen, Anlong Ming, Mingshuai Zhao, Mengying Yu, Shida Gao, Xiangfeng Wang, Feng Xue, Jun Shi, Yong Yang, Yong A, Yixiang Jin, Dingzhe Li, Aryan Shukla, Liam Frija-Altarac, Matthew Toews, Hui Geng, Tianjiao Wan, Zijian Gao, Qisheng Xu, Kele Xu, Zijian Zang, Jameer Babu Pinjari, Kuldeep Purohit, Mykola Lavreniuk, Jing Cao, Shenyi Li, Kui Jiang, Junjun Jiang, Yong Huang

Comments NTIRE Workshop Challenge Report, CVPR 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2506.05780 2025-06-09 cs.CV cs.AI cs.LG cs.RO

Robust sensor fusion against on-vehicle sensor staleness

Meng Fan, Yifan Zuo, Patrick Blaes, Harley Montgomery, Subhasis Das

机构 * Zoox Inc(Zoox公司)

Comments This paper has been accepted by CVPR 2025 Precognition Workshop

详情

展开后加载摘要…

URL PDF HTML 收藏
2506.05763 2025-06-09 cs.CV

Where Is The Ball: 3D Ball Trajectory Estimation From 2D Monocular Tracking

Puntawat Ponglertnapakorn, Supasorn Suwajanakorn

机构 * VISTEC Rayong, Thailand(泰国Rayong VISTEC)

Comments 11th International Workshop on Computer Vision in Sports (CVsports) at CVPR 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2506.05348 2025-06-09 cs.CV

FreeTimeGS: Free Gaussian Primitives at Anytime and Anywhere for Dynamic Scene Reconstruction

Yifan Wang, Peishan Yang, Zhen Xu, Jiaming Sun, Zhanhua Zhang, Yong Chen, Hujun Bao, Sida Peng, Xiaowei Zhou

机构 * Zhejiang University(浙江大学) Geely Automobile Research Institute(吉利汽车研究院)

Comments CVPR 2025; Project page: https://zju3dv.github.io/freetimegs/

详情

展开后加载摘要…

URL PDF HTML 收藏
2504.10984 2025-06-09 cs.CV

Seeing like a Cephalopod: Colour Vision with a Monochrome Event Camera

Sami Arja, Nimrod Kruger, Alexandre Marcireau, Nicholas Owen Ralph, Saeed Afshar, Gregory Cohen

机构 * Western Sydney University(西悉尼大学)

Comments 15 pages, 14 figures, 1 table. Accepted at CVPR 2025 (Workshop on Event-based Vision)

详情

展开后加载摘要…

URL PDF HTML 收藏
2503.11423 2025-06-09 cs.CV cs.RO

TASTE-Rob: Advancing Video Generation of Task-Oriented Hand-Object Interaction for Generalizable Robotic Manipulation

Hongxiang Zhao, Xingchen Liu, Mutian Xu, Yiming Hao, Weikai Chen, Xiaoguang Han

机构 * SSE, CUHKSZ(CUHKSZ系统科学系) FNii, CUHKSZ(CUHKSZ功能神经信息研究所)

Comments CVPR 2025; Project Page: https://taste-rob.github.io

详情

展开后加载摘要…

URL PDF HTML 收藏
2506.05563 2025-06-09 cs.CV

VoxelSplat: Dynamic Gaussian Splatting as an Effective Loss for Occupancy and Flow Prediction

Ziyue Zhu, Shenlong Wang, Jin Xie, Jiang-jiang Liu, Jingdong Wang, Jian Yang

机构 * PCA Lab, VCIP, College of Computer Science, Nankai University(PCA实验室、VCIP、计算机科学学院、南开大学)

Comments Accepted by CVPR 2025 Project Page: https://zzy816.github.io/VoxelSplat-Demo/

详情

展开后加载摘要…

URL PDF HTML 收藏
2506.05431 2025-06-09 cs.CV cs.AI cs.LG

Robustness Evaluation for Video Models with Reinforcement Learning

Ashwin Ramesh Babu, Sajad Mousavi, Vineet Gundecha, Sahand Ghorbanpour, Avisek Naug, Antonio Guillen, Ricardo Luna Gutierrez, Soumyendu Sarkar

Comments Accepted at the IEEE/CVF Conference on Computer Vision and Pattern Recognition Workshops (CVPRW) 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2506.05429 2025-06-09 cs.CV cs.AI cs.CL cs.LG

Coordinated Robustness Evaluation Framework for Vision-Language Models

Ashwin Ramesh Babu, Sajad Mousavi, Vineet Gundecha, Sahand Ghorbanpour, Avisek Naug, Antonio Guillen, Ricardo Luna Gutierrez, Soumyendu Sarkar

机构 * Hewlett Packard Enterprise (Hewlett Packard Labs)(惠普企业公司(惠普实验室))

Comments Accepted: IEEE/CVF Conference on Computer Vision and Pattern Recognition Workshops (CVPRW) 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2506.05397 2025-06-09 cs.GR cs.AI

Gen4D: Synthesizing Humans and Scenes in the Wild

Jerrin Bright, Zhibo Wang, Yuhao Chen, Sirisha Rambhatla, John Zelek, David Clausi

机构 * Vision and Image Processing Lab(视觉与图像处理实验室) Critical ML Lab(关键机器学习实验室) University of Waterloo(滑铁卢大学)

Comments Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition (CVPR) Workshops

详情

展开后加载摘要…

URL PDF HTML 收藏
2505.22458 2025-06-09 cs.CV

Universal Domain Adaptation for Semantic Segmentation

Seun-An Choe, Keon-Hee Park, Jinwoo Choi, Gyeong-Moon Park

机构 * Kyung Hee University(京亨大学) Korea University(韩国大学)

Comments Accepted by CVPR 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2503.19868 2025-06-09 cs.IR cs.AI cs.CV cs.LG

GENIUS: A Generative Framework for Universal Multimodal Search

Sungyeon Kim, Xinliang Zhu, Xiaofan Lin, Muhammet Bastan, Douglas Gray, Suha Kwak

机构 * Amazon(亚马逊公司) POSTECH

Comments Accepted to CVPR 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2412.05278 2025-06-09 cs.CV cs.GR

Birth and Death of a Rose

Chen Geng, Yunzhi Zhang, Shangzhe Wu, Jiajun Wu

机构 * Stanford University(斯坦福大学)

Comments CVPR 2025 Oral. Project website: https://chen-geng.com/rose4d

详情

展开后加载摘要…

URL PDF HTML 收藏
2303.14005 2025-06-09 cs.CV cs.AI

Category Query Learning for Human-Object Interaction Classification

Chi Xie, Fangao Zeng, Yue Hu, Shuang Liang, Yichen Wei

Comments Accepted by CVPR 2023

详情

展开后加载摘要…

URL PDF HTML 收藏
2506.04790 2025-06-06 cs.CV cs.IR cs.LG

LotusFilter: Fast Diverse Nearest Neighbor Search via a Learned Cutoff Table

Yusuke Matsui

机构 * The University of Tokyo(东京大学)

Comments CVPR 2025. GitHub: https://github.com/matsui528/lotf

详情

展开后加载摘要…

URL PDF HTML 收藏
2506.04453 2025-06-06 eess.IV cs.CR cs.CV cs.LG

Gradient Inversion Attacks on Parameter-Efficient Fine-Tuning

Hasin Us Sami, Swapneel Sen, Amit K. Roy-Chowdhury, Srikanth V. Krishnamurthy, Basak Guler

机构 * University of California, Riverside(加州大学河滨分校)

Comments 2025 IEEE/CVF Conference on Computer Vision and Pattern Recognition (CVPR 2025)

详情

展开后加载摘要…

URL PDF HTML 收藏
2506.04421 2025-06-06 cs.CV cs.AI cs.LG

HMAR: Efficient Hierarchical Masked Auto-Regressive Image Generation

Hermann Kumbong, Xian Liu, Tsung-Yi Lin, Ming-Yu Liu, Xihui Liu, Ziwei Liu, Daniel Y. Fu, Christopher Ré, David W. Romero

机构 * Stanford University(斯坦福大学) NVIDIA CUHK(中国香港大学) HKU(香港大学) NTU(国立新加坡大学) UCSD(加州大学圣地亚哥分校) Together AI

Comments Accepted to CVPR 2025. Project Page: https://research.nvidia.com/labs/dir/hmar/

详情

展开后加载摘要…

URL PDF HTML 收藏
2506.04379 2025-06-06 cs.CV cs.AI q-bio.NC

Visualizing and Controlling Cortical Responses Using Voxel-Weighted Activation Maximization

Matthew W. Shinkle, Mark D. Lescroart

机构 * Department of Cognitive and Brain Sciences(认知与脑科学系) University of Nevada, Reno(内华达大学里诺分校)

Comments Accepted to the Mechanistic Interpretability for Vision (MIV) Workshop at the 2025 Conference on Computer Vision and Pattern Recognition (CVPR) conference

详情

展开后加载摘要…

URL PDF HTML 收藏
2506.01199 2025-06-06 cs.AI cs.RO

Test Automation for Interactive Scenarios via Promptable Traffic Simulation

Augusto Mondelli, Yueshan Li, Alessandro Zanardi, Emilio Frazzoli

机构 * ETH Zurich(苏黎世联邦理工学院)

Comments Accepted by CVPR 2025 Workshop Data-Driven Autonomous Driving Simulation (track 1)

详情

展开后加载摘要…

URL PDF HTML 收藏
2505.04788 2025-06-06 cs.CV

Convex Relaxation for Robust Vanishing Point Estimation in Manhattan World

Bangyan Liao, Zhenjun Zhao, Haoang Li, Yi Zhou, Yingping Zeng, Hao Li, Peidong Liu

Comments Accepted to CVPR 2025 as Award Candidate & Oral Presentation. The first two authors contributed equally to this work. Code: https://github.com/WU-CVGL/GlobustVP

详情

展开后加载摘要…

URL PDF HTML 收藏
2504.16062 2025-06-06 cs.RO cs.CV

ForesightNav: Learning Scene Imagination for Efficient Exploration

Hardik Shah, Jiaxu Xing, Nico Messikommer, Boyang Sun, Marc Pollefeys, Davide Scaramuzza

机构 * ETH Zurich(苏黎世联邦理工学院) University of Zurich(苏黎世大学) Microsoft(微软公司)

Journal ref IEEE/CVF Conference on Computer Vision and Pattern Recognition Workshops (CVPRW), Nashville, 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2504.11199 2025-06-06 cs.CV

Video Summarization with Large Language Models

Min Jung Lee, Dayoung Gong, Minsu Cho

机构 * Pohang University of Science and Technology (POSTECH)(釜山科学技术大学) GenGenAI

Comments Accepted to CVPR 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2504.10746 2025-06-06 cs.CV cs.AI cs.LG cs.SD eess.AS

Hearing Anywhere in Any Environment

Xiulong Liu, Anurag Kumar, Paul Calamia, Sebastia V. Amengual, Calvin Murdock, Ishwarya Ananthabhotla, Philip Robinson, Eli Shlizerman, Vamsi Krishna Ithapu, Ruohan Gao

机构 * University of Washington(华盛顿大学) Meta University of Maryland, College Park(马里兰大学帕克分校)

Comments CVPR 2025; Project Page: https://dragonliu1995.github.io/hearinganywhereinanyenvironment/

详情

展开后加载摘要…

URL PDF HTML 收藏
2506.04174 2025-06-05 cs.CV

FlexGS: Train Once, Deploy Everywhere with Many-in-One Flexible 3D Gaussian Splatting

Hengyu Liu, Yuehao Wang, Chenxin Li, Ruisi Cai, Kevin Wang, Wuyang Li, Pavlo Molchanov, Peihao Wang, Zhangyang Wang

机构 * The Chinese University of Hong Kong(香港中文大学) University of Texas at Austin(德克萨斯大学奥斯汀分校) Nvidia(英伟达)

Comments CVPR 2025; Project Page: https://flexgs.github.io

详情

展开后加载摘要…

URL PDF HTML 收藏