arXivDaily arXiv每日学术速递 周一至周五更新

期刊&会议

International Conference on Computer Vision · 会议 · Computer Vision

共收录 4770
2508.00913 2025-08-05 cs.CV cs.LG

TESPEC: Temporally-Enhanced Self-Supervised Pretraining for Event Cameras

Mohammad Mohammadi, Ziyi Wu, Igor Gilitschenski

机构 * University of Toronto(多伦多大学) Vector Institute(向量研究所)

Comments Accepted at IEEE/CVF International Conference on Computer Vision (ICCV) 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.00599 2025-08-05 cs.CV

DPoser-X: Diffusion Model as Robust 3D Whole-body Human Pose Prior

Junzhe Lu, Jing Lin, Hongkun Dou, Ailing Zeng, Yue Deng, Xian Liu, Zhongang Cai, Lei Yang, Yulun Zhang, Haoqian Wang, Ziwei Liu

机构 * Tsinghua University(清华大学) Nanyang Technological University(南洋理工大学) Beihang University(北航) NVIDIA Research(NVIDIA研究) SenseTime Research(商汤科技研究院) Shanghai Jiao Tong University(上海交通大学)

Comments ICCV 2025 (oral); Code released: https://github.com/moonbow721/DPoser

详情

展开后加载摘要…

URL PDF HTML 收藏
2507.19451 2025-08-05 cs.CV

GS-Occ3D: Scaling Vision-only Occupancy Reconstruction with Gaussian Splatting

Baijun Ye, Minghui Qin, Saining Zhang, Moonjun Gong, Shaoting Zhu, Zebang Shen, Luan Zhang, Lu Zhang, Hao Zhao, Hang Zhao

机构 * IIIS, THU(清华大学人工智能学院) Shanghai Qi Zhi Institute(上海启智研究院) AIR, THU(清华大学人工智能研究院) BAAI(北京人工智能研究院) Mercedes-Benz Group China Ltd.(梅赛德斯-奔驰集团中国有限公司)

Comments ICCV 2025. Project Page: https://gs-occ3d.github.io/

详情

展开后加载摘要…

URL PDF HTML 收藏
2507.15542 2025-08-05 cs.CV

HOLa: Zero-Shot HOI Detection with Low-Rank Decomposed VLM Feature Adaptation

Qinqian Lei, Bo Wang, Robby T. Tan

机构 * National University of Singapore(国立新加坡大学) University of Mississippi(密苏里大学) ASUS Intelligent Cloud Services (AICS)(ASUS智能云服务(AICS))

Comments Accepted by ICCV 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2507.12441 2025-08-05 cs.CV cs.LG

Describe Anything Model for Visual Question Answering on Text-rich Images

Yen-Linh Vu, Dinh-Thang Duong, Truong-Binh Duong, Anh-Khoi Nguyen, Thanh-Huy Nguyen, Le Thien Phuc Nguyen, Jianhua Xing, Xingjian Li, Tianyang Wang, Ulas Bagci, Min Xu

机构 * AI VIETNAM Lab(AI越南实验室) Carnegie Mellon University(卡内基梅隆大学) University of Wisconsin - Madison(威斯康星大学麦迪逊分校) University of Pittsburgh(匹兹堡大学) University of Alabama at Birmingham(阿拉巴马大学伯明翰分校) Northwestern University(西北大学)

Comments 11 pages, 5 figures. Accepted to VisionDocs @ ICCV 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2507.10340 2025-08-05 cs.CV

Text Embedding Knows How to Quantize Text-Guided Diffusion Models

Hongjae Lee, Myungjun Son, Dongjea Kang, Seung-Won Jung

机构 * Korea University(韩国大学)

Comments ICCV 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2507.05256 2025-08-05 cs.CV

SegmentDreamer: Towards High-fidelity Text-to-3D Synthesis with Segmented Consistency Trajectory Distillation

Jiahao Zhu, Zixuan Chen, Guangcong Wang, Xiaohua Xie, Yi Zhou

机构 * Sun Yat-sen University(中山大学) Great Bay University(大贝大学) Pazhou Lab (Huangpu)(琶洲实验室) Guangdong Province Key Laboratory of Information Security Technology(广东省信息安全技术重点实验室)

Comments Accepted by ICCV 2025, project page: https://zjhjojo.github.io/segmentdreamer/

详情

展开后加载摘要…

URL PDF HTML 收藏
2507.04183 2025-08-05 cs.CV

Voyaging into Perpetual Dynamic Scenes from a Single View

Fengrui Tian, Tianjiao Ding, Jinqi Luo, Hancheng Min, René Vidal

机构 * University of Pennsylvania(宾夕法尼亚大学)

Comments Accepted by International Conference on Computer Vision (ICCV) 2025. Project Page: https://tianfr.github.io/DynamicVoyager

详情

展开后加载摘要…

URL PDF HTML 收藏
2507.03257 2025-08-05 cs.CV

LACONIC: A 3D Layout Adapter for Controllable Image Creation

Léopold Maillard, Tom Durand, Adrien Ramanana Rahary, Maks Ovsjanikov

机构 * LIX, École Polytechnique, IP Paris(巴黎高等理工学院LIX研究所) Dassault Systèmes(达索系统)

Comments ICCV 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2506.21835 2025-08-05 cs.CV

ProSAM: Enhancing the Robustness of SAM-based Visual Reference Segmentation with Probabilistic Prompts

Xiaoqi Wang, Clint Sebastian, Wenbin He, Liu Ren

机构 * Bosch Research North America(博世北美研究院) Bosch Center for Artificial Intelligence (BCAI)(博世人工智能中心)

Comments ICCV 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2506.21549 2025-08-05 cs.CV

SiM3D: Single-instance Multiview Multimodal and Multisetup 3D Anomaly Detection Benchmark

Alex Costanzino, Pierluigi Zama Ramirez, Luigi Lella, Matteo Ragaglia, Alessandro Oliva, Giuseppe Lisanti, Luigi Di Stefano

机构 * CVLab, University of Bologna, Italy(博洛尼亚大学计算机视觉实验室) SACMI Imola, Italy(意大利SACMI公司)

Comments Accepted at ICCV 2025. Project page: https://alex-costanzino.github.io/SiM3D/

详情

展开后加载摘要…

URL PDF HTML 收藏
2506.21260 2025-08-05 cs.CV

DuET: Dual Incremental Object Detection via Exemplar-Free Task Arithmetic

Munish Monga, Vishal Chudasama, Pankaj Wasnik, Biplab Banerjee

机构 * Sony Research India(索尼印度研究实验室) Indian Institute of Technology, Bombay(印度理工学院,孟买)

Comments Accepted at ICCV 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2506.13298 2025-08-05 cs.CV cs.AI

Fair Generation without Unfair Distortions: Debiasing Text-to-Image Generation with Entanglement-Free Attention

Jeonghoon Park, Juyoung Lee, Chaeyeon Chung, Jaeseong Lee, Jaegul Choo, Jindong Gu

机构 * KAIST(韩国科学技术院) Kakao Corp.(韩国 Kakao 公司) Yonsei University(延世大学) University of Oxford(牛津大学)

Comments Accepted to ICCV 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2505.22664 2025-08-05 cs.CV

Zero-Shot Vision Encoder Grafting via LLM Surrogates

Kaiyu Yue, Vasu Singla, Menglin Jia, John Kirchenbauer, Rifaa Qadri, Zikui Cai, Abhinav Bhatele, Furong Huang, Tom Goldstein

机构 * University of Maryland(马里兰大学) Meta

Comments ICCV 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2504.13490 2025-08-05 cs.CV

Early Timestep Zero-Shot Candidate Selection for Instruction-Guided Image Editing

Joowon Kim, Ziseok Lee, Donghyeon Cho, Sanghyun Jo, Yeonsung Jung, Kyungsu Kim, Eunho Yang

机构 * KAIST(韩国科学技术院) Seoul National University(首尔国立大学) OGQ

Comments Accepted to ICCV 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2504.01308 2025-08-05 cs.CV

Safeguarding Vision-Language Models: Mitigating Vulnerabilities to Gaussian Noise in Perturbation-based Attacks

Jiawei Wang, Yushen Zuo, Yuanjun Chai, Zhendong Liu, Yicheng Fu, Yichun Feng, Kin-Man Lam

机构 * University of Science and Technology of China(中国科学技术大学) The Hong Kong Polytechnic University(香港理工大学) University of Washington(华盛顿大学) Nanjing University(南京大学) Stanford University(斯坦福大学) University of Chinese Academy of Sciences(中国科学院大学)

Comments ICCV 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2503.19351 2025-08-05 cs.CV

Multi-Object Sketch Animation by Scene Decomposition and Motion Planning

Jingyu Liu, Zijie Xin, Yuhan Fu, Ruixiang Zhao, Bangxiang Lan, Xirong Li

Comments Accepted by ICCV 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2503.17544 2025-08-05 cs.CV cs.AI

PRIMAL: Physically Reactive and Interactive Motor Model for Avatar Learning

Yan Zhang, Yao Feng, Alpár Cseke, Nitin Saini, Nathan Bajandas, Nicolas Heron, Michael J. Black

机构 * Meshcapade Max Planck Institute for Intelligent Systems(马克斯·普朗克智能系统研究所) Stanford University(斯坦福大学)

Comments ICCV'25 camera ready; main paper and appendix; 19 pages in total

详情

展开后加载摘要…

URL PDF HTML 收藏
2503.13652 2025-08-05 cs.CV

Web Artifact Attacks Disrupt Vision Language Models

Maan Qraitem, Piotr Teterwak, Kate Saenko, Bryan A. Plummer

机构 * Boston University(波士顿大学)

Comments Accepted at ICCV 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2503.12937 2025-08-05 cs.AI cs.CL cs.CV cs.LG

R1-VL: Learning to Reason with Multimodal Large Language Models via Step-wise Group Relative Policy Optimization

Jingyi Zhang, Jiaxing Huang, Huanjin Yao, Shunyu Liu, Xikun Zhang, Shijian Lu, Dacheng Tao

机构 * Nanyang Technological University(南洋理工大学)

Comments ICCV 2025 Camera Ready

详情

展开后加载摘要…

URL PDF HTML 收藏
2503.08668 2025-08-05 cs.CV

SSVQ: Unleashing the Potential of Vector Quantization with Sign-Splitting

Shuaiting Li, Juncan Deng, Chenxuan Wang, Kedong Xu, Rongtao Deng, Hong Gu, Haibin Shen, Kejie Huang

机构 * Zhejiang University(浙江大学) vivo Mobile Communication Co., Ltd(vivo移动通信有限公司)

Comments ICCV'25 camera ready

详情

展开后加载摘要…

URL PDF HTML 收藏
2503.06271 2025-08-05 cs.CV

SplatTalk: 3D VQA with Gaussian Splatting

Anh Thai, Songyou Peng, Kyle Genova, Leonidas Guibas, Thomas Funkhouser

机构 * Georgia Institute of Technology(佐治亚理工学院) Google DeepMind(谷歌DeepMind)

Comments Accepted at ICCV 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2502.06957 2025-08-05 cs.CV

GAS: Generative Avatar Synthesis from a Single Image

Yixing Lu, Junting Dong, Youngjoong Kwon, Qin Zhao, Bo Dai, Fernando De la Torre

机构 * Carnegie Mellon University(卡内基梅隆大学) Shanghai AI Laboratory(上海人工智能实验室) Stanford University(斯坦福大学)

Comments ICCV 2025; Project Page: https://humansensinglab.github.io/GAS/

详情

展开后加载摘要…

URL PDF HTML 收藏
2502.04981 2025-08-05 cs.CV

AutoOcc: Automatic Open-Ended Semantic Occupancy Annotation via Vision-Language Guided Gaussian Splatting

Xiaoyu Zhou, Jingqi Wang, Yongtao Wang, Yufei Wei, Nan Dong, Ming-Hsuan Yang

机构 * Wangxuan Institute of Computer Technology, Peking University(北京大学计算机技术研究院) Chongqing Changan Automobile Co., Ltd(重庆长安汽车有限公司) University of California, Merced(加州大学默塞德分校)

Comments ICCV 2025 Hightlight (main conference)

详情

展开后加载摘要…

URL PDF HTML 收藏
2501.18898 2025-08-05 cs.CV cs.GR

GestureLSM: Latent Shortcut based Co-Speech Gesture Generation with Spatial-Temporal Modeling

Pinxin Liu, Luchuan Song, Junhua Huang, Haiyang Liu, Chenliang Xu

机构 * University of Rochester(罗切斯特大学) University of Tokyo(东京大学)

Comments Accepted to ICCV 2025. Project Page: https://andypinxinliu.github.io/GestureLSM

详情

展开后加载摘要…

URL PDF HTML 收藏
2501.07730 2025-08-05 cs.CV

Democratizing Text-to-Image Masked Generative Models with Compact Text-Aware One-Dimensional Tokens

Dongwon Kim, Ju He, Qihang Yu, Chenglin Yang, Xiaohui Shen, Suha Kwak, Liang-Chieh Chen

机构 * ByteDance Seed(字节跳动种子) POSTECH

Comments ICCV 2025. Project page at https://tacju.github.io/projects/maskgen.html

详情

展开后加载摘要…

URL PDF HTML 收藏
2411.16725 2025-08-05 cs.CV

$\textit{Revelio}$: Interpreting and leveraging semantic information in diffusion models

Dahye Kim, Xavier Thomas, Deepti Ghadiyaram

机构 * Boston University(波士顿大学) Runway

Comments ICCV 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2411.15472 2025-08-05 cs.CV cs.AI cs.GR

KinMo: Kinematic-aware Human Motion Understanding and Generation

Pengfei Zhang, Pinxin Liu, Pablo Garrido, Hyeongwoo Kim, Bindita Chaudhuri

机构 * University of California, Irvine(加州大学欧文分校) University of Rochester(罗切斯特大学) Imperial College, London(伦敦帝国学院) Flawless AI

Comments Accepted to ICCV 2025; Project page: https://andypinxinliu.github.io/KinMo

详情

展开后加载摘要…

URL PDF HTML 收藏
2409.10876 2025-08-05 eess.IV cs.CV eess.SP

Coordinate-based Speed of Sound Recovery for Aberration-Corrected Photoacoustic Computed Tomography

Tianao Li, Manxiu Cui, Cheng Ma, Emma Alexander

机构 * Northwestern University(西北大学) California Institute of Technology(加州理工学院) Tsinghua University(清华大学) NSF-Simons AI Institute for the Sky (SkAI)(NSF-模拟斯通AI天空研究所(SkAI))

Comments Accepted to IEEE/CVF International Conference on Computer Vision (ICCV), 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2409.01071 2025-08-05 cs.CV cs.CL

VideoLLaMB: Long Streaming Video Understanding with Recurrent Memory Bridges

Yuxuan Wang, Yiqi Song, Cihang Xie, Yang Liu, Zilong Zheng

机构 * NLCo Lab, State Key Laboratory of General Artificial Intelligence, BIGAI(NLCo实验室、国家一般人工智能重点实验室、BIGAI) School of Computer Science & Technology, Beijing Institute of Technology(计算机科学与技术学院、北京理工大学) Computer Science and Engineering, University of California(计算机科学与工程系、加州大学) Wangxuan Institute of Computer Technology, Peking University(王轩计算机技术研究所、北京大学)

Comments To appear at ICCV 2025

详情

展开后加载摘要…

URL PDF HTML 收藏