arXivDaily arXiv每日学术速递 周一至周五更新

期刊&会议

International Conference on Computer Vision · 会议 · Computer Vision

共收录 4770
2503.08685 2025-07-29 cs.CV

"Principal Components" Enable A New Language of Images

Xin Wen, Bingchen Zhao, Ismail Elezi, Jiankang Deng, Xiaojuan Qi

Comments Accepted by ICCV 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2503.06339 2025-07-29 cs.LG cs.CV

Learning to Unlearn while Retaining: Combating Gradient Conflicts in Machine Unlearning

Gaurav Patel, Qiang Qiu

Comments Accepted at ICCV 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2502.04469 2025-07-29 cs.CV cs.AI

Ask and Remember: A Questions-Only Replay Strategy for Continual Visual Question Answering

Imad Eddine Marouf, Enzo Tartaglione, Stephane Lathuiliere, Joost van de Weijer

机构 * LTCI, Télécom-Paris, Institut Polytechnique de Paris(LTCI, Télécom-Paris, Institut Polytechnique de Paris) Inria, LJK, Univ. Grenoble Alpes(Inria, LJK, 火车头 Grenoble Alpes) Universitat Autónoma de Barcelona(Autonomous University of Barcelona)

Comments ICCV 2025, 8 pages. Code: https://github.com/IemProg/QUAD

详情

展开后加载摘要…

URL PDF HTML 收藏
2412.06334 2025-07-29 cs.CV

TriDi: Trilateral Diffusion of 3D Humans, Objects, and Interactions

Ilya A. Petrov, Riccardo Marin, Julian Chibane, Gerard Pons-Moll

机构 * University of Tübingen, Germany(图宾根大学) Tübingen AI Center, Germany(图宾根人工智能中心) Technical University of Munich, Germany(慕尼黑技术大学) Munich Center for Machine Learning, Germany(慕尼黑机器学习中心) Max Planck Institute for Informatics, Saarland Informatics Campus, Germany(马克斯·普朗克信息研究所)

Comments 2025 IEEE/CVF International Conference on Computer Vision (ICCV)

详情

展开后加载摘要…

URL PDF HTML 收藏
2412.05983 2025-07-29 cs.CV

Chimera: Improving Generalist Model with Domain-Specific Experts

Tianshuo Peng, Mingsheng Li, Jiakang Yuan, Hongbin Zhou, Renqiu Xia, Renrui Zhang, Lei Bai, Song Mao, Bin Wang, Aojun Zhou, Botian Shi, Tao Chen, Bo Zhang, Xiangyu Yue

机构 * Shanghai Artificial Intelligence Laboratory(上海人工智能实验室) MMLab College of Future Information Technology, Fudan University(未来信息技术学院,复旦大学) Shanghai Jiao Tong University(上海交通大学) Shanghai Innovation Institute(上海创新研究院)

Comments Accepted by ICCV-2025, Chimera Homepage: https://alpha-innovator.github.io/chimera_page

详情

展开后加载摘要…

URL PDF HTML 收藏
2412.03515 2025-07-29 cs.CV

Distilling Diffusion Models to Efficient 3D LiDAR Scene Completion

Shengyuan Zhang, An Zhao, Ling Yang, Zejian Li, Chenye Meng, Haoran Xu, Tianrun Chen, AnYang Wei, Perry Pengyun GU, Lingyun Sun

Comments This paper is accepted by ICCV'25(Oral), the model and code are publicly available on https://github.com/happyw1nd/ScoreLiDAR

详情

展开后加载摘要…

URL PDF HTML 收藏
2412.00623 2025-07-29 cs.CV

A Lesson in Splats: Teacher-Guided Diffusion for 3D Gaussian Splats Generation with 2D Supervision

Chensheng Peng, Ido Sobol, Masayoshi Tomizuka, Kurt Keutzer, Chenfeng Xu, Or Litany

机构 * UC Berkeley(伯克利大学) Technion(技术学院) NVIDIA(英伟达)

Comments ICCV 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2411.17762 2025-07-29 cs.CV

MUSE-VL: Modeling Unified VLM through Semantic Discrete Encoding

Rongchang Xie, Chen Du, Ping Song, Chang Liu

Comments ICCV 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2410.14462 2025-07-29 cs.CV

LUDVIG: Learning-Free Uplifting of 2D Visual Features to Gaussian Splatting Scenes

Juliette Marrie, Romain Menegaux, Michael Arbel, Diane Larlus, Julien Mairal

机构 * Univ. Grenoble Alpes, Inria, CNRS, Grenoble INP, LJK(格勒诺布尔阿尔卑斯大学、法国国家科学研究中心、格勒诺布尔INP、LJK)

Comments Published at ICCV 2025. Project page: https://juliettemarrie.github.io/ludvig

详情

展开后加载摘要…

URL PDF HTML 收藏
2408.07703 2025-07-29 cs.CV

Knowledge Distillation with Refined Logits

Wujie Sun, Defang Chen, Siwei Lyu, Genlang Chen, Chun Chen, Can Wang

机构 * State Key Laboratory of Blockchain and Data Security, Zhejiang University(区块链与数据安全国家重点实验室,浙江大学) School of Software Technology, Zhejiang University(浙江大学软件技术学院) Hangzhou High-Tech Zone (Binjiang) Institute of Blockchain and Data Security(杭州高新技术区(滨江区)区块链与数据安全研究院) University at Buffalo, State University of New York(纽约州立大学布法罗分校) NingboTech University(宁波科技学院)

Comments ICCV 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2311.10887 2025-07-29 cs.CV cs.AI

Point Cloud Self-supervised Learning via 3D to Multi-view Masked Learner

Zhimin Chen, Xuewei Chen, Xiao Guo, Yingwei Li, Longlong Jing, Liang Yang, Bing Li

机构 * Clemson University(克莱姆森大学) Michigan State University(密歇根州立大学) Johns Hopkins University(约翰霍普金斯大学) The City University of New York(纽约城市大学)

Comments Accepted by ICCV 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2507.19754 2025-07-29 cs.CV

Latest Object Memory Management for Temporally Consistent Video Instance Segmentation

Seunghun Lee, Jiwan Seo, Minwoo Choi, Kiljoon Han, Jaehoon Jeong, Zane Durante, Ehsan Adeli, Sang Hyun Park, Sunghoon Im

机构 * DGIST, Daegu, Republic of Korea(韩国大邱科学技术院) Stanford University(斯坦福大学)

Comments ICCV 2025. Code: https://github.com/Seung-Hun-Lee/LOMM

详情

展开后加载摘要…

URL PDF HTML 收藏
2507.05390 2025-07-29 cs.CV eess.IV

From General to Specialized: The Need for Foundational Models in Agriculture

Vishal Nedungadi, Xingguo Xiong, Aike Potze, Ron Van Bree, Tao Lin, Marc Rußwurm, Ioannis N. Athanasiadis

机构 * Wageningen University and Research(瓦赫宁根大学和研究学院) Zhejiang University(浙江大学)

Comments Accepted to the SEA Workshop (Sustainability with Earth Observation & AI) at ICCV 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2504.16907 2025-07-29 cs.CV cs.AI

BadVideo: Stealthy Backdoor Attack against Text-to-Video Generation

Ruotong Wang, Mingli Zhu, Jiarong Ou, Rui Chen, Xin Tao, Pengfei Wan, Baoyuan Wu

机构 * The Chinese University of Hong Kong, Shenzhen(香港中文大学(深圳)) Kling Team, Kuaishou Technology(快手科技 Kling 团队)

Comments Accepted by ICCV 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2411.12789 2025-07-29 cs.CV

Efficient Physics Simulation for 3D Scenes via MLLM-Guided Gaussian Splatting

Haoyu Zhao, Hao Wang, Xingyue Zhao, Hao Fei, Hongqiu Wang, Chengjiang Long, Hua Zou

机构 * School of Computer Science, Wuhan University(武汉大学计算机学院) Wuhan National Laboratory for Optoelectronics, Huazhong University of Science and Technology(华中科技大学光电研究院) Meta Reality Lab(Meta现实实验室) Xi’an Jiao Tong University(西安交通大学) National University of Singapore(新加坡国立大学) The Department of Systems Hub, Hong Kong University of Science and Technology (Guangzhou)(香港科技大学系统枢纽部门(广州))

Comments ICCV 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2507.19481 2025-07-28 cs.CV

HairCUP: Hair Compositional Universal Prior for 3D Gaussian Avatars

Byungjun Kim, Shunsuke Saito, Giljoo Nam, Tomas Simon, Jason Saragih, Hanbyul Joo, Junxuan Li

Comments ICCV 2025. Project Page: https://bjkim95.github.io/haircup/

详情

展开后加载摘要…

URL PDF HTML 收藏
2507.19360 2025-07-28 cs.CV

EA-ViT: Efficient Adaptation for Elastic Vision Transformer

Chen Zhu, Wangbo Zhao, Huiwen Zhang, Samir Khaki, Yuhao Zhou, Weidong Tang, Shuo Wang, Zhihang Yuan, Yuzhang Shang, Xiaojiang Peng, Kai Wang, Dawei Yang

机构 * National University of Singapore(国立新加坡大学) Xidian University(西安电子科技大学) University of Toronto(多伦多大学) Houmo AI UCF Shenzhen Technology University(深圳技术大学)

Comments Published as a conference paper at ICCV 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2507.19359 2025-07-28 cs.CV

SemGes: Semantics-aware Co-Speech Gesture Generation using Semantic Coherence and Relevance Learning

Lanmiao Liu, Esam Ghaleb, Aslı Özyürek, Zerrin Yumak

机构 * Max Planck Institute for Psycholinguistics(马克斯·普朗克心理学研究所) Donders Institute for Brain Cognition and Behaviour(多纳尔斯脑认知与行为研究所) Utrecht University(乌得勒支大学)

Comments Accepted to IEEE/CVF International Conference on Computer Vision (ICCV) 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2507.19292 2025-07-28 cs.CV

PINO: Person-Interaction Noise Optimization for Long-Duration and Customizable Motion Generation of Arbitrary-Sized Groups

Sakuya Ota, Qing Yu, Kent Fujiwara, Satoshi Ikehata, Ikuro Sato

机构 * Institute of Science Tokyo(东京科学研究所) LY Corporation(LY公司) National Institute of Informatics (NII)(日本信息处理学会)

Comments Accepted to ICCV 2025, Project page: https://sinc865.github.io/pino/

详情

展开后加载摘要…

URL PDF HTML 收藏
2507.19239 2025-07-28 cs.CV

CoopTrack: Exploring End-to-End Learning for Efficient Cooperative Sequential Perception

Jiaru Zhong, Jiahao Wang, Jiahui Xu, Xiaofan Li, Zaiqing Nie, Haibao Yu

机构 * Institute for AI Industry Research, Tsinghua University(人工智能产业研究所,清华大学) The Hong Kong Polytechnic University(香港理工大学) The University of Hong Kong(香港大学) School of Vehicle and Mobility, Tsinghua University(车辆与移动技术学院,清华大学) Baidu Inc.(百度公司)

Comments Accepted by ICCV 2025 (Highlight)

详情

展开后加载摘要…

URL PDF HTML 收藏
2507.19140 2025-07-28 cs.CV

Balancing Conservatism and Aggressiveness: Prototype-Affinity Hybrid Network for Few-Shot Segmentation

Tianyu Zou, Shengwu Xiong, Ruilin Yao, Yi Rong

机构 * School of Computer Science and Artificial Intelligence, Wuhan University of Technology(武汉理工大学计算机科学与人工智能学院) Interdisciplinary Artificial Intelligence Research Institute, Wuhan College(武汉学院交叉人工智能研究院) Sanya Science and Education Innovation Park, Wuhan University of Technology(武汉理工大学三亚科学教育创新园) School of Artificial Intelligence, University of Chinese Academy of Sciences(中国科学院大学人工智能学院) Foundation Model Research Center, Institute of Automation, Chinese Academy of Sciences(中国科学院自动化研究所基础模型研究中心)

Comments 8 pages, 7 figures

Journal ref ICCV 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2507.19131 2025-07-28 cs.CV

MixA-Q: Revisiting Activation Sparsity for Vision Transformers from a Mixed-Precision Quantization Perspective

Weitian Wang, Rai Shubham, Cecilia De La Parra, Akash Kumar

机构 * Robert Bosch GmbH(博世公司) Ruhr University Bochum(博尔塔尔大学波恩)

Comments Accepted to ICCV 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2507.19125 2025-07-28 eess.IV cs.CV cs.MM

Learned Image Compression with Hierarchical Progressive Context Modeling

Yuqi Li, Haotian Zhang, Li Li, Dong Liu

机构 * MOE Key Laboratory of Brain-Inspired Intelligent Perception and Cognition(脑启发智能感知与认知国家重点实验室) University of Science and Technology of China(中国科学技术大学)

Comments 17 pages, ICCV 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2507.19002 2025-07-28 cs.CV

Enhancing Reward Models for High-quality Image Generation: Beyond Text-Image Alignment

Ying Ba, Tianyu Zhang, Yalong Bai, Wenyi Mo, Tao Liang, Bing Su, Ji-Rong Wen

机构 * Gaoling School of Artificial Intelligence(中关村人工智能学院) Renmin University of China(中国人民大学) Beijing Key Laboratory of Research on Large Models and Intelligent Governance(北京大模型与智能治理重点实验室) Engineering Research Center of Next-Generation Intelligent Search and Recommendation, MOE(下一代智能搜索与推荐工程研究中心,教育部) iN2X

Comments Accepted to ICCV 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2507.18997 2025-07-28 cs.CV

UPP: Unified Point-Level Prompting for Robust Point Cloud Analysis

Zixiang Ai, Zhenyu Cui, Yuxin Peng, Jiahuan Zhou

机构 * Wangxuan Institute of Computer Technology, Peking University(计算机技术研究院,北京大学)

Comments Accepted by ICCV 2025 as a Poster

详情

展开后加载摘要…

URL PDF HTML 收藏
2507.18838 2025-07-28 cs.CV cs.AI stat.ML

Flow Stochastic Segmentation Networks

Fabio De Sousa Ribeiro, Omar Todd, Charles Jones, Avinash Kori, Raghav Mehta, Ben Glocker

Comments Accepted at ICCV 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2507.18678 2025-07-28 cs.CV cs.AI

Towards Scalable Spatial Intelligence via 2D-to-3D Data Lifting

Xingyu Miao, Haoran Duan, Quanhao Qian, Jiuniu Wang, Yang Long, Ling Shao, Deli Zhao, Ran Xu, Gongjie Zhang

机构 * Durham University(杜ham大学) DAMO Academy, Alibaba Group(达摩院,阿里集团) Tsinghua University(清华大学) UCAS-Terminus AI Lab(北航-terminate人工智能实验室)

Comments ICCV 2025 (Highlight)

详情

展开后加载摘要…

URL PDF HTML 收藏
2507.18653 2025-07-28 cs.CV cs.LG

Adapt, But Don't Forget: Fine-Tuning and Contrastive Routing for Lane Detection under Distribution Shift

Mohammed Abdul Hafeez Khan, Parth Ganeriwala, Sarah M. Lehman, Siddhartha Bhattacharyya, Amy Alvarez, Natasha Neogi

机构 * Florida Institute of Technology(佛罗里达理工学院) NASA Langley Research Center(美国国家航空航天局兰利研究中心)

Comments Accepted to ICCV 2025, 2COOOL Workshop. Total 14 pages, 5 tables, and 4 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2507.18192 2025-07-28 cs.CV

TeEFusion: Blending Text Embeddings to Distill Classifier-Free Guidance

Minghao Fu, Guo-Hua Wang, Xiaohao Chen, Qing-Guo Chen, Zhao Xu, Weihua Luo, Kaifu Zhang

机构 * School of Artificial Intelligence, Nanjing University(南京大学人工智能学院) National Key Laboratory for Novel Software Technology, Nanjing University(南京大学新型软件技术国家实验室) Alibaba International Digital Commerce Group(阿里巴巴国际数字商业集团)

Comments Accepted by ICCV 2025. The code is publicly available at https://github.com/AIDC-AI/TeEFusion

详情

展开后加载摘要…

URL PDF HTML 收藏
2507.06230 2025-07-28 cs.CV

Feed-Forward SceneDINO for Unsupervised Semantic Scene Completion

Aleksandar Jevtić, Christoph Reich, Felix Wimbauer, Oliver Hahn, Christian Rupprecht, Stefan Roth, Daniel Cremers

Comments ICCV 2025. Christoph Reich and Aleksandar Jevtić - both authors contributed equally. Code: https://github.com/tum-vision/scenedino Project page: https://visinf.github.io/scenedino

详情

展开后加载摘要…

URL PDF HTML 收藏