arXivDaily arXiv每日学术速递 周一至周五更新

期刊&会议

International Conference on Computer Vision · 会议 · Computer Vision

共收录 4770
2501.06838 2025-07-31 eess.IV cs.CV

Generalized and Efficient 2D Gaussian Splatting for Arbitrary-scale Super-Resolution

Du Chen, Liyi Chen, Zhengqiang Zhang, Lei Zhang

机构 * The Hong Kong Polytechnic University(香港理工大学) OPPO Research Institute(OPPO研究院)

Comments Accepted by ICCV 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2411.17765 2025-07-31 cs.CV

I2VControl: Disentangled and Unified Video Motion Synthesis Control

Wanquan Feng, Tianhao Qi, Jiawei Liu, Mingzhen Sun, Pengqi Tu, Tianxiang Ma, Fei Dai, Songtao Zhao, Siyu Zhou, Qian He

机构 * Intelligent Creation Team, ByteDance(字节跳动智能创作团队) University of Science and Technology of China (USTC)(中国科学技术大学) Institute of Automation, Chinese Academy of Sciences (CASIA)(中国科学院自动化研究所)

Comments Accepted to ICCV 2025. Project page: https://wanquanf.github.io/I2VControl

详情

展开后加载摘要…

URL PDF HTML 收藏
2411.16072 2025-07-31 cs.CV

Language Driven Occupancy Prediction

Zhu Yu, Bowen Pang, Lizhe Liu, Runmin Zhang, Qiang Li, Si-Yuan Cao, Maochun Luo, Mingxia Chen, Sheng Yang, Hui-Liang Shen

机构 * Zhejiang University(浙江大学) Unmanned Vehicle Dept., CaiNiao Inc., Alibaba Group(无人车辆部门,菜鸟公司,阿里巴巴集团) Ningbo Global Innovation Center, Zhejiang University(宁波全球创新中心,浙江大学)

Comments ICCV 2025; Project Page: https://github.com/pkqbajng/LOcc

详情

展开后加载摘要…

URL PDF HTML 收藏
2406.05400 2025-07-31 cs.CV math.DG

Metric Convolutions: A Unifying Theory to Adaptive Image Convolutions

Thomas Dagès, Michael Lindenbaum, Alfred M. Bruckstein

机构 * Technion – Israel Institute of Technology(以色列技术学院) Technical University of Munich(慕尼黑技术大学) Munich Center for Machine Learning(慕尼黑机器学习中心)

Comments Updated version, Accepted for publication at the IEEE/CVF International Conference on Computer Vision (ICCV) 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2507.22100 2025-07-31 cs.CV

Trade-offs in Image Generation: How Do Different Dimensions Interact?

Sicheng Zhang, Binzhu Xie, Zhonghao Yan, Yuli Zhang, Donghao Zhou, Xiaofei Chen, Shi Qiu, Jiaqi Liu, Guoyang Xie, Zhichao Lu

机构 * Khalifa University(卡利法大学) The Chinese University of Hong Kong(香港中文大学) Queen Mary University of London(伦敦大学玛丽女王学院) Xi’an Jiaotong-Liverpool University(西安交通大学利物浦大学) City University of Hong Kong(香港城市大学)

Comments Accepted in ICCV 2025, Codebase: https://github.com/fesvhtr/TRIG

详情

展开后加载摘要…

URL PDF HTML 收藏
2507.22076 2025-07-31 cs.LG

Test-time Prompt Refinement for Text-to-Image Models

Mohammad Abdul Hafeez Khan, Yash Jain, Siddhartha Bhattacharyya, Vibhav Vineet

机构 * Florida Institute of Technology(佛罗里达理工学院) Microsoft Research(微软研究院)

Comments Accepted to ICCV 2025, MARS2 Workshop. Total 14 pages, 12 figures and 3 tables

详情

展开后加载摘要…

URL PDF HTML 收藏
2412.20104 2025-07-31 cs.CV cs.AI cs.LG cs.RO

SyncDiff: Synchronized Motion Diffusion for Multi-Body Human-Object Interaction Synthesis

Wenkun He, Yun Liu, Ruitao Liu, Li Yi

机构 * Tsinghua University(清华大学) Shanghai Qi Zhi Institute(上海启智研究院) Shanghai Artificial Intelligence Laboratory(上海人工智能实验室)

Comments 27 pages, 10 figures, 20 tables. Accepted by ICCV 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2507.22061 2025-07-30 cs.CV

MOVE: Motion-Guided Few-Shot Video Object Segmentation

Kaining Ying, Hengrui Hu, Henghui Ding

机构 * Fudan University(复旦大学)

Comments ICCV 2025, Project Page: https://henghuiding.com/MOVE/

详情

展开后加载摘要…

URL PDF HTML 收藏
2507.21960 2025-07-30 cs.CV

PanoSplatt3R: Leveraging Perspective Pretraining for Generalized Unposed Wide-Baseline Panorama Reconstruction

Jiahui Ren, Mochu Xiang, Jiajun Zhu, Yuchao Dai

机构 * School of Electronics and Information, Northwestern Polytechnical University and Shaanxi Key Laboratory of Information Acquisition and Processing(电子工程学院、西北工业大学和陕西省信息获取与处理重点实验室)

Comments Accepted to ICCV 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2507.21844 2025-07-30 cs.CV

Cross-Architecture Distillation Made Simple with Redundancy Suppression

Weijia Zhang, Yuehao Liu, Wu Ran, Chao Ma

机构 * Shanghai Jiao Tong University(上海交通大学)

Comments Accepted by ICCV 2025 (Highlight)

详情

展开后加载摘要…

URL PDF HTML 收藏
2507.21742 2025-07-30 cs.CV

Adversarial Reconstruction Feedback for Robust Fine-grained Generalization

Shijie Wang, Jian Shi, Haojie Li

机构 * College of Computer and Engineering, Shandong University of Science and Technology(山东科技大学计算机与工程学院) School of Software, Dalian University of Technology(大连理工大学软件学院)

Comments ICCV 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2507.21665 2025-07-30 cs.CV

Automated Detection of Antarctic Benthic Organisms in High-Resolution In Situ Imagery to Aid Biodiversity Monitoring

Cameron Trotter, Huw Griffiths, Tasnuva Ming Khan, Rowan Whittle

机构 * British Antarctic Survey(英国南极调查局) University of Cambridge(剑桥大学)

Comments Accepted to ICCV 2025's Joint Workshop on Marine Vision (ICCVW, CVAUI&AAMVEM). Main paper (11 pages, 3 figures, 3 tables) plus supplementary (7 pages, 5 figures, 2 tables)

详情

展开后加载摘要…

URL PDF HTML 收藏
2507.21588 2025-07-30 cs.AI cs.CV

Progressive Homeostatic and Plastic Prompt Tuning for Audio-Visual Multi-Task Incremental Learning

Jiong Yin, Liang Li, Jiehua Zhang, Yuhan Gao, Chenggang Yan, Xichun Sheng

机构 * Hangzhou Dianzi University(杭州电子科技大学) Institute of Computing Technology, Chinese Academy of Sciences(中国科学院计算技术研究所) Xi’an Jiaotong University(西安交通大学) Macao Polytechnic University(澳门 polytechnic university)

Comments Accepted by ICCV 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2507.21494 2025-07-30 cs.LG

Latte: Collaborative Test-Time Adaptation of Vision-Language Models in Federated Learning

Wenxuan Bao, Ruxi Deng, Ruizhong Qiu, Tianxin Wei, Hanghang Tong, Jingrui He

机构 * University of Illinois Urbana-Champaign(伊利诺伊大学厄巴纳-香槟分校)

Comments Accepted by ICCV 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2507.21489 2025-07-30 cs.CV

Describe, Adapt and Combine: Empowering CLIP Encoders for Open-set 3D Object Retrieval

Zhichuan Wang, Yang Zhou, Zhe Liu, Rui Yu, Song Bai, Yulong Wang, Xinwei He, Xiang Bai

机构 * Huazhong Agricultural University(华中农业大学) Shenzhen University(深圳大学) The University of Hong Kong(香港大学) University of Louisville(路易斯安那大学) ByteDance(字节跳动) Huazhong University of Science and Technology(华中科技大学)

Comments Accepted to ICCV 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2507.21371 2025-07-30 cs.CV

Top2Pano: Learning to Generate Indoor Panoramas from Top-Down View

Zitong Zhang, Suranjan Gautam, Rui Yu

机构 * University of Louisville(路易斯维尔大学)

Comments ICCV 2025. Project page: https://top2pano.github.io/

详情

展开后加载摘要…

URL PDF HTML 收藏
2507.20987 2025-07-30 cs.CV cs.AI

JWB-DH-V1: Benchmark for Joint Whole-Body Talking Avatar and Speech Generation Version 1

Xinhan Di, Kristin Qi, Pengqian Yu

机构 * Computer Science, University of Massachusetts Boston(马萨诸塞大学波士顿分校计算机科学系) National University of Singapore(新加坡国立大学)

Comments WiCV @ ICCV 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2507.20536 2025-07-30 cs.CV cs.AI cs.HC

T2I-Copilot: A Training-Free Multi-Agent Text-to-Image System for Enhanced Prompt Interpretation and Interactive Generation

Chieh-Yun Chen, Min Shi, Gong Zhang, Humphrey Shi

Comments ICCV 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2507.19141 2025-07-30 cs.CV

DASH: 4D Hash Encoding with Self-Supervised Decomposition for Real-Time Dynamic Scene Rendering

Jie Chen, Zhangchi Hu, Peixi Wu, Huyue Zhu, Hebei Li, Xiaoyan Sun

机构 * University of Science and Technology of China(中国科学技术大学) Institute of Artificial Intelligence, Hefei Comprehensive National Science Center(合肥综合国家科学中心人工智能研究院)

Comments ICCV 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2507.12857 2025-07-30 cs.CV

SCORE: Scene Context Matters in Open-Vocabulary Remote Sensing Instance Segmentation

Shiqi Huang, Shuting He, Huaiyuan Qin, Bihan Wen

机构 * Nanyang Technological University(南洋理工大学) MoE Key Laboratory of Interdisciplinary Research of Computation and Economics, Shanghai University of Finance and Economics(上海财经大学经济与计算交叉研究重点实验室) Institute for Infocomm Research (I 2 R), A*STAR, Singapore(新加坡资讯与通讯研究院)

Comments ICCV 2025 (Highlight), code see https://github.com/HuangShiqi128/SCORE

详情

展开后加载摘要…

URL PDF HTML 收藏
2507.10095 2025-07-30 cs.CV

FIX-CLIP: Dual-Branch Hierarchical Contrastive Learning via Synthetic Captions for Better Understanding of Long Text

Bingchao Wang, Zhiwei Ning, Jianyu Ding, Xuanang Gao, Yin Li, Dongsheng Jiang, Jie Yang, Wei Liu

机构 * Shanghai Jiao Tong University(上海交通大学) Huawei Inc.(华为公司)

Comments Accepted by ICCV 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2506.02751 2025-07-30 cs.CV

RobustSplat: Decoupling Densification and Dynamics for Transient-Free 3DGS

Chuanyu Fu, Yuqi Zhang, Kunbin Yao, Guanying Chen, Yuan Xiong, Chuan Huang, Shuguang Cui, Xiaochun Cao

机构 * Sun Yat-sen University(中山大学) FNii-Shenzhen(FNii-深圳) SSE, CUHKSZ(CUHKSZ信息科技学院) Guangdong Key Laboratory of Information Security Technology(广东省信息安全技术重点实验室)

Comments ICCV 2025. Project page: https://fcyycf.github.io/RobustSplat/

详情

展开后加载摘要…

URL PDF HTML 收藏
2503.16418 2025-07-30 cs.CV cs.LG

InfiniteYou: Flexible Photo Recrafting While Preserving Your Identity

Liming Jiang, Qing Yan, Yumin Jia, Zichuan Liu, Hao Kang, Xin Lu

机构 * ByteDance Intelligent Creation Project(字节跳动智能创作项目)

Comments ICCV 2025 (Highlight). Project page: https://bytedance.github.io/InfiniteYou/ Code and model: https://github.com/bytedance/InfiniteYou

详情

展开后加载摘要…

URL PDF HTML 收藏
2503.09631 2025-07-30 cs.GR eess.IV

V2M4: 4D Mesh Animation Reconstruction from a Single Monocular Video

Jianqi Chen, Biao Zhang, Xiangjun Tang, Peter Wonka

Comments Accepted by ICCV 2025. Project page: https://windvchen.github.io/V2M4/

详情

展开后加载摘要…

URL PDF HTML 收藏
2501.08654 2025-07-30 cs.CV

ZeroStereo: Zero-shot Stereo Matching from Single Images

Xianqi Wang, Hao Yang, Gangwei Xu, Junda Cheng, Min Lin, Yong Deng, Jinliang Zang, Yurui Chen, Xin Yang

机构 * Huazhong University of Science and Technology(华中科技大学) Autel Robotics(Autel机器人公司) Optics Valley Laboratory(光学谷实验室)

Comments Accepted to ICCV 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2501.04873 2025-07-30 cs.CV cs.AI cs.LG

Back Home: A Computer Vision Solution to Seashell Identification for Ecological Restoration

Alexander Valverde, Luis Solano, André Montoya

机构 * FIFCO

Comments ICCV 2025 (CV4E Workshop)

详情

展开后加载摘要…

URL PDF HTML 收藏
2412.06340 2025-07-30 cs.CV

UniPaint: Unified Space-time Video Inpainting via Mixture-of-Experts

Zhen Wan, Chenyang Qi, Zhiheng Liu, Tao Gui, Yue Ma

机构 * Fudan University(复旦大学) HKUST(香港科技大学) HKU(香港大学)

Comments ICCV 1st Workshop on Human-Interactive Generation and Editing (poster)

详情

展开后加载摘要…

URL PDF HTML 收藏
2412.03347 2025-07-30 cs.CV cs.AI

DIVE: Taming DINO for Subject-Driven Video Editing

Yi Huang, Wei Xiong, He Zhang, Chaoqi Chen, Jianzhuang Liu, Mingfu Yan, Shifeng Chen

机构 * Shenzhen Institutes of Advanced Technology, Chinese Academy of Sciences(深圳先进技术研究院,中国科学院) vivo AI Lab(vivo AI实验室) NVIDIA Adobe Research(Adobe研究院) Shenzhen University(深圳大学) Southeast University(东南大学) Shenzhen University of Advanced Technology(深圳大学先进技术学院)

Comments Accepted by ICCV 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2412.03248 2025-07-30 cs.CV cs.AI cs.CL

AIM: Adaptive Inference of Multi-Modal LLMs via Token Merging and Pruning

Yiwu Zhong, Zhuoming Liu, Yin Li, Liwei Wang

机构 * The Chinese University of Hong Kong(香港中文大学) University of Wisconsin-Madison(威斯康星大学麦迪逊分校)

Comments Accepted to ICCV 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2411.17799 2025-07-30 cs.CV cs.CL

Signs as Tokens: A Retrieval-Enhanced Multilingual Sign Language Generator

Ronglai Zuo, Rolandos Alexandros Potamias, Evangelos Ververas, Jiankang Deng, Stefanos Zafeiriou

机构 * Imperial College London(伦敦帝国学院)

Comments Accepted by ICCV 2025

详情

展开后加载摘要…

URL PDF HTML 收藏