arXivDaily arXiv每日学术速递 周一至周五更新

期刊&会议

International Conference on Computer Vision · 会议 · Computer Vision

共收录 4770
2507.15709 2025-08-12 cs.CV

Efficient Face Image Quality Assessment via Self-training and Knowledge Distillation

Wei Sun, Weixia Zhang, Linhan Cao, Jun Jia, Xiangyang Zhu, Dandan Zhu, Xiongkuo Min, Guangtao Zhai

机构 * East China Normal University(东华师范大学) Shanghai Jiao Tong University(上海交通大学) Shanghai Artificial Intelligence Laboratory(上海人工智能实验室)

Comments Efficient-FIQA achieved first place in the ICCV VQualA 2025 Face Image Quality Assessment Challenge

详情

展开后加载摘要…

URL PDF HTML 收藏
2505.23186 2025-08-12 cs.CV

HiGarment: Cross-modal Harmony Based Diffusion Model for Flat Sketch to Realistic Garment Image

Junyi Guo, Jingxuan Zhang, Fangyu Wu, Huanda Lu, Qiufeng Wang, Wenmian Yang, Eng Gee Lim, Dongming Lu

机构 * Xi’an Jiaotong Liverpool University(西安交通大学利物浦大学) NingboTech University(宁波科技学院) Beijing Normal University(北京师范大学) Zhejiang University(浙江大学)

Comments ICCV 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2505.05591 2025-08-12 cs.CV

QuickSplat: Fast 3D Surface Reconstruction via Learned Gaussian Initialization

Yueh-Cheng Liu, Lukas Höllein, Matthias Nießner, Angela Dai

Comments ICCV 2025. Project page: https://liu115.github.io/quicksplat, Video: https://youtu.be/2IA_gnFvFG8

详情

展开后加载摘要…

URL PDF HTML 收藏
2412.05148 2025-08-12 cs.CV cs.AI cs.LG

LoRA.rar: Learning to Merge LoRAs via Hypernetworks for Subject-Style Conditioned Image Generation

Donald Shenaj, Ondrej Bohdal, Mete Ozay, Pietro Zanuttigh, Umberto Michieli

Comments ICCV 2025. Project page: https://donaldssh.github.io/LoRA.rar

详情

展开后加载摘要…

URL PDF HTML 收藏
2410.12777 2025-08-12 cs.CV cs.CL cs.CR cs.LG

Meta-Unlearning on Diffusion Models: Preventing Relearning Unlearned Concepts

Hongcheng Gao, Tianyu Pang, Chao Du, Taihang Hu, Zhijie Deng, Min Lin

机构 * Sea AI Lab, Singapore(新加坡Sea AI实验室) University of Chinese Academy of Sciences(中国科学院大学) Shanghai Jiao Tong University(上海交通大学) Nankai University(南开大学)

Comments ICCV 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2403.09333 2025-08-12 cs.CV cs.AI

Griffon v2: Advancing Multimodal Perception with High-Resolution Scaling and Visual-Language Co-Referring

Yufei Zhan, Shurong Zheng, Yousong Zhu, Hongyin Zhao, Fan Yang, Ming Tang, Jinqiao Wang

机构 * Foundation Model Research Center, Institute of Automation, Chinese Academy of Sciences(中国科学院自动化研究所基础模型研究中心) School of Artificial Intelligence, University of Chinese Academy of Sciences(中国科学院大学人工智能学院) Peng Cheng Laboratory, Shenzhen, China(鹏城实验室) Wuhan AI Research, Wuhan, China(武汉人工智能研究所)

Comments Accepted by ICCV 2025. Codes and datasets are released at https://github.com/jefferyZhan/Griffon

详情

展开后加载摘要…

URL PDF HTML 收藏
2311.13621 2025-08-12 cs.CV

EA-KD: Entropy-based Adaptive Knowledge Distillation

Chi-Ping Su, Ching-Hsun Tseng, Bin Pu, Lei Zhao, Jiewen Yang, Zhuangzhuang Chen, Shin-Jye Lee

Comments Accepted to ICCV 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.07089 2025-08-12 cs.CV cs.RO

ForeSight: Multi-View Streaming Joint Object Detection and Trajectory Forecasting

Sandro Papais, Letian Wang, Brian Cheong, Steven L. Waslander

机构 * University of Toronto(多伦多大学)

Comments Accepted to ICCV 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.06895 2025-08-12 cs.CV cs.AI

BASIC: Boosting Visual Alignment with Intrinsic Refined Embeddings in Multimodal Large Language Models

Jianting Tang, Yubo Wang, Haoyu Cao, Linli Xu

机构 * University of Science and Technology of China(中国科学技术大学) State Key Laboratory of Cognitive Intelligence(认知智能国家重点实验室)

Comments Accepted to ICCV 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.06546 2025-08-12 cs.CV eess.IV

Statistical Confidence Rescoring for Robust 3D Scene Graph Generation from Multi-View Images

Qi Xun Yeo, Yanyan Li, Gim Hee Lee

机构 * Department of Computer Science, National University of Singapore(计算机科学系,新加坡国立大学)

Comments This paper has been accepted in ICCV 25

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.02278 2025-08-12 cs.CV

SGAD: Semantic and Geometric-aware Descriptor for Local Feature Matching

Xiangzeng Liu, Chi Wang, Guanglu Shi, Xiaodong Zhang, Qiguang Miao, Miao Fan

机构 * Xidian University(西电大学) Navinfo Europe B.V(Navinfo欧洲有限公司)

Comments Accepted to ICCV 2025 (Highlight)

详情

展开后加载摘要…

URL PDF HTML 收藏
2507.23597 2025-08-12 cs.CV

MoGA: 3D Generative Avatar Prior for Monocular Gaussian Avatar Reconstruction

Zijian Dong, Longteng Duan, Jie Song, Michael J. Black, Andreas Geiger

机构 * ETH Zürich(苏黎世联邦理工学院) University of Tübingen, Tübingen AI Center(图宾根大学图宾根人工智能中心) Max Planck Institute for Intelligent Systems, Tübingen(智能系统马克斯·普朗克研究所)

Comments ICCV 2025 (Highlight), Project Page: https://zj-dong.github.io/MoGA/

详情

展开后加载摘要…

URL PDF HTML 收藏
2507.06272 2025-08-12 cs.CV cs.AI

LIRA: Inferring Segmentation in Large Multi-modal Models with Local Interleaved Region Assistance

Zhang Li, Biao Yang, Qiang Liu, Shuo Zhang, Zhiyin Ma, Liang Yin, Linger Deng, Yabo Sun, Yuliang Liu, Xiang Bai

机构 * Huazhong University of Science and Technology(华中科技大学)

Comments ICCV 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2507.03541 2025-08-12 cs.CV cs.AI

Foundation versus Domain-specific Models: Performance Comparison, Fusion, and Explainability in Face Recognition

Redwan Sony, Parisa Farmanifard, Arun Ross, Anil K. Jain

机构 * Michigan State University(密歇根州立大学)

Comments Accepted at the International Conference on Computer Vision (ICCV) 2025 Workshop

详情

展开后加载摘要…

URL PDF HTML 收藏
2506.21249 2025-08-12 cs.CV

Temporal Rate Reduction Clustering for Human Motion Segmentation

Xianghan Meng, Zhengyu Tong, Zhiyuan Huang, Chun-Guang Li

机构 * Beijing University of Posts and Telecommunications(北京邮电大学)

Comments The paper is accepted by ICCV 2025. The first two authors are equally contributed. Camera-ready version uploaded

详情

展开后加载摘要…

URL PDF HTML 收藏
2504.07951 2025-08-12 cs.CV

Scaling Laws for Native Multimodal Models

Mustafa Shukor, Enrico Fini, Victor Guilherme Turrisi da Costa, Matthieu Cord, Joshua Susskind, Alaaeldin El-Nouby

Comments ICCV 2025 (Oral). 28 figures, 13 tables

详情

展开后加载摘要…

URL PDF HTML 收藏
2503.13236 2025-08-12 cs.LG cs.CV

Gradient Extrapolation for Debiased Representation Learning

Ihab Asaad, Maha Shadaydeh, Joachim Denzler

机构 * Computer Vision Group, Friedrich Schiller University Jena(杰纳大学计算机视觉组)

Comments Accepted at International Conference on Computer Vision, ICCV 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2503.12929 2025-08-12 cs.CV

AR-1-to-3: Single Image to Consistent 3D Object Generation via Next-View Prediction

Xuying Zhang, Yupeng Zhou, Kai Wang, Yikai Wang, Zhen Li, Shaohui Jiao, Daquan Zhou, Qibin Hou, Ming-Ming Cheng

机构 * VCIP, CS, Nankai University(南开大学计算机科学与技术学院) NKIARI, Shenzhen Futian(深圳南山人工智能研究院) Tsinghua University(清华大学) ByteDance Inc.(字节跳动公司)

Comments Accepted at ICCV 2025; Project page: https://github.com/HVision-NKU/AR123

详情

展开后加载摘要…

URL PDF HTML 收藏
2503.06134 2025-08-12 cs.CV

X2I: Seamless Integration of Multimodal Understanding into Diffusion Transformer via Attention Distillation

Jian Ma, Qirong Peng, Xu Guo, Chen Chen, Haonan Lu, Zhenyu Yang

机构 * OPPO AI Center(OPPO人工智能中心) Tsinghua University(清华大学)

Comments Accepted to ICCV 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2412.16919 2025-08-12 cs.CV

TAR3D: Creating High-Quality 3D Assets via Next-Part Prediction

Xuying Zhang, Yutong Liu, Yangguang Li, Renrui Zhang, Yufei Liu, Kai Wang, Wanli Ouyang, Zhiwei Xiong, Peng Gao, Qibin Hou, Ming-Ming Cheng

机构 * VCIP, CS, Nankai University(南开大学计算机科学与技术学院) NKIARI, Shenzhen Futian(深圳未来科技研究院) USTC(University of Science and Technology of China) CUHK MMLab(香港中文大学MMLab) VAST(中国科学院自动化研究所) Shanghai AI Lab(上海人工智能实验室)

Comments Accepted at ICCV 2025. Project page: https://github.com/HVision-NKU/TAR3D

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.06494 2025-08-11 cs.CV

LightSwitch: Multi-view Relighting with Material-guided Diffusion

Yehonathan Litman, Fernando De la Torre, Shubham Tulsiani

机构 * Carnegie Mellon University(卡内基梅隆大学)

Comments ICCV 2025, Project page & Code: https://yehonathanlitman.github.io/light_switch/

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.06492 2025-08-11 cs.CV cs.CL

Effective Training Data Synthesis for Improving MLLM Chart Understanding

Yuwei Yang, Zeyu Zhang, Yunzhong Hou, Zhuowan Li, Gaowen Liu, Ali Payani, Yuan-Sen Ting, Liang Zheng

机构 * Australian National University(澳大利亚国立大学) Ohio State University(俄亥俄州立大学) Cisco(思科公司) Johns Hopkins University(约翰霍普金斯大学)

Comments Accepted by ICCV 2025 (poster). 26 pages, 17 figures

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.06325 2025-08-11 cs.CR cs.CV

Anti-Tamper Protection for Unauthorized Individual Image Generation

Zelin Li, Ruohan Zong, Yifan Liu, Ruichen Yao, Yaokun Liu, Yang Zhang, Dong Wang

机构 * University of Illinois Urbana-Champaign(伊利诺伊大学厄巴纳-香槟分校)

Comments 22 pages ,22 figures, Paper has been accepted by ICCV'2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.06301 2025-08-11 cs.LG cs.AI cs.CV cs.DC

FedMeNF: Privacy-Preserving Federated Meta-Learning for Neural Fields

Junhyeog Yun, Minui Hong, Gunhee Kim

机构 * Seoul National University(首尔国立大学)

Comments ICCV 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.06160 2025-08-11 cs.CV

Fewer Denoising Steps or Cheaper Per-Step Inference: Towards Compute-Optimal Diffusion Model Deployment

Zhenbang Du, Yonggan Fu, Lifu Wang, Jiayi Qian, Xiao Luo, Yingyan, Lin

机构 * Georgia Institute of Technology(佐治亚理工学院)

Comments Accepted by ICCV 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.06125 2025-08-11 cs.CV

SC-Captioner: Improving Image Captioning with Self-Correction by Reinforcement Learning

Lin Zhang, Xianfang Zeng, Kangcong Li, Gang Yu, Tao Chen

机构 * College of Future Information Technology, Fudan University(复旦大学未来信息科技学院) StepFun Shanghai Innovation Institute(上海创新研究院)

Comments ICCV 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.06033 2025-08-11 cs.CV

InstantEdit: Text-Guided Few-Step Image Editing with Piecewise Rectified Flow

Yiming Gong, Zhen Zhu, Minjia Zhang

机构 * University of Illinois at Urbana-Champaign(伊利诺伊大学厄巴纳-香槟分校)

Comments ICCV 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.06014 2025-08-11 cs.CV

ExploreGS: Explorable 3D Scene Reconstruction with Virtual Camera Samplings and Diffusion Priors

Minsu Kim, Subin Jeon, In Cho, Mijin Yoo, Seon Joo Kim

机构 * Yonsei University(延世大学)

Comments 10 pages, 6 Figures, ICCV 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.05976 2025-08-11 cs.CV cs.RO

PASG: A Closed-Loop Framework for Automated Geometric Primitive Extraction and Semantic Anchoring in Robotic Manipulation

Zhihao Zhu, Yifan Zheng, Siyu Pan, Yaohui Jin, Yao Mu

机构 * MoE key Lab of Artificial Intelligence, AI Institute, Shanghai Jiao Tong University(人工智能教育部重点实验室、人工智能研究院、上海交通大学)

Comments Accepted to ICCV 2025. 8 pages main paper, 8 figures, plus supplementary material

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.05001 2025-08-11 cs.CV cs.LG cs.PF

CRAM: Large-scale Video Continual Learning with Bootstrapped Compression

Shivani Mall, Joao F. Henriques

机构 * Visual Geometry Group, University of Oxford(牛津大学视觉几何组)

Journal ref International Conference on Computer Vision, ICCV 2025

详情

展开后加载摘要…

URL PDF HTML 收藏