arXivDaily arXiv每日学术速递 周一至周五更新

期刊&会议

International Conference on Computer Vision · 会议 · Computer Vision

2025-08-12 至 2025-08-12 共收录 32
2508.08254 2025-08-12 cs.CV

Learning an Implicit Physics Model for Image-based Fluid Simulation

Emily Yue-Ting Jia, Jiageng Mao, Zhiyuan Gao, Yajie Zhao, Yue Wang

机构 * University of Southern California(南加州大学)

Comments Accepted at ICCV 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.08165 2025-08-12 cs.CV cs.LG

Integrating Task-Specific and Universal Adapters for Pre-Trained Model-based Class-Incremental Learning

Yan Wang, Da-Wei Zhou, Han-Jia Ye

机构 * School of Artificial Intelligence, Nanjing University(人工智能学院,南京大学) National Key Laboratory for Novel Software Technology, Nanjing University(新型软件技术国家重点实验室,南京大学)

Comments Accepted to ICCV 2025. Code is available at: https://github.com/LAMDA-CL/ICCV2025-TUNA

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.07989 2025-08-12 cs.CV cs.HC

The Escalator Problem: Identifying Implicit Motion Blindness in AI for Accessibility

Xiantao Zhang

机构 * Beihang University(北航大学)

Comments 9 pages, 3 figures, 2 tables. Accepted at CV4A11y, ICCV 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.07877 2025-08-12 cs.CV cs.AI

Selective Contrastive Learning for Weakly Supervised Affordance Grounding

WonJun Moon, Hyun Seok Seong, Jae-Pil Heo

机构 * Sungkyunkwan University(全北大学)

Comments Accepted to ICCV 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.07850 2025-08-12 cs.CV

Morphological Analysis of Semiconductor Microstructures using Skeleton Graphs

Noriko Nitta, Rei Miyata, Naoto Oishi

机构 * Kochi University of Technology(Kochi技术大学) National Institute of Technology, Kochi College(Kochi国立技术大学学院)

Comments CV4MS: Computer Vision for Materials Science, Workshop in conjunction with the IEEE/CVF ICCV 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.07847 2025-08-12 cs.CV cs.AI

Deep Space Weather Model: Long-Range Solar Flare Prediction from Multi-Wavelength Images

Shunya Nagashima, Komei Sugiura

机构 * Keio University(Keio大学)

Comments ICCV 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.07775 2025-08-12 cs.CV

Forecasting Continuous Non-Conservative Dynamical Systems in SO(3)

Lennart Bastian, Mohammad Rashed, Nassir Navab, Tolga Birdal

机构 * Technical University of Munich(慕尼黑技术大学) Munich Center of Machine Learning(慕尼黑机器学习中心) Imperial College London(伦敦帝国学院)

Comments ICCV 2025 Oral

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.07747 2025-08-12 cs.CV

Grouped Speculative Decoding for Autoregressive Image Generation

Junhyuk So, Juncheol Shin, Hyunho Kook, Eunhyeok Park

机构 * Department of Computer Science and Engineering, POSTECH(计算机科学与工程系,POSTECH) Graduate School of Artificial Intelligence, POSTECH(人工智能研究生院,POSTECH)

Comments Accepted to the ICCV 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.07647 2025-08-12 cs.CV

LaRender: Training-Free Occlusion Control in Image Generation via Latent Rendering

Xiaohang Zhan, Dingming Liu

机构 * Tencent(腾讯)

Comments Accepted by ICCV 2025 (oral). Project page: https://xiaohangzhan.github.io/projects/larender/

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.07519 2025-08-12 cs.CV

Exploring Multimodal Diffusion Transformers for Enhanced Prompt-based Image Editing

Joonghyuk Shin, Alchan Hwang, Yujin Kim, Daneul Kim, Jaesik Park

机构 * Seoul National University(首尔国立大学)

Comments ICCV 2025. Project webpage: https://joonghyuk.com/exploring-mmdit-web/

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.02516 2025-08-12 cs.CV

Engagement Prediction of Short Videos with Large Multimodal Models

Wei Sun, Linhan Cao, Yuqin Cao, Weixia Zhang, Wen Wen, Kaiwei Zhang, Zijian Chen, Fangfang Lu, Xiongkuo Min, Guangtao Zhai

机构 * East China Normal University(华东师范大学) Shanghai Jiao Tong University(上海交通大学) City University of Hong Kong(香港城市大学) Shanghai University of Electric Power(上海电力大学)

Comments The proposed method achieves first place in the ICCV VQualA 2025 EVQA-SnapUGC Challenge on short-form video engagement prediction

详情

展开后加载摘要…

URL PDF HTML 收藏
2507.19993 2025-08-12 cs.CV

FROSS: Faster-than-Real-Time Online 3D Semantic Scene Graph Generation from RGB-D Images

Hao-Yu Hou, Chun-Yi Lee, Motoharu Sonogashira, Yasutomo Kawanishi

机构 * National Tsing Hua University(国立清华大学) National Taiwan University(国立台湾大学) RIKEN(日本研究机构)

Comments International Conference on Computer Vision (ICCV 2025)

详情

展开后加载摘要…

URL PDF HTML 收藏
2507.15709 2025-08-12 cs.CV

Efficient Face Image Quality Assessment via Self-training and Knowledge Distillation

Wei Sun, Weixia Zhang, Linhan Cao, Jun Jia, Xiangyang Zhu, Dandan Zhu, Xiongkuo Min, Guangtao Zhai

机构 * East China Normal University(东华师范大学) Shanghai Jiao Tong University(上海交通大学) Shanghai Artificial Intelligence Laboratory(上海人工智能实验室)

Comments Efficient-FIQA achieved first place in the ICCV VQualA 2025 Face Image Quality Assessment Challenge

详情

展开后加载摘要…

URL PDF HTML 收藏
2505.23186 2025-08-12 cs.CV

HiGarment: Cross-modal Harmony Based Diffusion Model for Flat Sketch to Realistic Garment Image

Junyi Guo, Jingxuan Zhang, Fangyu Wu, Huanda Lu, Qiufeng Wang, Wenmian Yang, Eng Gee Lim, Dongming Lu

机构 * Xi’an Jiaotong Liverpool University(西安交通大学利物浦大学) NingboTech University(宁波科技学院) Beijing Normal University(北京师范大学) Zhejiang University(浙江大学)

Comments ICCV 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2505.05591 2025-08-12 cs.CV

QuickSplat: Fast 3D Surface Reconstruction via Learned Gaussian Initialization

Yueh-Cheng Liu, Lukas Höllein, Matthias Nießner, Angela Dai

Comments ICCV 2025. Project page: https://liu115.github.io/quicksplat, Video: https://youtu.be/2IA_gnFvFG8

详情

展开后加载摘要…

URL PDF HTML 收藏
2412.05148 2025-08-12 cs.CV cs.AI cs.LG

LoRA.rar: Learning to Merge LoRAs via Hypernetworks for Subject-Style Conditioned Image Generation

Donald Shenaj, Ondrej Bohdal, Mete Ozay, Pietro Zanuttigh, Umberto Michieli

Comments ICCV 2025. Project page: https://donaldssh.github.io/LoRA.rar

详情

展开后加载摘要…

URL PDF HTML 收藏
2410.12777 2025-08-12 cs.CV cs.CL cs.CR cs.LG

Meta-Unlearning on Diffusion Models: Preventing Relearning Unlearned Concepts

Hongcheng Gao, Tianyu Pang, Chao Du, Taihang Hu, Zhijie Deng, Min Lin

机构 * Sea AI Lab, Singapore(新加坡Sea AI实验室) University of Chinese Academy of Sciences(中国科学院大学) Shanghai Jiao Tong University(上海交通大学) Nankai University(南开大学)

Comments ICCV 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2403.09333 2025-08-12 cs.CV cs.AI

Griffon v2: Advancing Multimodal Perception with High-Resolution Scaling and Visual-Language Co-Referring

Yufei Zhan, Shurong Zheng, Yousong Zhu, Hongyin Zhao, Fan Yang, Ming Tang, Jinqiao Wang

机构 * Foundation Model Research Center, Institute of Automation, Chinese Academy of Sciences(中国科学院自动化研究所基础模型研究中心) School of Artificial Intelligence, University of Chinese Academy of Sciences(中国科学院大学人工智能学院) Peng Cheng Laboratory, Shenzhen, China(鹏城实验室) Wuhan AI Research, Wuhan, China(武汉人工智能研究所)

Comments Accepted by ICCV 2025. Codes and datasets are released at https://github.com/jefferyZhan/Griffon

详情

展开后加载摘要…

URL PDF HTML 收藏
2311.13621 2025-08-12 cs.CV

EA-KD: Entropy-based Adaptive Knowledge Distillation

Chi-Ping Su, Ching-Hsun Tseng, Bin Pu, Lei Zhao, Jiewen Yang, Zhuangzhuang Chen, Shin-Jye Lee

Comments Accepted to ICCV 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.07089 2025-08-12 cs.CV cs.RO

ForeSight: Multi-View Streaming Joint Object Detection and Trajectory Forecasting

Sandro Papais, Letian Wang, Brian Cheong, Steven L. Waslander

机构 * University of Toronto(多伦多大学)

Comments Accepted to ICCV 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.06895 2025-08-12 cs.CV cs.AI

BASIC: Boosting Visual Alignment with Intrinsic Refined Embeddings in Multimodal Large Language Models

Jianting Tang, Yubo Wang, Haoyu Cao, Linli Xu

机构 * University of Science and Technology of China(中国科学技术大学) State Key Laboratory of Cognitive Intelligence(认知智能国家重点实验室)

Comments Accepted to ICCV 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.06546 2025-08-12 cs.CV eess.IV

Statistical Confidence Rescoring for Robust 3D Scene Graph Generation from Multi-View Images

Qi Xun Yeo, Yanyan Li, Gim Hee Lee

机构 * Department of Computer Science, National University of Singapore(计算机科学系,新加坡国立大学)

Comments This paper has been accepted in ICCV 25

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.02278 2025-08-12 cs.CV

SGAD: Semantic and Geometric-aware Descriptor for Local Feature Matching

Xiangzeng Liu, Chi Wang, Guanglu Shi, Xiaodong Zhang, Qiguang Miao, Miao Fan

机构 * Xidian University(西电大学) Navinfo Europe B.V(Navinfo欧洲有限公司)

Comments Accepted to ICCV 2025 (Highlight)

详情

展开后加载摘要…

URL PDF HTML 收藏
2507.23597 2025-08-12 cs.CV

MoGA: 3D Generative Avatar Prior for Monocular Gaussian Avatar Reconstruction

Zijian Dong, Longteng Duan, Jie Song, Michael J. Black, Andreas Geiger

机构 * ETH Zürich(苏黎世联邦理工学院) University of Tübingen, Tübingen AI Center(图宾根大学图宾根人工智能中心) Max Planck Institute for Intelligent Systems, Tübingen(智能系统马克斯·普朗克研究所)

Comments ICCV 2025 (Highlight), Project Page: https://zj-dong.github.io/MoGA/

详情

展开后加载摘要…

URL PDF HTML 收藏
2507.06272 2025-08-12 cs.CV cs.AI

LIRA: Inferring Segmentation in Large Multi-modal Models with Local Interleaved Region Assistance

Zhang Li, Biao Yang, Qiang Liu, Shuo Zhang, Zhiyin Ma, Liang Yin, Linger Deng, Yabo Sun, Yuliang Liu, Xiang Bai

机构 * Huazhong University of Science and Technology(华中科技大学)

Comments ICCV 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2507.03541 2025-08-12 cs.CV cs.AI

Foundation versus Domain-specific Models: Performance Comparison, Fusion, and Explainability in Face Recognition

Redwan Sony, Parisa Farmanifard, Arun Ross, Anil K. Jain

机构 * Michigan State University(密歇根州立大学)

Comments Accepted at the International Conference on Computer Vision (ICCV) 2025 Workshop

详情

展开后加载摘要…

URL PDF HTML 收藏
2506.21249 2025-08-12 cs.CV

Temporal Rate Reduction Clustering for Human Motion Segmentation

Xianghan Meng, Zhengyu Tong, Zhiyuan Huang, Chun-Guang Li

机构 * Beijing University of Posts and Telecommunications(北京邮电大学)

Comments The paper is accepted by ICCV 2025. The first two authors are equally contributed. Camera-ready version uploaded

详情

展开后加载摘要…

URL PDF HTML 收藏
2504.07951 2025-08-12 cs.CV

Scaling Laws for Native Multimodal Models

Mustafa Shukor, Enrico Fini, Victor Guilherme Turrisi da Costa, Matthieu Cord, Joshua Susskind, Alaaeldin El-Nouby

Comments ICCV 2025 (Oral). 28 figures, 13 tables

详情

展开后加载摘要…

URL PDF HTML 收藏
2503.13236 2025-08-12 cs.LG cs.CV

Gradient Extrapolation for Debiased Representation Learning

Ihab Asaad, Maha Shadaydeh, Joachim Denzler

机构 * Computer Vision Group, Friedrich Schiller University Jena(杰纳大学计算机视觉组)

Comments Accepted at International Conference on Computer Vision, ICCV 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2503.12929 2025-08-12 cs.CV

AR-1-to-3: Single Image to Consistent 3D Object Generation via Next-View Prediction

Xuying Zhang, Yupeng Zhou, Kai Wang, Yikai Wang, Zhen Li, Shaohui Jiao, Daquan Zhou, Qibin Hou, Ming-Ming Cheng

机构 * VCIP, CS, Nankai University(南开大学计算机科学与技术学院) NKIARI, Shenzhen Futian(深圳南山人工智能研究院) Tsinghua University(清华大学) ByteDance Inc.(字节跳动公司)

Comments Accepted at ICCV 2025; Project page: https://github.com/HVision-NKU/AR123

详情

展开后加载摘要…

URL PDF HTML 收藏