arXivDaily arXiv每日学术速递 周一至周五更新

期刊&会议

International Conference on Computer Vision · 会议 · Computer Vision

共收录 4770
2511.00411 2025-11-04 cs.LG cs.AI cs.CV

Enhancing Adversarial Transferability by Balancing Exploration and Exploitation with Gradient-Guided Sampling

Zenghao Niu, Weicheng Xie, Siyang Song, Zitong Yu, Feng Liu, Linlin Shen

机构 * School of Computer Science & Software Engineering, Shenzhen University, China(深圳大学计算机科学与软件工程学院) Guangdong Laboratory of Artificial Intelligence and Digital Economy (SZ), Shenzhen, China(广东省人工智能与数字经济发展实验室(深圳)) Guangdong Provincial Key Laboratory of Intelligent Information Processing, Shenzhen University, China(广东省智能信息处理省级重点实验室) School of Computer Science, University of Exeter, U.K.(埃克塞特大学计算机科学学院) Department of Computing and Information Technology, Great Bay University, China(大鹏大学计算与信息技术系) Computer Vision Institute, School of Artificial Intelligence, Shenzhen University, China(人工智能学院计算机视觉研究所)

Comments accepted by iccv 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2511.00191 2025-11-04 cs.CV cs.AI cs.LG

A Retrospect to Multi-prompt Learning across Vision and Language

Ziliang Chen, Xin Huang, Quanlong Guan, Liang Lin, Weiqi Luo

机构 * Jinan University(济南大学) Sun Yat-sen University(中山大学) Pazhou Laboratory(琶洲实验室)

Comments ICCV

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.14831 2025-11-04 cs.CV

Scaling Tumor Segmentation: Best Lessons from Real and Synthetic Data

Qi Chen, Xinze Zhou, Chen Liu, Hao Chen, Wenxuan Li, Zekun Jiang, Ziyan Huang, Yuxuan Zhao, Dexin Yu, Junjun He, Yefeng Zheng, Ling Shao, Alan Yuille, Zongwei Zhou

机构 * Johns Hopkins University(约翰霍普金斯大学) UCAS-Terminus AI Lab, University of Chinese Academy of Sciences(中国科学院大学Terminus AI实验室) Hong Kong Polytechnic University(香港理工大学) University of Cambridge(剑桥大学) Sichuan University(四川大学) Shanghai Jiao Tong University(上海交通大学) Shanghai AI Laboratory(上海人工智能实验室) Qilu Hospital of Shandong University(山东大学齐鲁医院) Westlake University(西湖大学)

Comments ICCV 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.08271 2025-11-04 cs.GR cs.CV cs.LG

SViM3D: Stable Video Material Diffusion for Single Image 3D Generation

Andreas Engelhardt, Mark Boss, Vikram Voleti, Chun-Han Yao, Hendrik P. A. Lensch, Varun Jampani

机构 * Stability AI University of Tübingen(图宾根大学)

Comments Accepted by International Conference on Computer Vision (ICCV 2025). Project page: http://svim3d.aengelhardt.com

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.09780 2025-11-04 cs.CV cs.AI

Combinative Matching for Geometric Shape Assembly

Nahyuk Lee, Juhong Min, Junhong Lee, Chunghyun Park, Minsu Cho

机构 * POSTECH Samsung Research America(三星美国研究院) RLWRLD

Comments Accepted to ICCV 2025 (Highlight)

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.27632 2025-11-03 cs.CV cs.AI

Sketch-to-Layout: Sketch-Guided Multimodal Layout Generation

Riccardo Brioschi, Aleksandr Alekseev, Emanuele Nevali, Berkay Döner, Omar El Malki, Blagoj Mitrevski, Leandro Kieliger, Mark Collier, Andrii Maksai, Jesse Berent, Claudiu Musat, Efi Kokiopoulou

机构 * EPFL(苏黎世联邦理工学院) Google DeepMind(谷歌DeepMind)

Comments 15 pages, 18 figures, GitHub link: https://github.com/google-deepmind/sketch_to_layout, accept at ICCV 2025 Workshop (HiGen)

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.21199 2025-11-03 cs.CV

3rd Place Solution to Large-scale Fine-grained Food Recognition

Yang Zhong, Yifan Yao, Tong Luo, Youcai Zhang, Yaqian Li

Journal ref ICCV Workshop LargeFineFoodAI (2021)

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.21198 2025-11-03 cs.CV

3rd Place Solution to ICCV LargeFineFoodAI Retrieval

Yang Zhong, Zhiming Wang, Zhaoyang Li, Jinyu Ma, Xiang Li

Journal ref ICCV Workshop LargeFineFoodAI (2021)

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.04539 2025-11-03 cs.GR cs.CV

C3Editor: Achieving Controllable Consistency in 2D Model for 3D Editing

Zeng Tao, Zheng Ding, Zeyuan Chen, Xiang Zhang, Leizhi Li, Zhuowen Tu

机构 * Fudan University(复旦大学) UC San Diego(加州大学圣达菲分校)

Comments ICCV 2025 Workshop Wild3D

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.26151 2025-10-31 cs.CV cs.AI

MV-MLM: Bridging Multi-View Mammography and Language for Breast Cancer Diagnosis and Risk Prediction

Shunjie-Fabian Zheng, Hyeonjun Lee, Thijs Kooi, Ali Diba

机构 * Department of Medicine I, LMU University Hospital, LMU Munich, Germany(慕尼黑大学医学部第一部门,LMU大学医院,慕尼黑,德国) Lunit Inc.(Lunit公司)

Comments Accepted to Computer Vision for Automated Medical Diagnosis (CVAMD) Workshop at ICCV 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2506.21046 2025-10-31 cs.CV cs.CR

Boosting Generative Adversarial Transferability with Self-supervised Vision Transformer Features

Shangbo Wu, Yu-an Tan, Ruinan Ma, Wencong Ma, Dehua Zhu, Yuanzhang Li

机构 * School of Cyberspace Science and Technology, Beijing Institute of Technology(网络安全科学与技术学院,北京理工大学) School of Computer Science and Technology, Beijing Institute of Technology(计算机科学与技术学院,北京理工大学)

Comments 14 pages, 9 figures, accepted at ICCV 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2501.08325 2025-10-31 cs.CV

GameFactory: Creating New Games with Generative Interactive Videos

Jiwen Yu, Yiran Qin, Xintao Wang, Pengfei Wan, Di Zhang, Xihui Liu

机构 * The University of Hong Kong(香港大学) Kuaishou Technology(快手科技)

Comments ICCV 2025 Highlight, Project Page: https://yujiwen.github.io/gamefactory

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.25051 2025-10-30 cs.CV cs.LG

Breast Cancer VLMs: Clinically Practical Vision-Language Train-Inference Models

Shunjie-Fabian Zheng, Hyeonjun Lee, Thijs Kooi, Ali Diba

机构 * Department of Medicine I, LMU University Hospital, LMU Munich, Germany(慕尼黑大学医学院第一医学部,LMU慕尼黑大学医院) Lunit Inc.(Lunit公司)

Comments Accepted to Computer Vision for Automated Medical Diagnosis (CVAMD) Workshop at ICCV 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.24907 2025-10-30 cs.CV cs.AI cs.LG

Understanding Multi-View Transformers

Michal Stary, Julien Gaubil, Ayush Tewari, Vincent Sitzmann

机构 * TUM(慕尼黑技术大学) Claude Bernard University Lyon 1(里昂第一大学) University of Cambridge(剑桥大学) MIT CSAIL(麻省理工学院计算机科学与人工智能实验室)

Comments Presented at the ICCV 2025 E2E3D Workshop

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.02293 2025-10-29 cs.CV cs.LG

Towards Real Unsupervised Anomaly Detection Via Confident Meta-Learning

Muhammad Aqeel, Shakiba Sharifi, Marco Cristani, Francesco Setti

机构 * Dept. of Engineering for Innovation Medicine, University of Verona(创新医学工程系,威尼斯大学) Qualyco S.r.l.(Qualyco公司)

Comments Accepted to IEEE/CVF International Conference on Computer Vision (ICCV2025)

详情

展开后加载摘要…

URL PDF HTML 收藏
2503.17071 2025-10-29 cs.CV

Superpowering Open-Vocabulary Object Detectors for X-ray Vision

Pablo Garcia-Fernandez, Lorenzo Vaquero, Mingxuan Liu, Feng Xue, Daniel Cores, Nicu Sebe, Manuel Mucientes, Elisa Ricci

机构 * University of Santiago de Compostela(圣地亚哥-德孔波斯特拉大学) Fondazione Bruno Kessler(布鲁诺·凯斯勒基金会) University of Trento(特伦托大学)

Comments Accepted at ICCV 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.24052 2025-10-29 cs.RO cs.AI

SynAD: Enhancing Real-World End-to-End Autonomous Driving Models through Synthetic Data Integration

Jongsuk Kim, Jaeyoung Lee, Gyojin Han, Dongjae Lee, Minki Jeong, Junmo Kim

机构 * KAIST(韩国科学技术院) AI Center, Samsung Electronics(三星电子人工智能中心)

Journal ref International Conference on Computer Vision, ICCV 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.24034 2025-10-29 cs.CV

AutoPrompt: Automated Red-Teaming of Text-to-Image Models via LLM-Driven Adversarial Prompts

Yufan Liu, Wanqian Zhang, Huashan Chen, Lin Wang, Xiaojun Jia, Zheng Lin, Weiping Wang

机构 * Institute of Information Engineering, Chinese Academy of Sciences(中国科学院信息工程研究所) School of Cyber Security, University of Chinese Academy of Sciences(中国科学院大学网络安全学院) School of Cyberspace, Hangzhou Dianzi University(杭州电子科技大学网络空间学院) Nanyang Technological University(南洋理工大学)

Comments Accepted by ICCV 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.17550 2025-10-29 cs.AI cs.CV cs.LG

Is It Certainly a Deepfake? Reliability Analysis in Detection & Generation Ecosystem

Neslihan Kose, Anthony Rhodes, Umur Aybars Ciftci, Ilke Demir

机构 * Intel Labs(英特尔实验室) Binghamton University(宾夕法尼亚州立大学) Cauth AI(Cauth人工智能)

Comments Accepted for publication at the ICCV 2025 workshop - STREAM

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.15256 2025-10-29 cs.CV

Normal and Abnormal Pathology Knowledge-Augmented Vision-Language Model for Anomaly Detection in Pathology Images

Jinsol Song, Jiamu Wang, Anh Tien Nguyen, Keunho Byeon, Sangjeong Ahn, Sung Hak Lee, Jin Tae Kwak

机构 * Korea University(韩国大学) The Catholic University of Korea(韩国天主大学)

Comments Accepted to ICCV 2025. Code is available at: https://github.com/QuIIL/ICCV2025_Ano-NAViLa

详情

展开后加载摘要…

URL PDF HTML 收藏
2506.22802 2025-10-29 cs.LG cs.CR cs.CV

Riemannian-Geometric Fingerprints of Generative Models

Hae Jin Song, Laurent Itti

Comments ICCV 2025 Highlight paper

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.14081 2025-10-28 cs.CV cs.GR

Capture, Canonicalize, Splat: Zero-Shot 3D Gaussian Avatars from Unstructured Phone Images

Emanuel Garbin, Guy Adam, Oded Krams, Zohar Barzelay, Eran Guendelman, Michael Schwarz, Matteo Presutto, Moran Vatelmacher, Yigal Shenkman, Eli Peker, Itai Druker, Uri Patish, Yoav Blum, Max Bluvstein, Junxuan Li, Rawal Khirodkar, Shunsuke Saito

机构 * Meta

Comments This work received the Best Paper Honorable Mention at the AMFG Workshop, ICCV 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.22528 2025-10-28 cs.CV

AesCrop: Aesthetic-driven Cropping Guided by Composition

Yen-Hong Wong, Lai-Kuan Wong

机构 * Faculty of Computing and Informatics(计算与信息学院)

Comments Accepted at the IEEE/CVF International Conference on Computer Vision (ICCV) Workshops, 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.22217 2025-10-28 cs.CV

Enpowering Your Pansharpening Models with Generalizability: Unified Distribution is All You Need

Yongchuan Cui, Peng Liu, Hui Zhang

机构 * Aerospace Information Research Institute, Chinese Academy of Sciences(中国科学院航空航天信息研究所) School of Electronic, Electrical and Communication Engineering, University of Chinese Academy of Sciences(中国科学院大学电子电气与通信工程学院) School of Engineering Medicine, Beihang University(北航工程医学学院)

Comments Accepted to ICCV 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2507.00372 2025-10-28 cs.CV eess.IV

Efficient Depth- and Spatially-Varying Image Simulation for Defocus Deblur

Xinge Yang, Chuong Nguyen, Wenbin Wang, Kaizhang Kang, Wolfgang Heidrich, Xiaoxing Li

机构 * KAUST(卡斯土尼亚大学) Meta Reality Labs(Meta现实实验室)

Journal ref Proceedings of the IEEE/CVF International Conference on Computer Vision (ICCV) Workshops 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2503.16421 2025-10-28 cs.CV cs.AI cs.LG cs.MM

MagicMotion: Controllable Video Generation with Dense-to-Sparse Trajectory Guidance

Quanhao Li, Zhen Xing, Rui Wang, Hui Zhang, Qi Dai, Zuxuan Wu

机构 * Shanghai Key Lab of Intell. Info. Processing, School of CS, Fudan University(上海智能信息处理关键实验室,复旦大学计算机学院) Shanghai Collaborative Innovation Center of Intelligent Visual Computing(上海智能视觉计算协同创新中心) Microsoft Research Asia(微软亚洲研究院)

Comments Accepted by ICCV 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2503.14350 2025-10-28 cs.CV cs.AI cs.CL

VEGGIE: Instructional Editing and Reasoning Video Concepts with Grounded Generation

Shoubin Yu, Difan Liu, Ziqiao Ma, Yicong Hong, Yang Zhou, Hao Tan, Joyce Chai, Mohit Bansal

机构 * Adobe Research(Adobe研究机构) University of Michigan(密歇根大学) UNC Chapel Hill(北卡罗来纳大学教堂山分校)

Comments ICCV 2025; First three authors contributed equally. Project page: https://veggie-gen.github.io/

详情

展开后加载摘要…

URL PDF HTML 收藏
2502.09110 2025-10-28 cs.CV

Pulling Back the Curtain: Unsupervised Adversarial Detection via Contrastive Auxiliary Networks

Eylon Mizrahi, Raz Lapid, Moshe Sipper

机构 * Ben-Gurion University(本·古里安大学) DeepKeep

Comments Accepted for Oral Presentation at SafeMM-AI @ ICCV 2025 (Spotlight)

Journal ref ICCV 2025 Workshop on SafeMM-AI: Safe and Trustworthy Multimodal AI Systems

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.21840 2025-10-28 cs.CV cs.GR

Improving the Physics of Video Generation with VJEPA-2 Reward Signal

Jianhao Yuan, Xiaofeng Zhang, Felix Friedrich, Nicolas Beltran-Velez, Melissa Hall, Reyhane Askari-Hemmat, Xiaochuang Han, Nicolas Ballas, Michal Drozdzal, Adriana Romero-Soriano

机构 * FAIR, Meta Superintelligence Labs(FAIR、Meta超智能实验室) University of Oxford(牛津大学) Mila - Québec AI Institute(魁北克人工智能研究所) Université de Montréal(蒙特利尔大学) Columbia University(哥伦比亚大学) McGill University(麦吉尔大学) Canada CIFAR AI Chair(加拿大CIFAR人工智能 chair)

Comments 2 pages

Journal ref Winning entry of the ICCV 2025 Physics IQ Challenge

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.21809 2025-10-28 cs.CV cs.RO

Embodied Navigation with Auxiliary Task of Action Description Prediction

Haru Kondoh, Asako Kanezaki

机构 * Institute of Science Tokyo(东京科学研究所) RIKEN AIP(日本科学技术研究所AIP)

Comments ICCV 2025 Poster

详情

展开后加载摘要…

URL PDF HTML 收藏