arXivDaily arXiv每日学术速递 周一至周五更新

期刊&会议

International Conference on Computer Vision · 会议 · Computer Vision

2025-10-14 至 2025-10-14 共收录 18
2510.11605 2025-10-14 cs.CV

ACE-G: Improving Generalization of Scene Coordinate Regression Through Query Pre-Training

Leonard Bruns, Axel Barroso-Laguna, Tommaso Cavallari, Áron Monszpart, Sowmya Munukutla, Victor Adrian Prisacariu, Eric Brachmann

Comments ICCV 2025, Project page: https://nianticspatial.github.io/ace-g/

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.11437 2025-10-14 eess.IV

GADA: Graph Attention-based Detection Aggregation for Ultrasound Video Classification

Li Chen, Naveen Balaraju, Jochen Kruecker, Balasundar Raju, Alvin Chen

Comments ICCV CVAMD 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.11346 2025-10-14 cs.CV cs.AI

Uncertainty-Aware ControlNet: Bridging Domain Gaps with Synthetic Image Generation

Joshua Niemeijer, Jan Ehrhardt, Heinz Handels, Hristina Uzunova

机构 * German Aerospace Center (DLR)(德国航空航天中心) University of Lübeck(吕贝克大学) German Research Center for Artificial Intelligence (DFKI)(德国人工智能研究中心)

Comments Accepted for presentation at ICCV Workshops 2025, "The 4th Workshop on What is Next in Multimodal Foundation Models?" (MMFM)

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.11107 2025-10-14 cs.CV

MoMaps: Semantics-Aware Scene Motion Generation with Motion Maps

Jiahui Lei, Kyle Genova, George Kopanas, Noah Snavely, Leonidas Guibas

机构 * Google DeepMind(谷歌DeepMind) University of Pennsylvania(宾夕法尼亚大学) Google(谷歌)

Comments Accepted at ICCV 2025, project page: https://jiahuilei.com/projects/momap/

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.11017 2025-10-14 cs.CV

High-Resolution Spatiotemporal Modeling with Global-Local State Space Models for Video-Based Human Pose Estimation

Runyang Feng, Hyung Jin Chang, Tze Ho Elden Tse, Boeun Kim, Yi Chang, Yixing Gao

机构 * School of Artificial Intelligence, Jilin University(吉林大学人工智能学院) Engineering Research Center of Knowledge-Driven Human-Machine Intelligence, Ministry of Education, China(知识驱动人机智能工程研究中心,中华人民共和国教育部,中国) School of Computer Science, University of Birmingham(伯明翰大学计算机科学学院) National University of Singapore(新加坡国立大学) Dankook University(Dankook 大学)

Comments This paper is accepted to ICCV 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.10793 2025-10-14 cs.CV

ImHead: A Large-scale Implicit Morphable Model for Localized Head Modeling

Rolandos Alexandros Potamias, Stathis Galanakis, Jiankang Deng, Athanasios Papaioannou, Stefanos Zafeiriou

机构 * Imperial College London(伦敦帝国学院)

Comments ICCV 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.01852 2025-10-14 cs.CV cs.MM

Context Guided Transformer Entropy Modeling for Video Compression

Junlong Tong, Wei Zhang, Yaohui Jin, Xiaoyu Shen

机构 * Shanghai Jiao Tong University(上海交通大学) Ningbo Key Laboratory of Spatial Intelligence and Digital Derivative(宁波空间智能与数字衍生关键实验室) Institute of Digital Twin(数字孪生研究院)

Comments ICCV 2025. This is an update to the camera-ready version

详情

展开后加载摘要…

URL PDF HTML 收藏
2507.09279 2025-10-14 cs.CV cs.AI cs.CL

Prompt4Trust: A Reinforcement Learning Prompt Augmentation Framework for Clinically-Aligned Confidence Calibration in Multimodal Large Language Models

Anita Kriz, Elizabeth Laura Janes, Xing Shen, Tal Arbel

机构 * McGill University(麦吉尔大学) Mila – Quebec AI Institute(魁北克AI研究所)

Comments Accepted to ICCV 2025 Workshop CVAMD

详情

展开后加载摘要…

URL PDF HTML 收藏
2507.01738 2025-10-14 cs.CV

DeRIS: Decoupling Perception and Cognition for Enhanced Referring Image Segmentation through Loopback Synergy

Ming Dai, Wenxuan Cheng, Jiang-jiang Liu, Sen Yang, Wenxiao Cai, Yanpeng Sun, Wankou Yang

机构 * Southeast University(东南大学) Baidu VIS(百度VIS) Stanford University(斯坦福大学)

Comments Accepted by ICCV 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2504.09426 2025-10-14 cs.CV cs.AI cs.CL

BabyVLM: Data-Efficient Pretraining of VLMs Inspired by Infant Learning

Shengao Wang, Arjun Chandra, Aoming Liu, Venkatesh Saligrama, Boqing Gong

机构 * Boston University(波士顿大学)

Comments Accepted to ICCV 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2502.01720 2025-10-14 cs.CV cs.GR cs.LG

Generating Multi-Image Synthetic Data for Text-to-Image Customization

Nupur Kumari, Xi Yin, Jun-Yan Zhu, Ishan Misra, Samaneh Azadi

机构 * Carnegie Mellon University(卡内基梅隆大学) Meta

Comments ICCV 2025. Project webpage: https://www.cs.cmu.edu/~syncd-project/

详情

展开后加载摘要…

URL PDF HTML 收藏
2411.14961 2025-10-14 cs.LG cs.CV

LoRA-FAIR: Federated LoRA Fine-Tuning with Aggregation and Initialization Refinement

Jieming Bian, Lei Wang, Letian Zhang, Jie Xu

机构 * University of Florida(佛罗里达大学) Middle Tennessee State University(中田纳西州立大学)

Comments ICCV 2025, Code is available: https://github.com/jmbian/LoRA-FAIR

详情

展开后加载摘要…

URL PDF HTML 收藏
2406.04765 2025-10-14 cs.CV cs.MM

SMC++: Masked Learning of Unsupervised Video Semantic Compression

Yuan Tian, Xiaoyue Ling, Cong Geng, Qiang Hu, Guo Lu, Guangtao Zhai

机构 * Shanghai Artificial Intelligence Laboratory(上海人工智能实验室) JIUTIAN Research(JIUTIAN研究) Cooperative Medianet Innovation Center(协作中继创新中心) Shanghai Jiao Tong University(上海交通大学) Institute of Image Communication and Network Engineering(图像通信与网络工程研究所)

Comments Accepted to TPAMI; Substantial Extension of ICCV 2023 paper

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.10191 2025-10-14 cs.CV

Fairness Without Labels: Pseudo-Balancing for Bias Mitigation in Face Gender Classification

Haohua Dong, Ana Manzano Rodríguez, Camille Guinaudeau, Shin'ichi Satoh

机构 * National Institute of Informatics, Japan(日本信息机构国家研究所) INESC-ID, Instituto Superior Técnico, University of Lisbon, Portugal(里斯本大学技术高等学院、INESC-ID,葡萄牙) University of Amsterdam, The Netherlands(荷兰阿姆斯特丹大学) LIMSI, CNRS / Université Paris-Saclay, France(法国CNRS/巴黎萨克雷大学LIMSI)

Comments 8 pages. Accepted for publication in the ICCV 2025 Workshop Proceedings (2nd FAILED Workshop). Also available on HAL (hal-05210445v1)

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.10100 2025-10-14 cs.CV cs.LG

Cooperative Pseudo Labeling for Unsupervised Federated Classification

Kuangpu Guo, Lijun Sheng, Yongcan Yu, Jian Liang, Zilei Wang, Ran He

机构 * University of Science and Technology of China(中国科学技术大学) NLPR & MAIS, Institute of Automation, Chinese Academy of Sciences(中国科学院自动化研究所) School of Artificial Intelligence, University of Chinese Academy of Sciences(中国科学院大学人工智能学院)

Comments Accepted by ICCV 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2507.13229 2025-10-14 cs.CV cs.AI cs.RO

{S\textsuperscript{2}M\textsuperscript{2}}: Scalable Stereo Matching Model for Reliable Depth Estimation

Junhong Min, Youngpil Jeon, Jimin Kim, Minyong Choi

机构 * Samsung Electronics(三星电子)

Comments 8 pages, 5 figures, ICCV accepted paper

详情

展开后加载摘要…

URL PDF HTML 收藏
2503.07856 2025-10-14 cs.CV

Blind Video Super-Resolution based on Implicit Kernels

Qiang Zhu, Yuxuan Jiang, Shuyuan Zhu, Fan Zhang, David Bull, Bing Zeng

机构 * University of Electronic Science and Technology of China(电子科技大学) University of Bristol(布里斯托大学)

Comments Accepted by ICCV 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2411.14384 2025-10-14 cs.CV cs.GR

Baking Gaussian Splatting into Diffusion Denoiser for Fast and Scalable Single-stage Image-to-3D Generation and Reconstruction

Yuanhao Cai, He Zhang, Kai Zhang, Yixun Liang, Mengwei Ren, Fujun Luan, Qing Liu, Soo Ye Kim, Jianming Zhang, Zhifei Zhang, Yuqian Zhou, Yulun Zhang, Xiaokang Yang, Zhe Lin, Alan Yuille

机构 * Johns Hopkins University(约翰霍普金斯大学) Adobe Research(Adobe研究) HKUST(香港科技大学) Shanghai Jiao Tong University(上海交通大学)

Comments ICCV 2025; A novel one-stage 3DGS-based diffusion for 3D object generation and scene reconstruction from a single view in ~6 seconds

详情

展开后加载摘要…

URL PDF HTML 收藏