arXivDaily arXiv每日学术速递 周一至周五更新

期刊&会议

International Conference on Computer Vision · 会议 · Computer Vision

共收录 4770
2508.09372 2025-08-14 cs.CV cs.AI cs.IR cs.LG

A Signer-Invariant Conformer and Multi-Scale Fusion Transformer for Continuous Sign Language Recognition

Md Rezwanul Haque, Md. Milon Islam, S M Taslim Uddin Raju, Fakhri Karray

机构 * University of Waterloo(滑铁卢大学) Mohamed bin Zayed University of Artificial Intelligence(穆罕默德·本·扎耶德人工智能大学)

Comments Accepted for the IEEE/CVF International Conference on Computer Vision (ICCV), Honolulu, Hawaii, USA. 1st MSLR Workshop 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.09362 2025-08-14 cs.CV cs.AI cs.LG

FusionEnsemble-Net: An Attention-Based Ensemble of Spatiotemporal Networks for Multimodal Sign Language Recognition

Md. Milon Islam, Md Rezwanul Haque, S M Taslim Uddin Raju, Fakhri Karray

机构 * University of Waterloo(滑铁卢大学) Mohamed bin Zayed University of Artificial Intelligence(穆罕默德·本·扎耶德人工智能大学)

Comments Accepted for the IEEE/CVF International Conference on Computer Vision (ICCV), Honolulu, Hawaii, USA. 1st MSLR Workshop 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.09262 2025-08-14 cs.CV cs.LG

Harnessing Input-Adaptive Inference for Efficient VLN

Dongwoo Kang, Akhil Perincherry, Zachary Coalson, Aiden Gabriel, Stefan Lee, Sanghyun Hong

机构 * Oregon State University(俄勒冈州立大学)

Comments Accepted to ICCV 2025 [Poster]

详情

展开后加载摘要…

URL PDF HTML 收藏
2507.01367 2025-08-14 cs.CV

3D Gaussian Splatting Driven Multi-View Robust Physical Adversarial Camouflage Generation

Tianrui Lou, Xiaojun Jia, Siyuan Liang, Jiawei Liang, Ming Zhang, Yanjun Xiao, Xiaochun Cao

机构 * Sun Yat-Sen University(中山大学) Peng Cheng Laboratory(鹏城实验室) Nanyang Technological University(南洋理工大学) National University of Singapore(新加坡国立大学) National Key Laboratory of Science and Technology on Information System Security(国家信息系统安全科学技术重点实验室) Nsfocus

Comments Accepted by ICCV 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.09137 2025-08-13 cs.CV

HumanOLAT: A Large-Scale Dataset for Full-Body Human Relighting and Novel-View Synthesis

Timo Teufel, Pulkit Gera, Xilong Zhou, Umar Iqbal, Pramod Rao, Jan Kautz, Vladislav Golyanik, Christian Theobalt

机构 * Max Planck Institute for Informatics(马克斯·普朗克信息研究所) NVIDIA(英伟达)

Comments TT and PG contributed equally; accepted at ICCV 2025; project page: https://vcai.mpi-inf.mpg.de/projects/HumanOLAT/

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.09062 2025-08-13 cs.GR cs.CV cs.LG

VertexRegen: Mesh Generation with Continuous Level of Detail

Xiang Zhang, Yawar Siddiqui, Armen Avetisyan, Chris Xie, Jakob Engel, Henry Howard-Jenkins

机构 * UC San Diego(圣迭戈大学) Meta Reality Labs Research(Meta现实实验室)

Comments ICCV 2025. Project Page: https://vertexregen.github.io/

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.09000 2025-08-13 cs.CV

UniConvNet: Expanding Effective Receptive Field while Maintaining Asymptotically Gaussian Distribution for ConvNets of Any Scale

Yuhao Wang, Wei Xi

机构 * Xi’an Jiaotong University(西安交通大学)

Comments ICCV 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.08949 2025-08-13 cs.CV

Lay2Story: Extending Diffusion Transformers for Layout-Togglable Story Generation

Ao Ma, Jiasong Feng, Ke Cao, Jing Wang, Yun Wang, Quanwei Zhang, Zhanjie Zhang

机构 * JD.com, Inc.(京东公司)

Comments Accepted by ICCV 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.08891 2025-08-13 cs.CV

Preview WB-DH: Towards Whole Body Digital Human Bench for the Generation of Whole-body Talking Avatar Videos

Chaoyi Wang, Yifan Yang, Jun Pei, Lijie Xia, Jianpo Liu, Xiaobing Yuan, Xinhan Di

Comments This paper has been accepted by ICCV 2025 Workshop MMFM4

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.08867 2025-08-13 cs.CV

GaussianUpdate: Continual 3D Gaussian Splatting Update for Changing Environments

Lin Zeng, Boming Zhao, Jiarui Hu, Xujie Shen, Ziqiang Dang, Hujun Bao, Zhaopeng Cui

机构 * State Key Lab of CAD & CG(计算机辅助设计与图形学国家重点实验室)

Comments Accepted to ICCV 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.08811 2025-08-13 cs.CV

Revisiting Efficient Semantic Segmentation: Learning Offsets for Better Spatial and Class Feature Alignment

Shi-Chen Zhang, Yunheng Li, Yu-Huan Wu, Qibin Hou, Ming-Ming Cheng

机构 * VCIP, CS, Nankai University(VCIP、计算机科学系、南开大学)

Comments Accepted at ICCV 2025. Project page: https://github.com/HVision-NKU/OffSeg

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.08589 2025-08-13 cs.CV

DocThinker: Explainable Multimodal Large Language Models with Rule-based Reinforcement Learning for Document Understanding

Wenwen Yu, Zhibo Yang, Yuliang Liu, Xiang Bai

机构 * Huazhong University of Science and Technology(华中科技大学) Alibaba Group(阿里巴巴集团)

Comments ICCV 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.08498 2025-08-13 cs.CV

CObL: Toward Zero-Shot Ordinal Layering without User Prompting

Aneel Damaraju, Dean Hazineh, Todd Zickler

机构 * Harvard University, School of Engineering and Applied Sciences(哈佛大学工程与应用科学学院)

Comments ICCV 2025: Project page with demo, datasets, and code: https://vision.seas.harvard.edu/cobl/

详情

展开后加载摘要…

URL PDF HTML 收藏
2507.01384 2025-08-13 cs.CV

MUG: Pseudo Labeling Augmented Audio-Visual Mamba Network for Audio-Visual Video Parsing

Langyu Wang, Bingke Zhu, Yingying Chen, Yiyuan Zhang, Ming Tang, Jinqiao Wang

机构 * Foundation Model Research Center, Institute of Automation, Chinese Academy of Sciences, China(中国科学院自动化研究所基础模型研究中心)

Comments Accpted by ICCV 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2503.21055 2025-08-13 cs.CV

What Changed and What Could Have Changed? State-Change Counterfactuals for Procedure-Aware Video Representation Learning

Chi-Hsi Kung, Frangil Ramirez, Juhyung Ha, Yi-Ting Chen, David Crandall, Yi-Hsuan Tsai

机构 * Indiana University(印第安纳大学) National Yang-Ming Chiao-Tung University(国家阳明交通大学) Atmanity Inc.(Atmanity公司)

Comments 16 pages, 4 figures

Journal ref International Conference on Computer Vision, ICCV 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2412.20977 2025-08-13 cs.AI cs.CV cs.RO

UnrealZoo: Enriching Photo-realistic Virtual Worlds for Embodied AI

Fangwei Zhong, Kui Wu, Churan Wang, Hao Chen, Hai Ci, Zhoujun Li, Yizhou Wang

机构 * School of Artificial Intelligence, Beijing Normal University(北京师范大学人工智能学院) State Key Laboratory of Complex & Critical Software Environment, Beihang University(北航复杂与关键软件环境国家重点实验室) School of Computer Science, Institute for Artificial Intelligence, State Key Laboratory of General Artificial Intelligence, Peking University(北京大学计算机学院) City University of Macau(澳门城市大学) National University of Singapore(新加坡国立大学) Beijing Institute for General Artificial Intelligence (BIGAI)(北京通用人工智能研究院)

Comments ICCV 2025 (Highlight), Project page: http://unrealzoo.site/

详情

展开后加载摘要…

URL PDF HTML 收藏
2412.02141 2025-08-13 cs.CV cs.CL

WSI-LLaVA: A Multimodal Large Language Model for Whole Slide Image

Yuci Liang, Xinheng Lyu, Wenting Chen, Meidan Ding, Jipeng Zhang, Xiangjian He, Song Wu, Xiaohan Xing, Sen Yang, Xiyue Wang, Linlin Shen

机构 * Shenzhen University(深圳大学) University of Nottingham Ningbo China(诺丁汉大学宁波分校) City University of Hong Kong(香港城市大学) Stanford University(斯坦福大学) Hong Kong University of Science and Technology(香港科学与技术大学)

Comments ICCV 2025, 38 pages, 22 figures, 35 tables

详情

展开后加载摘要…

URL PDF HTML 收藏
2410.09865 2025-08-13 cs.CV

SynFER: Towards Boosting Facial Expression Recognition with Synthetic Data

Xilin He, Cheng Luo, Xiaole Xian, Bing Li, Muhammad Haris Khan, Zongyuan Ge, Weicheng Xie, Siyang Song, Linlin Shen, Bernard Ghanem, Xiangyu Yue

机构 * School of Computer Science & Software Engineering, Shenzhen University, China(深圳大学计算机科学与软件工程学院) Guangdong Laboratory of Artificial Intelligence and Digital Economy (SZ), Shenzhen, China(广东省人工智能与数字经济实验室) Guangdong Provincial Key Laboratory of Intelligent Information Processing, Shenzhen University, China(广东省智能信息处理重点实验室) KAUST(卡塔尔科技大学) Monash University(墨尔本大学) School of Computer Science, University of Exeter, UK(埃克塞特大学计算机科学学院) Computer Vision Institute, School of Artificial Intelligence, Shenzhen University, China(深圳大学人工智能学院计算机视觉研究所) MBZUAI(穆罕默德·本·拉什德智能研究院) The Chinese University of Hong Kong(香港中文大学)

Comments ICCV 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.08254 2025-08-12 cs.CV

Learning an Implicit Physics Model for Image-based Fluid Simulation

Emily Yue-Ting Jia, Jiageng Mao, Zhiyuan Gao, Yajie Zhao, Yue Wang

机构 * University of Southern California(南加州大学)

Comments Accepted at ICCV 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.08165 2025-08-12 cs.CV cs.LG

Integrating Task-Specific and Universal Adapters for Pre-Trained Model-based Class-Incremental Learning

Yan Wang, Da-Wei Zhou, Han-Jia Ye

机构 * School of Artificial Intelligence, Nanjing University(人工智能学院,南京大学) National Key Laboratory for Novel Software Technology, Nanjing University(新型软件技术国家重点实验室,南京大学)

Comments Accepted to ICCV 2025. Code is available at: https://github.com/LAMDA-CL/ICCV2025-TUNA

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.07989 2025-08-12 cs.CV cs.HC

The Escalator Problem: Identifying Implicit Motion Blindness in AI for Accessibility

Xiantao Zhang

机构 * Beihang University(北航大学)

Comments 9 pages, 3 figures, 2 tables. Accepted at CV4A11y, ICCV 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.07877 2025-08-12 cs.CV cs.AI

Selective Contrastive Learning for Weakly Supervised Affordance Grounding

WonJun Moon, Hyun Seok Seong, Jae-Pil Heo

机构 * Sungkyunkwan University(全北大学)

Comments Accepted to ICCV 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.07850 2025-08-12 cs.CV

Morphological Analysis of Semiconductor Microstructures using Skeleton Graphs

Noriko Nitta, Rei Miyata, Naoto Oishi

机构 * Kochi University of Technology(Kochi技术大学) National Institute of Technology, Kochi College(Kochi国立技术大学学院)

Comments CV4MS: Computer Vision for Materials Science, Workshop in conjunction with the IEEE/CVF ICCV 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.07847 2025-08-12 cs.CV cs.AI

Deep Space Weather Model: Long-Range Solar Flare Prediction from Multi-Wavelength Images

Shunya Nagashima, Komei Sugiura

机构 * Keio University(Keio大学)

Comments ICCV 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.07775 2025-08-12 cs.CV

Forecasting Continuous Non-Conservative Dynamical Systems in SO(3)

Lennart Bastian, Mohammad Rashed, Nassir Navab, Tolga Birdal

机构 * Technical University of Munich(慕尼黑技术大学) Munich Center of Machine Learning(慕尼黑机器学习中心) Imperial College London(伦敦帝国学院)

Comments ICCV 2025 Oral

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.07747 2025-08-12 cs.CV

Grouped Speculative Decoding for Autoregressive Image Generation

Junhyuk So, Juncheol Shin, Hyunho Kook, Eunhyeok Park

机构 * Department of Computer Science and Engineering, POSTECH(计算机科学与工程系,POSTECH) Graduate School of Artificial Intelligence, POSTECH(人工智能研究生院,POSTECH)

Comments Accepted to the ICCV 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.07647 2025-08-12 cs.CV

LaRender: Training-Free Occlusion Control in Image Generation via Latent Rendering

Xiaohang Zhan, Dingming Liu

机构 * Tencent(腾讯)

Comments Accepted by ICCV 2025 (oral). Project page: https://xiaohangzhan.github.io/projects/larender/

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.07519 2025-08-12 cs.CV

Exploring Multimodal Diffusion Transformers for Enhanced Prompt-based Image Editing

Joonghyuk Shin, Alchan Hwang, Yujin Kim, Daneul Kim, Jaesik Park

机构 * Seoul National University(首尔国立大学)

Comments ICCV 2025. Project webpage: https://joonghyuk.com/exploring-mmdit-web/

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.02516 2025-08-12 cs.CV

Engagement Prediction of Short Videos with Large Multimodal Models

Wei Sun, Linhan Cao, Yuqin Cao, Weixia Zhang, Wen Wen, Kaiwei Zhang, Zijian Chen, Fangfang Lu, Xiongkuo Min, Guangtao Zhai

机构 * East China Normal University(华东师范大学) Shanghai Jiao Tong University(上海交通大学) City University of Hong Kong(香港城市大学) Shanghai University of Electric Power(上海电力大学)

Comments The proposed method achieves first place in the ICCV VQualA 2025 EVQA-SnapUGC Challenge on short-form video engagement prediction

详情

展开后加载摘要…

URL PDF HTML 收藏
2507.19993 2025-08-12 cs.CV

FROSS: Faster-than-Real-Time Online 3D Semantic Scene Graph Generation from RGB-D Images

Hao-Yu Hou, Chun-Yi Lee, Motoharu Sonogashira, Yasutomo Kawanishi

机构 * National Tsing Hua University(国立清华大学) National Taiwan University(国立台湾大学) RIKEN(日本研究机构)

Comments International Conference on Computer Vision (ICCV 2025)

详情

展开后加载摘要…

URL PDF HTML 收藏