arXivDaily arXiv每日学术速递 周一至周五更新

期刊&会议

International Conference on Computer Vision · 会议 · Computer Vision

共收录 4770
2507.23543 2025-08-11 cs.CV cs.AI

ART: Adaptive Relation Tuning for Generalized Relation Prediction

Gopika Sudhakaran, Hikaru Shindo, Patrick Schramowski, Simone Schaub-Meyer, Kristian Kersting, Stefan Roth

机构 * Department of Computer Science, TU Darmstadt(图宾根大学计算机科学系) Hessian Center for AI (hessian.AI)(海德堡人工智能中心) German Research Center for AI (DFKI)(德国人工智能研究中心)

Comments Accepted for publication in ICCV 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2503.18525 2025-08-11 cs.RO

RoboTron-Nav: A Unified Framework for Embodied Navigation Integrating Perception, Planning, and Prediction

Yufeng Zhong, Chengjian Feng, Feng Yan, Fanfan Liu, Liming Zheng, Lin Ma

机构 * Meituan(美团)

Comments ICCV 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2503.11439 2025-08-11 cs.CV

COIN: Confidence Score-Guided Distillation for Annotation-Free Cell Segmentation

Sanghyun Jo, Seo Jin Lee, Seungwoo Lee, Seohyung Hong, Hyungseok Seo, Kyungsu Kim

机构 * SNU(首尔国立大学)

Comments Accepted at ICCV 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2503.06364 2025-08-11 cs.CV cs.GR

Generative Video Bi-flow

Chen Liu, Tobias Ritschel

机构 * University College London(伦敦大学学院)

Comments ICCV 2025. Project Page at https://ryushinn.github.io/ode-video

详情

展开后加载摘要…

URL PDF HTML 收藏
2501.13667 2025-08-11 cs.CV

MPG-SAM 2: Adapting SAM 2 with Mask Priors and Global Context for Referring Video Object Segmentation

Fu Rong, Meng Lan, Qian Zhang, Lefei Zhang

机构 * National Engineering Research Center for Multimedia Software, School of Computer Science, Wuhan University(国家多媒体软件工程研究中心,计算机学院,武汉大学) Hong Kong University of Science and Technology(香港科学与技术大学) Horizon Robotics

Comments ICCV 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2409.08566 2025-08-11 cs.CV

Hybrid-TTA: Continual Test-time Adaptation via Dynamic Domain Shift Detection

Hyewon Park, Hyejin Park, Jueun Ko, Dongbo Min

机构 * Ewha Womans University(成均馆大学)

Comments Accepted by ICCV 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2405.18018 2025-08-11 cs.CV

A Calibration Tool for Refractive Underwater Vision

Felix Seegräber, Mengkun She, Felix Woelk, Kevin Köser

机构 * Kiel University(基尔大学) University of Applied Sciences Kiel(基尔应用科学大学)

Comments 9 pages, 5 figures, the paper is accepted to the ICCV 2025 Workshop CVAUI-AAMVEM

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.05857 2025-08-11 cs.CV

Multi-view Gaze Target Estimation

Qiaomu Miao, Vivek Raju Golani, Jingyi Xu, Progga Paromita Dutta, Minh Hoai, Dimitris Samaras

机构 * Stony Brook University(石溪大学) The University of Adelaide(阿德莱德大学)

Comments Accepted to ICCV 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.05689 2025-08-11 cs.CV cs.CR cs.LG

Boosting Adversarial Transferability via Residual Perturbation Attack

Jinjia Peng, Zeze Tao, Huibing Wang, Meng Wang, Yang Wang

机构 * School of Cyber Security and Computer(网络安全与计算机学院) College of Information and Science Technology(信息与科学学院) School of Computer and Information Engineering(计算机与信息工程学院)

Comments Accepted to ieee/cvf international conference on computer vision (ICCV2025)

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.04682 2025-08-11 cs.CV

TurboTrain: Towards Efficient and Balanced Multi-Task Learning for Multi-Agent Perception and Prediction

Zewei Zhou, Seth Z. Zhao, Tianhui Cai, Zhiyu Huang, Bolei Zhou, Jiaqi Ma

机构 * University of California, Los Angeles(加州大学洛杉矶分校)

Comments ICCV 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2406.09105 2025-08-11 cs.CV cs.AI cs.CL cs.LG

INS-MMBench: A Comprehensive Benchmark for Evaluating LVLMs' Performance in Insurance

Chenwei Lin, Hanjia Lyu, Xian Xu, Jiebo Luo

机构 * School of Computer Science Fudan University(复旦大学计算机科学学院) Department of Computer Science University of Rochester(罗切斯特大学计算机科学系) School of Economics Fudan University(复旦大学经济学院)

Comments To appear in the International Conference on Computer Vision, ICCV 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.05527 2025-08-08 cs.CV

AI vs. Human Moderators: A Comparative Evaluation of Multimodal LLMs in Content Moderation for Brand Safety

Adi Levi, Or Levi, Sardhendu Mishra, Jonathan Morra

机构 * Zefr Inc(Zefr公司)

Comments Accepted to the Computer Vision in Advertising and Marketing (CVAM) workshop at ICCV 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.05505 2025-08-08 cs.CV

Symmetry Understanding of 3D Shapes via Chirality Disentanglement

Weikang Wang, Tobias Weißberg, Nafie El Amrani, Florian Bernard

机构 * University of Bonn(波恩大学) Lamarr Institute(拉马尔研究所)

Comments Accepted at ICCV 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.05316 2025-08-08 cs.LG cs.CV

Divide-and-Conquer for Enhancing Unlabeled Learning, Stability, and Plasticity in Semi-supervised Continual Learning

Yue Duan, Taicai Chen, Lei Qi, Yinghuan Shi

机构 * Nanjing University(南京大学) Southeast University(东南大学)

Comments Accepted by ICCV 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.02903 2025-08-08 cs.CV eess.IV stat.ML

RDDPM: Robust Denoising Diffusion Probabilistic Model for Unsupervised Anomaly Segmentation

Mehrdad Moradi, Kamran Paynabar

机构 * H. Milton Stewart School of Industrial and Systems Engineering, Georgia Institute of Technology(H. Milton Stewart工业与系统工程学院,佐治亚理工学院)

Comments 10 pages, 5 figures. Accepted to the ICCV 2025 Workshop on Vision-based Industrial InspectiON (VISION)

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.02095 2025-08-08 cs.CV cs.AI

VLM4D: Towards Spatiotemporal Awareness in Vision Language Models

Shijie Zhou, Alexander Vilesov, Xuehai He, Ziyu Wan, Shuwang Zhang, Aditya Nagachandra, Di Chang, Dongdong Chen, Xin Eric Wang, Achuta Kadambi

机构 * UCLA(美国大学洛杉矶分校) Microsoft(微软公司) UCSC(加州大学圣塔克拉拉分校) USC(美国大学洛杉矶分校)

Comments ICCV 2025, Project Website: https://vlm4d.github.io/

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.00649 2025-08-08 cs.CV cs.CR

Revisiting Adversarial Patch Defenses on Object Detectors: Unified Evaluation, Large-Scale Dataset, and New Insights

Junhao Zheng, Jiahao Sun, Chenhao Lin, Zhengyu Zhao, Chen Ma, Chong Zhang, Cong Wang, Qian Wang, Chao Shen

机构 * Xi’an Jiaotong University(西安交通大学) City University of Hong Kong(香港城市大学) Wuhan University(武汉大学)

Comments Accepted by ICCV 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2507.21745 2025-08-08 cs.CV

Few-Shot Vision-Language Reasoning for Satellite Imagery via Verifiable Rewards

Aybora Koksal, A. Aydin Alatan

Comments ICCV 2025 Workshop on Curated Data for Efficient Learning (CDEL). 10 pages, 3 figures, 6 tables. Our model, training code and dataset will be at https://github.com/aybora/FewShotReasoning

详情

展开后加载摘要…

URL PDF HTML 收藏
2507.12714 2025-08-08 cs.CV cs.GR

NeuraLeaf: Neural Parametric Leaf Models with Shape and Deformation Disentanglement

Yang Yang, Dongni Mao, Hiroaki Santo, Yasuyuki Matsushita, Fumio Okura

机构 * The University of Osaka(大阪大学) Microsoft Research Asia – Tokyo(微软亚洲研究院-东京)

Comments IEEE/CVF International Conference on Computer Vision (ICCV 2025), Highlight, Project: https://neuraleaf-yang.github.io/

详情

展开后加载摘要…

URL PDF HTML 收藏
2507.01603 2025-08-08 cs.CV

DepthSync: Diffusion Guidance-Based Depth Synchronization for Scale- and Geometry-Consistent Video Depth Estimation

Yue-Jiang Dong, Wang Zhao, Jiale Xu, Ying Shan, Song-Hai Zhang

机构 * Tsinghua University(清华大学) ARC Lab, Tencent PCG(腾讯PCG实验室)

Comments Accepted by ICCV 2025; Project Homepage: https://yuejiangdong.github.io/depthsync

详情

展开后加载摘要…

URL PDF HTML 收藏
2503.15877 2025-08-08 cs.CV

Repurposing 2D Diffusion Models with Gaussian Atlas for 3D Generation

Tiange Xiang, Kai Li, Chengjiang Long, Christian Häne, Peihong Guo, Scott Delp, Ehsan Adeli, Li Fei-Fei

机构 * Stanford University(斯坦福大学) Meta Reality Labs(Meta现实实验室)

Comments ICCV 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2503.15451 2025-08-08 cs.CV

MotionStreamer: Streaming Motion Generation via Diffusion-based Autoregressive Model in Causal Latent Space

Lixing Xiao, Shunlin Lu, Huaijin Pi, Ke Fan, Liang Pan, Yueer Zhou, Ziyong Feng, Xiaowei Zhou, Sida Peng, Jingbo Wang

机构 * Zhejiang University(浙江大学) The Chinese University of Hong Kong (Shenzhen)(香港中文大学(深圳)) The University of Hong Kong(香港大学) Shanghai Jiao Tong University(上海交通大学) DeepGlint Shanghai AI Laboratory(上海人工智能实验室)

Comments ICCV 2025. Project Page: https://zju3dv.github.io/MotionStreamer/

详情

展开后加载摘要…

URL PDF HTML 收藏
2412.07689 2025-08-08 cs.CV cs.MM cs.RO

RoboTron-Drive: All-in-One Large Multimodal Model for Autonomous Driving

Zhijian Huang, Chengjian Feng, Feng Yan, Baihui Xiao, Zequn Jie, Yujie Zhong, Xiaodan Liang, Lin Ma

机构 * Shenzhen Campus of Sun Yat-sen University(中山大学深圳校区) Meituan(美团)

Comments ICCV 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2411.18651 2025-08-08 cs.CV cs.CL cs.LG

Verbalized Representation Learning for Interpretable Few-Shot Generalization

Cheng-Fu Yang, Da Yin, Wenbo Hu, Heng Ji, Nanyun Peng, Bolei Zhou, Kai-Wei Chang

机构 * University of California, Los Angeles(加州大学洛杉矶分校) University of Illinois Urbana-Champaign(伊利诺伊大学厄巴纳-香槟分校)

Comments Accepted to ICCV 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.04625 2025-08-07 cs.CV cs.CE

FinMMR: Make Financial Numerical Reasoning More Multimodal, Comprehensive, and Challenging

Zichen Tang, Haihong E, Jiacheng Liu, Zhongjun Yang, Rongjin Li, Zihua Rong, Haoyang He, Zhuodi Hao, Xinyang Hu, Kun Ji, Ziyan Ma, Mengyuan Ji, Jun Zhang, Chenghao Ma, Qianhe Zheng, Yang Liu, Yiling Huang, Xinyi Hu, Qing Huang, Zijian Xie, Shiyao Peng

机构 * Beijing University of Posts and Telecommunications(北京邮电大学)

Comments Accepted by ICCV 2025. arXiv admin note: text overlap with arXiv:2311.06602 by other authors

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.04681 2025-08-07 cs.CV

Perceiving and Acting in First-Person: A Dataset and Benchmark for Egocentric Human-Object-Human Interactions

Liang Xu, Chengqun Yang, Zili Lin, Fei Xu, Yifan Liu, Congsheng Xu, Yiyi Zhang, Jie Qin, Xingdong Sheng, Yunhui Liu, Xin Jin, Yichao Yan, Wenjun Zeng, Xiaokang Yang

机构 * MoE Key Lab of Artificial Intelligence, AI Institute, Shanghai Jiao Tong University(人工智能教育部重点实验室,上海交通大学AI研究院) Ningbo Institute of Digital Twin, Eastern Institute of Technology(宁波数字孪生研究所,东部技术研究所) Ningbo Key Laboratory of Spatial Intelligence and Digital Derivative(宁波空间智能与数字衍生关键实验室) MoE Key Lab of AI, School of Computer Science, Shanghai Jiao Tong University(人工智能教育部重点实验室,上海交通大学计算机学院) Nanjing University of Aeronautics and Astronautics(南京航空航天大学) Lenovo(联想公司)

Comments Accepted to ICCV 2025. Project Page: https://liangxuy.github.io/InterVLA/

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.04659 2025-08-07 cs.CV

PixCuboid: Room Layout Estimation from Multi-view Featuremetric Alignment

Gustav Hanning, Kalle Åström, Viktor Larsson

机构 * Lund University(隆德大学)

Comments Accepted at the ICCV 2025 Workshop on Large Scale Cross Device Localization

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.04648 2025-08-07 astro-ph.IM cs.CV

Super Resolved Imaging with Adaptive Optics

Robin Swanson, Esther Y. H. Lin, Masen Lamb, Suresh Sivanandam, Kiriakos N. Kutulakos

机构 * University of Toronto(多伦多大学) Dunlap Institute for Astronomy & Astrophysics(天文学与天体物理学敦洛普研究所) International Gemini Observatory(国际Gemini天文台) University of Victoria(维多利亚大学)

Comments Accepted to ICCV 2025 (IEEE/CVF International Conference on Computer Vision)

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.04642 2025-08-07 cs.RO cs.CV

RoboTron-Sim: Improving Real-World Driving via Simulated Hard-Case

Baihui Xiao, Chengjian Feng, Zhijian Huang, Feng yan, Yujie Zhong, Lin Ma

机构 * Meituan(美团) Shenzhen Campus of Sun Yat-sen University(中山大学深圳校区)

Comments ICCV 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.04546 2025-08-07 cs.CV

Hierarchical Event Memory for Accurate and Low-latency Online Video Temporal Grounding

Minghang Zheng, Yuxin Peng, Benyuan Sun, Yi Yang, Yang Liu

机构 * Wangxuan Institute of Computer Technology, Peking University(北京大学王宣计算机技术研究所) State Key Laboratory of General Artificial Intelligence, Peking University(北京大学通用人工智能国家重点实验室) Central Media Technology Institute, Huawei(华为中央媒体技术研究所)

Comments Accepted by ICCV 2025

详情

展开后加载摘要…

URL PDF HTML 收藏