arXivDaily arXiv每日学术速递 周一至周五更新

期刊&会议

International Conference on Computer Vision · 会议 · Computer Vision

2025-10-21 至 2025-10-21 共收录 18
2510.17384 2025-10-21 cs.CV

Closed-Loop Transfer for Weakly-supervised Affordance Grounding

Jiajin Tang, Zhengxuan Wei, Ge Zheng, Sibei Yang

机构 * ShanghaiTech University(上海科技大学) School of Computer Science and Engineering(计算机科学与工程学院)

Comments Accepted at ICCV 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.17332 2025-10-21 cs.CV

iDETEX: Empowering MLLMs for Intelligent DETailed EXplainable IQA

Zhaoran Zhao, Xinli Yue, Jianhui Sun, Yuhao Xie, Tao Shao, Liangchao Yao, Fan Xia, Yuetang Deng

机构 * Tencent, WeChat(腾讯,微信)

Comments Accepted to ICCV 2025 Workshop

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.17181 2025-10-21 cs.CV

Capturing Head Avatar with Hand Contacts from a Monocular Video

Haonan He, Yufeng Zheng, Jie Song

机构 * The Hong Kong University of Science and Technology (Guangzhou)(香港科技大学(广州)) The Hong Kong University of Science and Technology(香港科技大学) ETH Zürich(苏黎世联邦理工学院) Max Planck Institute for Intelligent Systems(智能系统研究所)

Comments ICCV 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.17078 2025-10-21 cs.CV

Towards a Generalizable Fusion Architecture for Multimodal Object Detection

Jad Berjawi, Yoann Dupas, Christophe C'erin

机构 * Université Grenoble Alpes(格勒诺布尔大学) Université Sorbonne Paris Nord(巴黎-萨克勒大学) INRIA(法国国家信息与自动化研究所)

Comments 8 pages, 8 figures, accepted at ICCV 2025 MIRA Workshop

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.17023 2025-10-21 cs.CV cs.MM

Enrich and Detect: Video Temporal Grounding with Multimodal LLMs

Shraman Pramanick, Effrosyni Mavroudi, Yale Song, Rama Chellappa, Lorenzo Torresani, Triantafyllos Afouras

机构 * FAIR, Meta(FAIR、Meta) Johns Hopkins University(约翰霍普金斯大学) Northeastern University(东北大学)

Comments ICCV 2025 (Highlights)

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.16774 2025-10-21 cs.LG cs.AI

Learning to play: A Multimodal Agent for 3D Game-Play

Yuguang Yue, Irakli Salia, Samuel Hunt, Christopher Green, Wenzhe Shi, Jonathan J Hunt

机构 * Player2

Comments International Conference on Computer Vision Workshop on Multi-Modal Reasoning for Agentic Intelligence

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.15208 2025-10-21 cs.CV

CARDIUM: Congenital Anomaly Recognition with Diagnostic Images and Unified Medical records

Daniela Vega, Hannah V. Ceballos, Javier S. Vera, Santiago Rodriguez, Alejandra Perez, Angela Castillo, Maria Escobar, Dario Londoño, Luis A. Sarmiento, Camila I. Castro, Nadiezhda Rodriguez, Juan C. Briceño, Pablo Arbeláez

机构 * Universidad de los Andes(andes大学) Fundación Santa Fe de Bogotá(圣费博多哥基金会)

Comments Accepted to CVAMD Workshop, ICCV 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.16972 2025-10-21 cs.CV cs.AI

The 1st Solution for 7th LSVOS RVOS Track: SaSaSa2VA

Quanzhu Niu, Dengxian Gong, Shihao Chen, Tao Zhang, Yikang Zhou, Haobo Yuan, Lu Qi, Xiangtai Li, Shunping Ji

机构 * Wuhan University(武汉大学) University of California, Merced(加州大学默塞德分校) Nanyang Technological University(南洋理工大学)

Comments The 1st place report of 7th LSVOS challenge RVOS track in ICCV 2025. The code is released in Sa2VA repository: https://github.com/bytedance/Sa2VA

详情

展开后加载摘要…

URL PDF HTML 收藏
2507.18060 2025-10-21 cs.CV

BokehDiff: Neural Lens Blur with One-Step Diffusion

Chengxuan Zhu, Qingnan Fan, Qi Zhang, Jinwei Chen, Huaqi Zhang, Chao Xu, Boxin Shi

机构 * National Key Lab of General AI, School of Intelligence Science and Technology, Peking University(国家关键人工智能实验室,智能科学与技术学院,北京大学) Vivo Mobile Communication Co., Ltd.(Vivo移动通信有限公司) State Key Laboratory of Multimedia Information Processing, School of Computer Science, Peking University(多媒体信息处理国家重点实验室,计算机科学学院,北京大学) National Engineering Research Center of Visual Technology, School of Computer Science, Peking University(视觉技术国家工程研究中心,计算机科学学院,北京大学)

Comments Accepted by ICCV 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2505.22911 2025-10-21 cs.CV

Hierarchical Material Recognition from Local Appearance

Matthew Beveridge, Shree K. Nayar

Comments ICCV 2025 Camera Ready

详情

展开后加载摘要…

URL PDF HTML 收藏
2504.11521 2025-10-21 cs.LG cs.RO

LANGTRAJ: Diffusion Model and Dataset for Language-Conditioned Trajectory Simulation

Wei-Jer Chang, Wei Zhan, Masayoshi Tomizuka, Manmohan Chandraker, Francesco Pittaluga

机构 * UC Berkeley(加州大学伯克利分校) NEC Labs America(NEC美国实验室) UC San Diego(加州大学圣地亚哥分校)

Comments ICCV 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.16457 2025-10-21 cs.CV cs.RO

NavQ: Learning a Q-Model for Foresighted Vision-and-Language Navigation

Peiran Xu, Xicheng Gong, Yadong MU

机构 * Peking University(北京大学)

Comments ICCV 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.16391 2025-10-21 physics.optics

Recover Biological Structure from Sparse-View Diffraction Images with Neural Volumetric Prior

Renzhi He, Haowen Zhou, Yubei Chen, Yi Xue

Comments ICCV 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.16377 2025-10-21 cs.CV

Demeter: A Parametric Model of Crop Plant Morphology from the Real World

Tianhang Cheng, Albert J. Zhai, Evan Z. Chen, Rui Zhou, Yawen Deng, Zitong Li, Kejie Zhao, Janice Shiu, Qianyu Zhao, Yide Xu, Xinlei Wang, Yuan Shen, Sheng Wang, Lisa Ainsworth, Kaiyu Guan, Shenlong Wang

机构 * University of Illinois Urbana-Champaign(伊利诺伊大学厄巴纳-香槟分校)

Comments ICCV 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.16319 2025-10-21 cs.CV

Stroke2Sketch: Harnessing Stroke Attributes for Training-Free Sketch Generation

Rui Yang, Huining Li, Yiyi Long, Xiaojun Wu, Shengfeng He

机构 * Huaqiao University(华侨大学) South China University of Technology(华南理工大学) Shaanxi Normal University(陕西师范大学) Beijing University of Aeronautics and Astronautics(北京航空航天大学) Singapore Management University(新加坡国立大学)

Comments ICCV 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.16118 2025-10-21 cs.CV

ObjectTransforms for Uncertainty Quantification and Reduction in Vision-Based Perception for Autonomous Vehicles

Nishad Sahu, Shounak Sural, Aditya Satish Patil, Ragunathan, Rajkumar

机构 * Carnegie Mellon University(卡内基梅隆大学) University Of Minnesota(明尼苏达大学)

Comments Accepted at International Conference on Computer Vision (ICCV) 2025 Workshops

详情

展开后加载摘要…

URL PDF HTML 收藏
2506.21514 2025-10-21 cs.CV

G$^{2}$D: Boosting Multimodal Learning with Gradient-Guided Distillation

Mohammed Rakib, Arunkumar Bagavathi

机构 * Oklahoma State University(俄克拉荷马州立大学)

Comments Accepted at ICCV 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2411.17041 2025-10-21 cs.CV cs.AI cs.LG

Free$^2$Guide: Training-Free Text-to-Video Alignment using Image LVLM

Jaemin Kim, Bryan Sangwoo Kim, Jong Chul Ye

机构 * Graduate School of AI, KAIST(人工智能研究生院,韩国科学技术院)

Comments ICCV 2025 accepted

详情

展开后加载摘要…

URL PDF HTML 收藏