arXivDaily arXiv每日学术速递 周一至周五更新

期刊&会议

International Conference on Computer Vision · 会议 · Computer Vision

共收录 4770
2510.15208 2025-10-21 cs.CV

CARDIUM: Congenital Anomaly Recognition with Diagnostic Images and Unified Medical records

Daniela Vega, Hannah V. Ceballos, Javier S. Vera, Santiago Rodriguez, Alejandra Perez, Angela Castillo, Maria Escobar, Dario Londoño, Luis A. Sarmiento, Camila I. Castro, Nadiezhda Rodriguez, Juan C. Briceño, Pablo Arbeláez

机构 * Universidad de los Andes(andes大学) Fundación Santa Fe de Bogotá(圣费博多哥基金会)

Comments Accepted to CVAMD Workshop, ICCV 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.16972 2025-10-21 cs.CV cs.AI

The 1st Solution for 7th LSVOS RVOS Track: SaSaSa2VA

Quanzhu Niu, Dengxian Gong, Shihao Chen, Tao Zhang, Yikang Zhou, Haobo Yuan, Lu Qi, Xiangtai Li, Shunping Ji

机构 * Wuhan University(武汉大学) University of California, Merced(加州大学默塞德分校) Nanyang Technological University(南洋理工大学)

Comments The 1st place report of 7th LSVOS challenge RVOS track in ICCV 2025. The code is released in Sa2VA repository: https://github.com/bytedance/Sa2VA

详情

展开后加载摘要…

URL PDF HTML 收藏
2507.18060 2025-10-21 cs.CV

BokehDiff: Neural Lens Blur with One-Step Diffusion

Chengxuan Zhu, Qingnan Fan, Qi Zhang, Jinwei Chen, Huaqi Zhang, Chao Xu, Boxin Shi

机构 * National Key Lab of General AI, School of Intelligence Science and Technology, Peking University(国家关键人工智能实验室,智能科学与技术学院,北京大学) Vivo Mobile Communication Co., Ltd.(Vivo移动通信有限公司) State Key Laboratory of Multimedia Information Processing, School of Computer Science, Peking University(多媒体信息处理国家重点实验室,计算机科学学院,北京大学) National Engineering Research Center of Visual Technology, School of Computer Science, Peking University(视觉技术国家工程研究中心,计算机科学学院,北京大学)

Comments Accepted by ICCV 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2505.22911 2025-10-21 cs.CV

Hierarchical Material Recognition from Local Appearance

Matthew Beveridge, Shree K. Nayar

Comments ICCV 2025 Camera Ready

详情

展开后加载摘要…

URL PDF HTML 收藏
2504.11521 2025-10-21 cs.LG cs.RO

LANGTRAJ: Diffusion Model and Dataset for Language-Conditioned Trajectory Simulation

Wei-Jer Chang, Wei Zhan, Masayoshi Tomizuka, Manmohan Chandraker, Francesco Pittaluga

机构 * UC Berkeley(加州大学伯克利分校) NEC Labs America(NEC美国实验室) UC San Diego(加州大学圣地亚哥分校)

Comments ICCV 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.16457 2025-10-21 cs.CV cs.RO

NavQ: Learning a Q-Model for Foresighted Vision-and-Language Navigation

Peiran Xu, Xicheng Gong, Yadong MU

机构 * Peking University(北京大学)

Comments ICCV 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.16391 2025-10-21 physics.optics

Recover Biological Structure from Sparse-View Diffraction Images with Neural Volumetric Prior

Renzhi He, Haowen Zhou, Yubei Chen, Yi Xue

Comments ICCV 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.16377 2025-10-21 cs.CV

Demeter: A Parametric Model of Crop Plant Morphology from the Real World

Tianhang Cheng, Albert J. Zhai, Evan Z. Chen, Rui Zhou, Yawen Deng, Zitong Li, Kejie Zhao, Janice Shiu, Qianyu Zhao, Yide Xu, Xinlei Wang, Yuan Shen, Sheng Wang, Lisa Ainsworth, Kaiyu Guan, Shenlong Wang

机构 * University of Illinois Urbana-Champaign(伊利诺伊大学厄巴纳-香槟分校)

Comments ICCV 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.16319 2025-10-21 cs.CV

Stroke2Sketch: Harnessing Stroke Attributes for Training-Free Sketch Generation

Rui Yang, Huining Li, Yiyi Long, Xiaojun Wu, Shengfeng He

机构 * Huaqiao University(华侨大学) South China University of Technology(华南理工大学) Shaanxi Normal University(陕西师范大学) Beijing University of Aeronautics and Astronautics(北京航空航天大学) Singapore Management University(新加坡国立大学)

Comments ICCV 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.16118 2025-10-21 cs.CV

ObjectTransforms for Uncertainty Quantification and Reduction in Vision-Based Perception for Autonomous Vehicles

Nishad Sahu, Shounak Sural, Aditya Satish Patil, Ragunathan, Rajkumar

机构 * Carnegie Mellon University(卡内基梅隆大学) University Of Minnesota(明尼苏达大学)

Comments Accepted at International Conference on Computer Vision (ICCV) 2025 Workshops

详情

展开后加载摘要…

URL PDF HTML 收藏
2506.21514 2025-10-21 cs.CV

G$^{2}$D: Boosting Multimodal Learning with Gradient-Guided Distillation

Mohammed Rakib, Arunkumar Bagavathi

机构 * Oklahoma State University(俄克拉荷马州立大学)

Comments Accepted at ICCV 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2411.17041 2025-10-21 cs.CV cs.AI cs.LG

Free$^2$Guide: Training-Free Text-to-Video Alignment using Image LVLM

Jaemin Kim, Bryan Sangwoo Kim, Jong Chul Ye

机构 * Graduate School of AI, KAIST(人工智能研究生院,韩国科学技术院)

Comments ICCV 2025 accepted

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.15868 2025-10-20 cs.CV

LightsOut: Diffusion-based Outpainting for Enhanced Lens Flare Removal

Shr-Ruei Tsai, Wei-Cheng Chang, Jie-Ying Lee, Chih-Hai Su, Yu-Lun Liu

机构 * National Yang Ming Chiao Tung University(国家阳明交通大学)

Comments ICCV 2025. Project page: https://ray-1026.github.io/lightsout/

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.15800 2025-10-20 cs.CV

ERNet: Efficient Non-Rigid Registration Network for Point Sequences

Guangzhao He, Yuxi Xiao, Zhen Xu, Xiaowei Zhou, Sida Peng

机构 * Zhejiang University(浙江大学)

Comments Accepted to ICCV 2025. Project Page: https://guangzhaohe.com/ernet

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.15749 2025-10-20 cs.CV

SEGA: A Stepwise Evolution Paradigm for Content-Aware Layout Generation with Design Prior

Haoran Wang, Bo Zhao, Jinghui Wang, Hanzhang Wang, Huan Yang, Wei Ji, Hao Liu, Xinyan Xiao

机构 * Baidu Inc.(百度公司) Nanjing University(南京大学) Harbin Institute of Technology(哈尔滨工业大学) Kuaishou Technology(快手科技)

Comments Accepted by ICCV-2025, Our project website is at: https://brucew91.github.io/SEGA.github.io/, 10 pages

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.15673 2025-10-20 cs.CV cs.AI

Valeo Near-Field: a novel dataset for pedestrian intent detection

Antonyo Musabini, Rachid Benmokhtar, Jagdish Bhanushali, Victor Galizzi, Bertrand Luvison, Xavier Perrotton

机构 * Valeo(瓦莱欧) BRAIN Division(BRAIN部门) Universite Paris-Saclay(巴黎-萨克雷大学) CEA(法国国家科学研究中心) List F-91120(法国)

Journal ref ICCV 2025 - 9th Workshop and Competition on Affective & Behavior Analysis in-the-wild (ABAW)

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.15026 2025-10-20 cs.CV

MOBIUS: Big-to-Mobile Universal Instance Segmentation via Multi-modal Bottleneck Fusion and Calibrated Decoder Pruning

Mattia Segu, Marta Tintore Gazulla, Yongqin Xian, Luc Van Gool, Federico Tombari

机构 * Google(谷歌) ETH Zurich(苏黎世联邦理工学院) INSAIT, Sofia University, St. Kliment Ohridski(INSAIT,索菲亚大学,圣克莱孟·奥赫里茨基)

Comments ICCV 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.13253 2025-10-20 cs.CV cs.AI cs.LG

End-to-End Multi-Modal Diffusion Mamba

Chunhao Lu, Qiang Lu, Meichen Dong, Jake Luo

机构 * China University of Petroleum-Beijing(中国石油大学(北京)) Leyard Optoelectronic(莱亚德光电) University of Wisconsin-Milwaukee(威斯康星大学密尔沃基分校)

Comments Accepted by ICCV 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2506.05342 2025-10-20 cs.CV cs.AI

Refer to Any Segmentation Mask Group With Vision-Language Prompts

Shengcao Cao, Zijun Wei, Jason Kuen, Kangning Liu, Lingzhi Zhang, Jiuxiang Gu, HyunJoon Jung, Liang-Yan Gui, Yu-Xiong Wang

机构 * University of Illinois Urbana-Champaign(伊利诺伊大学厄巴纳-香槟分校) Adobe

Comments ICCV 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2504.04126 2025-10-20 cs.CV cs.AI

Multi-identity Human Image Animation with Structural Video Diffusion

Zhenzhi Wang, Yixuan Li, Yanhong Zeng, Yuwei Guo, Dahua Lin, Tianfan Xue, Bo Dai

机构 * The Chinese University of Hong Kong(香港中文大学) Shanghai Artificial Intelligence Laboratory(上海人工智能实验室) The University of Hong Kong(香港大学)

Comments ICCV 2025 camera ready

详情

展开后加载摘要…

URL PDF HTML 收藏
2503.15671 2025-10-20 cs.CV

CHROME: Clothed Human Reconstruction with Occlusion-Resilience and Multiview-Consistency from a Single Image

Arindam Dutta, Meng Zheng, Zhongpai Gao, Benjamin Planche, Anwesha Choudhuri, Terrence Chen, Amit K. Roy-Chowdhury, Ziyan Wu

机构 * University of California, Riverside(加州大学河滨分校) United Imaging Intelligence(联合影像智能)

Comments Accepted at ICCV 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2503.14445 2025-10-20 cs.CV

Bolt3D: Generating 3D Scenes in Seconds

Stanislaw Szymanowicz, Jason Y. Zhang, Pratul Srinivasan, Ruiqi Gao, Arthur Brussee, Aleksander Holynski, Ricardo Martin-Brualla, Jonathan T. Barron, Philipp Henzler

机构 * Google Research(谷歌研究) VGG – University of Oxford(视觉感知集团 – 欧洲大学) Google DeepMind(谷歌DeepMind)

Comments ICCV 2025. Project page: https://szymanowiczs.github.io/bolt3d

详情

展开后加载摘要…

URL PDF HTML 收藏
2503.07465 2025-10-20 cs.CV

YOLOE: Real-Time Seeing Anything

Ao Wang, Lihao Liu, Hui Chen, Zijia Lin, Jungong Han, Guiguang Ding

机构 * School of Software, Tsinghua University(清华大学软件学院) BNRist, Tsinghua University(清华大学BNRist) Department of Automation, Tsinghua University(清华大学自动化系)

Comments ICCV 2025 Camera-ready Version

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.14976 2025-10-17 cs.CV cs.GR cs.RO

Ponimator: Unfolding Interactive Pose for Versatile Human-human Interaction Animation

Shaowei Liu, Chuan Guo, Bing Zhou, Jian Wang

Comments Accepted to ICCV 2025. Project page: https://stevenlsw.github.io/ponimator/

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.14960 2025-10-17 cs.CV cs.AI

C4D: 4D Made from 3D through Dual Correspondences

Shizun Wang, Zhenxiang Jiang, Xingyi Yang, Xinchao Wang

机构 * National University of Singapore(新加坡国立大学) The Hong Kong Polytechnic University(香港理工大学)

Comments Accepted to ICCV 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2408.11817 2025-10-17 cs.CV

GRAB: A Challenging GRaph Analysis Benchmark for Large Multimodal Models

Jonathan Roberts, Kai Han, Samuel Albanie

机构 * University of Cambridge(剑桥大学) The University of Hong Kong(香港大学)

Comments Accepted at ICCV 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2407.13764 2025-10-17 cs.CV

Shape of Motion: 4D Reconstruction from a Single Video

Qianqian Wang, Vickie Ye, Hang Gao, Weijia Zeng, Jake Austin, Zhengqi Li, Angjoo Kanazawa

机构 * UC Berkeley(加州大学伯克利分校) Google DeepMind(谷歌DeepMind) UC San Diego(加州大学圣地亚哥分校) Adobe Research(Adobe研究)

Comments ICCV 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.14705 2025-10-17 cs.CV

Leveraging Learned Image Prior for 3D Gaussian Compression

Seungjoo Shin, Jaesik Park, Sunghyun Cho

Comments Accepted to ICCV 2025 Workshop on ECLR

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.14672 2025-10-17 cs.CV

VTimeCoT: Thinking by Drawing for Video Temporal Grounding and Reasoning

Jinglei Zhang, Yuanfan Guo, Rolandos Alexandros Potamias, Jiankang Deng, Hang Xu, Chao Ma

机构 * Shanghai Jiao Tong University(上海交通大学) Noah’s Ark Lab(诺亚实验室) Imperial College London(伦敦帝国理工学院)

Comments Accepted by ICCV 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.14374 2025-10-17 cs.CV

Spatial Preference Rewarding for MLLMs Spatial Understanding

Han Qiu, Peng Gao, Lewei Lu, Xiaoqin Zhang, Ling Shao, Shijian Lu

机构 * S-Lab, Nanyang Technological University(南洋理工大学S实验室) Shanghai AI Laboratory(上海人工智能实验室) Sensetime Research(商汤科技研究院) Zhejiang University of Technology(浙江工业大学) UCAS-Terminus AI Lab,University of Chinese Academy of Sciences(中国科学院大学Terminus AI实验室)

Comments ICCV 2025

详情

展开后加载摘要…

URL PDF HTML 收藏