arXivDaily arXiv每日学术速递 周一至周五更新
arXiv周末暂无论文更新,休息一下吧,周末愉快~~

期刊&会议

International Conference on Computer Vision · 会议 · Computer Vision

2025-08-14 至 2025-08-14 共收录 17
2508.09973 2025-08-14 cs.CV

PERSONA: Personalized Whole-Body 3D Avatar with Pose-Driven Deformations from a Single Image

Geonhee Sim, Gyeongsik Moon

机构 * Dept. of CSE, Korea University(计算机科学与工程系,韩国大学)

Comments Accepted to ICCV 2025. https://mks0601.github.io/PERSONA/

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.09949 2025-08-14 cs.CV cs.LG

Stable Diffusion Models are Secretly Good at Visual In-Context Learning

Trevine Oorloff, Vishwanath Sindagi, Wele Gedara Chaminda Bandara, Ali Shafahi, Amin Ghiasi, Charan Prakash, Reza Ardekani

机构 * Apple(苹果公司) University of Maryland - College Park(马里兰大学-College Park)

Comments Accepted to ICCV 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.09936 2025-08-14 cs.CV cs.DL

Quo Vadis Handwritten Text Generation for Handwritten Text Recognition?

Vittorio Pippi, Konstantina Nikolaidou, Silvia Cascianelli, George Retsinas, Giorgos Sfikas, Rita Cucchiara, Marcus Liwicki

机构 * University of Modena and Reggio Emilia(摩德纳和雷吉奥艾米利亚大学) Luleå University of Technology(吕勒奥技术大学) National Technical University of Athens(雅典国家技术大学) University of West Attica(西阿提卡大学)

Comments Accepted at ICCV Workshop VisionDocs

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.09886 2025-08-14 cs.CV cs.AI cs.CL

COME: Dual Structure-Semantic Learning with Collaborative MoE for Universal Lesion Detection Across Heterogeneous Ultrasound Datasets

Lingyu Chen, Yawen Zeng, Yue Wang, Peng Wan, Guo-chen Ning, Hongen Liao, Daoqiang Zhang, Fang Chen

机构 * College of Artificial Intelligence, Nanjing University of Aeronautics and Astronautics(人工智能学院,南京航空航天大学) ByteDance Inc.(字节跳动公司) Tsinghua University(清华大学) Shanghai Jiaotong University(上海交通大学)

Comments ICCV 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.09830 2025-08-14 cs.CV cs.AI cs.GR cs.LG cs.RO

RayletDF: Raylet Distance Fields for Generalizable 3D Surface Reconstruction from Point Clouds or Gaussians

Shenxing Wei, Jinxi Li, Yafei Yang, Siyuan Zhou, Bo Yang

机构 * vLAR Group, The Hong Kong Polytechnic University(vLAR组,香港理工大学)

Comments ICCV 2025 Highlight. Shenxing and Jinxi are co-first authors. Code and data are available at: https://github.com/vLAR-group/RayletDF

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.09811 2025-08-14 cs.CV cs.AI cs.CE cs.LG cs.RO

TRACE: Learning 3D Gaussian Physical Dynamics from Multi-view Videos

Jinxi Li, Ziyang Song, Bo Yang

机构 * vLAR Group, The Hong Kong Polytechnic University(vLAR小组,香港理工大学)

Comments ICCV 2025. Code and data are available at: https://github.com/vLAR-group/TRACE

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.09661 2025-08-14 cs.CV

NegFaceDiff: The Power of Negative Context in Identity-Conditioned Diffusion for Synthetic Face Generation

Eduarda Caldeira, Naser Damer, Fadi Boutros

机构 * Fraunhofer IGD(弗劳恩霍夫研究所) TU Darmstadt(图宾根大学)

Comments Accepted at ICCV Workshops

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.04611 2025-08-14 cs.CV cs.RO

BridgeDepth: Bridging Monocular and Stereo Reasoning with Latent Alignment

Tongfan Guan, Jiaxin Guo, Chen Wang, Yun-Hui Liu

机构 * The Chinese University of Hong Kong(香港中文大学) University at Buffalo(布法罗大学) Spatial AI & Robotics Lab(空间人工智能与机器人实验室)

Comments ICCV 2025 Highlight

Journal ref IEEE/CVF International Conference on Computer Vision (ICCV), 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2501.14729 2025-08-14 cs.CV

HERMES: A Unified Self-Driving World Model for Simultaneous 3D Scene Understanding and Generation

Xin Zhou, Dingkang Liang, Sifan Tu, Xiwu Chen, Yikang Ding, Dingyuan Zhang, Feiyang Tan, Hengshuang Zhao, Xiang Bai

机构 * Huazhong University of Science and Technology(华中科技大学) MEGVII Technology(梅格维七科技) Mach Drive(马奇驱动) The University of Hong Kong(香港大学)

Comments Accepted by ICCV 2025. The code is available at https://github.com/LMD0311/HERMES

详情

展开后加载摘要…

URL PDF HTML 收藏
2501.06399 2025-08-14 cs.CV cs.AI cs.CR cs.CY cs.LG

GenAI Confessions: Black-box Membership Inference for Generative Image Models

Matyas Bohacek, Hany Farid

机构 * Stanford University(斯坦福大学) University of California, Berkeley(加州大学伯克利分校)

Comments https://genai-confessions.github.io

Journal ref ICCV-W 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2501.03012 2025-08-14 cs.AI cs.CL cs.CV

Analyzing Finetuning Representation Shift for Multimodal LLMs Steering

Pegah Khayatan, Mustafa Shukor, Jayneel Parekh, Arnaud Dapogny, Matthieu Cord

机构 * ISIR, Sorbonne Université(ISIR,索邦大学)

Comments ICCV 2025. The first three authors contributed equally. Project page and code: https://pegah- kh.github.io/projects/lmm-finetuning-analysis-and-steering/

详情

展开后加载摘要…

URL PDF HTML 收藏
2412.01787 2025-08-14 cs.CV cs.AI cs.LG

Pretrained Reversible Generation as Unsupervised Visual Representation Learning

Rongkun Xue, Jinouwen Zhang, Yazhe Niu, Dazhong Shen, Bingqi Ma, Yu Liu, Jing Yang

机构 * Xi’an Jiaotong University(西安交通大学) Shanghai AI Laboratory(上海人工智能实验室) SenseTime(商汤科技) The Chinese University of Hong Kong(香港中文大学) Nanjing University of Aeronautics and Astronautics(南京航空航天大学)

Comments Accepted by ICCV 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.09415 2025-08-14 cs.CV cs.AI

RampNet: A Two-Stage Pipeline for Bootstrapping Curb Ramp Detection in Streetscape Images from Open Government Metadata

John S. O'Meara, Jared Hwang, Zeyu Wang, Michael Saugstad, Jon E. Froehlich

机构 * Issaquah High School(伊萨夸高中) University of Washington(华盛顿大学)

Comments Accepted to the ICCV'25 Workshop on Vision Foundation Models and Generative AI for Accessibility: Challenges and Opportunities

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.09372 2025-08-14 cs.CV cs.AI cs.IR cs.LG

A Signer-Invariant Conformer and Multi-Scale Fusion Transformer for Continuous Sign Language Recognition

Md Rezwanul Haque, Md. Milon Islam, S M Taslim Uddin Raju, Fakhri Karray

机构 * University of Waterloo(滑铁卢大学) Mohamed bin Zayed University of Artificial Intelligence(穆罕默德·本·扎耶德人工智能大学)

Comments Accepted for the IEEE/CVF International Conference on Computer Vision (ICCV), Honolulu, Hawaii, USA. 1st MSLR Workshop 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.09362 2025-08-14 cs.CV cs.AI cs.LG

FusionEnsemble-Net: An Attention-Based Ensemble of Spatiotemporal Networks for Multimodal Sign Language Recognition

Md. Milon Islam, Md Rezwanul Haque, S M Taslim Uddin Raju, Fakhri Karray

机构 * University of Waterloo(滑铁卢大学) Mohamed bin Zayed University of Artificial Intelligence(穆罕默德·本·扎耶德人工智能大学)

Comments Accepted for the IEEE/CVF International Conference on Computer Vision (ICCV), Honolulu, Hawaii, USA. 1st MSLR Workshop 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2508.09262 2025-08-14 cs.CV cs.LG

Harnessing Input-Adaptive Inference for Efficient VLN

Dongwoo Kang, Akhil Perincherry, Zachary Coalson, Aiden Gabriel, Stefan Lee, Sanghyun Hong

机构 * Oregon State University(俄勒冈州立大学)

Comments Accepted to ICCV 2025 [Poster]

详情

展开后加载摘要…

URL PDF HTML 收藏
2507.01367 2025-08-14 cs.CV

3D Gaussian Splatting Driven Multi-View Robust Physical Adversarial Camouflage Generation

Tianrui Lou, Xiaojun Jia, Siyuan Liang, Jiawei Liang, Ming Zhang, Yanjun Xiao, Xiaochun Cao

机构 * Sun Yat-Sen University(中山大学) Peng Cheng Laboratory(鹏城实验室) Nanyang Technological University(南洋理工大学) National University of Singapore(新加坡国立大学) National Key Laboratory of Science and Technology on Information System Security(国家信息系统安全科学技术重点实验室) Nsfocus

Comments Accepted by ICCV 2025

详情

展开后加载摘要…

URL PDF HTML 收藏