arXivDaily arXiv每日学术速递 周一至周五更新

期刊&会议

Winter Conference on Applications of Computer Vision · 会议 · Computer Vision

2026-01-13 至 2026-01-13 共收录 7
2511.17068 2026-01-13 cs.CV cs.AI

ReBrain: Brain MRI Reconstruction from Sparse CT Slice via Retrieval-Augmented Diffusion

ReBrain: 通过检索增强扩散模型从稀疏CT切片重建脑部MRI

Junming Liu, Yifei Sun, Weihua Cheng, Yujin Kang, Yirong Chen, Ding Wang, Guosun Zeng

机构 * Tongji University(同济大学) Shanghai Artificial Intelligence Laboratory(上海人工智能实验室)

AI总结 ReBrain通过检索增强扩散模型从稀疏CT切片重建脑部MRI,利用BBDM和ControlNet实现结构连续性,提升稀疏条件下的跨模态重建性能。

Comments 16 pages, 12 figures, 7 tables; Accepted by WACV 2026

详情

展开后加载摘要…

URL PDF HTML 收藏
2507.14959 2026-01-13 cs.CV cs.PF

Polymorph: Energy-Efficient Multi-Label Classification for Video Streams on Embedded Devices

Polymorph: 为嵌入式设备上的视频流实现高效多标签分类

Saeid Ghafouri, Mohsen Fayyaz, Xiangchen Li, Deepu John, Bo Ji, Dimitrios Nikolopoulos, Hans Vandierendonck

机构 * Queen’s University Belfast(女王大学贝尔法斯特分校) Microsoft(微软) Virginia Tech(弗吉尼亚理工大学) University College Dublin(都柏林大学)

AI总结 Polymorph通过模块化轻量级LoRA适配器实现嵌入式设备上视频流的高效多标签分类,降低能耗并提升mAP性能。

Comments Accepted at the IEEE/CVF winter conference on applications of computer vision (WACV 2026)

详情

展开后加载摘要…

URL PDF HTML 收藏
2503.11742 2026-01-13 cs.CV cs.AI

Safe Vision-Language Models via Unsafe Weights Manipulation

通过不安全权重操作实现安全的视觉-语言模型

Moreno D'Incà, Elia Peruzzo, Xingqian Xu, Humphrey Shi, Nicu Sebe, Massimiliano Mancini

机构 * University of Trento(特伦托大学) NVIDIA(NVIDIA公司) Georgia Tech(佐治亚理工学院)

AI总结 本文提出UWM方法,通过不训练的方式提升视觉-语言模型在不安全查询上的安全性,同时在安全输入上表现更优。

Comments WACV 2026

详情

展开后加载摘要…

URL PDF HTML 收藏
2601.06537 2026-01-13 cs.CV

Towards Egocentric 3D Hand Pose Estimation in Unseen Domains

面向未见领域的自体视觉三维手姿态估计

Wiktor Mucha, Michael Wray, Martin Kampel

机构 * Computer Vision Lab, TU Wien(维也纳技术大学计算机视觉实验室) SoftServe Inc.(SoftServe公司) University of Bristol(布里斯托大学)

AI总结 V-HPOT通过虚拟相机空间和自监督优化提升跨领域三维手姿态估计性能,减少71%的平均姿态误差。

Comments Accepted at WACV 2026

详情

展开后加载摘要…

URL PDF HTML 收藏
2601.06484 2026-01-13 cs.CV cs.AI

Learning Domain Agnostic Latent Embeddings of 3D Faces for Zero-shot Animal Expression Transfer

学习领域无关的3D人脸潜在嵌入以实现零样本动物表情迁移

Yue Wang, Lawrence Amadi, Xiang Gao, Yazheng Chen, Yuanpeng Liu, Ning Lu, Xianfeng Gu

机构 * Stony Brook University(石英溪大学) Futurewei Technologies(未来讯技术)

AI总结 本文提出了一种零样本框架,通过学习领域无关的3D人脸潜在嵌入,实现人类表情到动物面部的跨物种表情迁移。

Comments WACV 2026 Workshop LENS

详情

展开后加载摘要…

URL PDF HTML 收藏
2601.06460 2026-01-13 cs.CV cs.AI cs.CL

Tone Matters: The Impact of Linguistic Tone on Hallucination in VLMs

语气至关重要:语言语气对VLMs幻觉影响的研究

Weihao Hong, Zhiyuan Jiang, Bingyu Shen, Xinlei Guan, Yangyi Feng, Meng Xu, Boyang Li

机构 * Department of Computer Science and Technology, Kean University(计算机科学与技术系,凯恩大学) Department of Computer Science and Engineering, University of Notre Dame(计算机科学与工程系,圣母大学)

AI总结 本文研究了提示语气对VLMs幻觉的影响,通过Ghost-100数据集发现幻觉率与提示强度非线性相关,揭示模型在处理结构性强制时的局限性。

Comments 10 pages, 6 figures, WACV Workshop

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.25077 2026-01-13 cs.CV eess.IV

Neighborhood Feature Pooling for Remote Sensing Image Classification

邻域特征池化用于遥感图像分类

Fahimeh Orvati Nia, Amirmohammad Mohammadi, Salim Al Kharsa, Pragati Naikare, Zigfried Hampel-Arias, Joshua Peeples

机构 * Dept. of Electrical & Computer Engineering, Texas A&M University, College Station, TX, USA(电子与计算机工程系,德克萨斯A&M大学) Dept. of Computer Science & Engineering, Texas A&M University, College Station, TX, USA(计算机科学与工程系,德克萨斯A&M大学) Los Alamos National Laboratory, Los Alamos, NM, USA(洛斯阿拉莫斯国家实验室)

AI总结 本文提出邻域特征池化方法,通过聚合局部相似性模式提升遥感图像分类性能,实验表明其在多个数据集上均优于传统池化策略。

Comments 10 pages, 4 figures, accepted at the IEEE/CVF Winter Conference on Applications of Computer Vision (WACV) 2026, 3rd Workshop on Computer Vision for Earth Observation (CV4EO)

详情

展开后加载摘要…

URL PDF HTML 收藏