arXivDaily arXiv每日学术速递 周一至周五更新

期刊&会议

Winter Conference on Applications of Computer Vision · 会议 · Computer Vision

2025-12-05 至 2025-12-05 共收录 5
2512.04927 2025-12-05 cs.CV

Virtually Unrolling the Herculaneum Papyri by Diffeomorphic Spiral Fitting

通过仿射螺旋拟合虚拟展开赫库拉尼姆莎草纸

Paul Henderson

机构 * University of Glasgow(格拉斯哥大学)

AI总结 通过仿射螺旋拟合方法实现莎草纸的虚拟展开,自动拟合表面模型以生成连续的2D展开表示。

Comments Accepted at WACV 2026

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.07828 2025-12-05 cs.CV

MMHOI: Modeling Complex 3D Multi-Human Multi-Object Interactions

MMHOI:建模复杂3D多人类多物体交互

Kaen Kogashi, Anoop Cherian, Meng-Yu Jennifer Kuo

机构 * Mitsubishi Electric Japan(三菱电机日本公司) Mitsubishi Electric Research Labs(三菱电机研究实验室) Nara Women’s University(奈良女子大学)

AI总结 MMHOI提出了一种大规模多人类多物体交互数据集和端到端Transformer网络,用于建模复杂3D人类-物体交互,实现了最先进的性能。

Comments Accepted to WACV 2026

详情

展开后加载摘要…

URL PDF HTML 收藏
2512.04487 2025-12-05 cs.CV

Controllable Long-term Motion Generation with Extended Joint Targets

通过扩展目标关节实现可控的长期运动生成

Eunjong Lee, Eunhee Kim, Sanghoon Hong, Eunho Jung, Jihoon Kim

机构 * Cinamon Inc.(Cinamon公司)

AI总结 COMET通过扩展目标关节实现可控的长期运动生成,利用高效Transformer基于条件VAE实现精确交互控制,并通过参考引导反馈机制确保长期稳定性。

Comments WACV 2026

详情

展开后加载摘要…

URL PDF HTML 收藏
2512.04356 2025-12-05 cs.CV cs.AI cs.CL cs.LG

Mitigating Object and Action Hallucinations in Multimodal LLMs via Self-Augmented Contrastive Alignment

通过自增强对比对齐缓解多模态大语言模型中的对象和动作幻觉

Kai-Po Chang, Wei-Yuan Cheng, Chi-Pin Huang, Fu-En Yang, Yu-Chiang Frank Wang

机构 * Graduate Institute of Communication Engineering, National Taiwan University(国家交通大学通信工程研究所) NVIDIA

AI总结 SANTA框架通过自增强对比对齐方法,有效缓解多模态大语言模型中的对象和动作幻觉问题。

Comments IEEE/CVF Winter Conference on Applications of Computer Vision (WACV) 2026. Project page: https://kpc0810.github.io/santa/

详情

展开后加载摘要…

URL PDF HTML 收藏
2411.19717 2025-12-05 cs.CV cs.AI cs.LG cs.RO

MonoPP: Metric-Scaled Self-Supervised Monocular Depth Estimation by Planar-Parallax Geometry in Automotive Applications

MonoPP: 基于平面-视差几何的汽车应用中利用度量尺度自监督单目深度估计

Gasser Elazab, Torben Gräber, Michael Unterreiner, Olaf Hellwich

机构 * CARIAD SE(CARIAD公司) Technische Universität Berlin(柏林技术大学)

AI总结 MonoPP通过平面-视差几何实现自监督单目深度估计,利用多帧和单帧网络及姿态网络,实现汽车应用中度量尺度深度预测的先进性能。

Comments Accepted at WACV 25, project page: https://mono-pp.github.io/

Journal ref Proceedings of the 2025 IEEE/CVF Winter Conference on Applications of Computer Vision (WACV), Tucson, AZ, USA, 26 February 2025, pp. 2777-2787

详情

展开后加载摘要…

URL PDF HTML 收藏