arXivDaily arXiv每日学术速递 周一至周五更新
arXiv周末暂无论文更新,休息一下吧,周末愉快~~

AI 大模型

视频大模型

视频理解、视频生成、视频语言模型和时序视觉推理。

2026-02-19 至 2026-02-19 共收录 2 信号源:cs.CV, eess.IV, cs.MM

1. 视频理解 2 篇

2602.16545 2026-02-19 cs.CV cs.LG 74%

Let's Split Up: Zero-Shot Classifier Edits for Fine-Grained Video Understanding

让我们拆分:用于细粒度视频理解的零样本分类器编辑

Kaiting Liu, Hazel Doughty

机构 * Leiden University(莱顿大学)

专题命中 视频理解 :video understanding(title);分类 cs.CV

AI总结 本文提出零样本分类器编辑方法,通过细粒度视频理解提升分类精度,优于视觉-语言基线。

Comments ICLR 2026

详情

展开后加载摘要…

URL PDF HTML 收藏
2602.16231 2026-02-19 cs.CV 57%

DataCube: A Video Retrieval Platform via Natural Language Semantic Profiling

DataCube: 通过自然语言语义分析实现的视频检索平台

Yiming Ju, Hanyu Zhao, Quanyue Ma, Donglin Hao, Chengwei Wu, Ming Li, Songjing Wang, Tengfei Pan

机构 * Beijing Academy of Artificial Intelligence(北京人工智能研究院)

专题命中 视频理解 :video understanding(abstract);分类 cs.CV

AI总结 DataCube通过自然语言语义分析实现视频检索,提供高效的视频处理和多维检索功能。

Comments This paper is under review for the IJCAI-ECAI 2026 Demonstrations Track

详情

展开后加载摘要…

URL PDF HTML 收藏