Multi-event Video-Text Retrieval
多事件视频-文本检索
机构 * LMU Munich(慕尼黑大学) ; Munich Center for Machine Learning(慕尼黑机器学习中心) ; University of Oxford(牛津大学)
专题命中 视频多模态 :multi-modal(abstract);分类 cs.CV
AI总结 本文提出多事件视频-文本检索任务,设计Me-Retriever模型,通过关键事件表示和新损失函数提升视频-文本检索性能。
Comments [fixed typos in equations] accepted to ICCV2023 Poster; some figures are not supported when viewed online, please download the file and view locally