Overview of Tencent Multi-modal Ads Video Understanding Challenge
Zhenzhi Wang, Liyu Wu, Zhimin Li, Jiangfeng Xiong, Qinglin Lu
专题命中
视频多模态
:multi-modal(title,abstract);分类 cs.CV
Comments8-page extended version of our challenge paper in ACM MM 2021. It presents the overview of grand challenge "Multi-modal Ads Video Understanding" in ACM MM 2021. Our grand challenge is also the Tencent Advertising Algorithm Competition (TAAC) 2021
CommentsThis is the accepted version of the following article: Kragh M, Underwood J. Multimodal obstacle detection in unstructured environments with conditional random fields. J Field Robotics. 2019, 1-20., which has been published in final form at https://doi.org/10.1002/rob.21866
NARU: A Benchmark for NARrative Evolution and Cultural Nuance Understanding in Japanese Extreme Long Video
NARU:用于理解日语超长视频中叙事演变与文化细微差别的基准
Yuheng Huang, Jianlang Chen, Jiayang Song, Hua Qi, Aza Kai, Vincent Markert, Edison Marrese-Taylor, Jianjun Zhao, Lei Ma
机构
*
The University of Tokyo(东京大学)
;
Kyushu University(九州大学)
;
Macau University of Science and Technology(澳门科技大学)
;
Infinimind Japan Inc.(Infinimind日本公司)
;
University of Alberta(阿尔伯塔大学)
Action Anticipation at a Glimpse: To What Extent Can Multimodal Cues Replace Video?
瞬间动作预见:多模态线索能替代视频到何种程度?
Manuel Benavent-Lledo, Konstantinos Bacharidis, Victoria Manousaki, Konstantinos Papoutsakis, Antonis Argyros, Jose Garcia-Rodriguez
机构
*
Universidad de Alicante(阿利坎特大学)
;
Foundation for Research and Technology-Hellas(希腊基础研究与技术基金会)
;
University of Crete(克里特大学)
;
Hellenic Mediterranean University(希腊地中海大学)
机构
*
School of Computer Science, Peking University(北京大学计算机科学学院)
;
Microsoft Research Asia(微软亚洲研究院)
;
Princeton University(普林斯顿大学)
;
Tsinghua University(清华大学)
Comments11 pages, 4 figures. Published in the Proceedings of the 1st Workshop on Shaping Future Human Connection: Social Augmentation through XR Technologies (SAXR 2026), April 13, 2026, Barcelona, Spain
Journal refProceedings of the 1st Workshop on Shaping Future Human Connection: Social Augmentation through XR Technologies (SAXR 2026), CEUR Workshop Proceedings, Vol. 4226, pp. 252-262, 2026
When One Modality Is Not Enough: Multimodal Sex and Life-Stage Classification of Red Deer from Aerial RGB-Thermal Video
当单模态不够时:基于航拍RGB-热红外视频的马鹿多模态性别与生命阶段分类
Hugo Markoff, Christoph Praschl, Ivan Ludoški, Sara Beery, Michael Ørsted, David C. Schedl
机构
*
Aalborg University(奥尔堡大学)
;
University of Applied Sciences Upper Austria(上奥地利应用科学大学)
;
University of Novi Sad(诺威萨德大学)
;
Massachusetts Institute of Technology(麻省理工学院)