Towards Temporal Compositional Reasoning in Long-Form Sports Videos
长期体育视频中的时间组成推理方法
Siyu Cao, Lu Zhang, Ruizhe Zeng, Zhi-yong Liu
机构
*
School of Artificial Intelligence, University of Chinese Academy of Sciences(中国科学院大学人工智能学院)
;
MAIS, Institute of Automation, Chinese Academy of Sciences(中国科学院自动化研究所MAIS)
机构
*
Graduate School of Fundamental Science and Engineering, Waseda University(早稻田大学基础科学与工程研究生院)
;
Generative AI Innovation Center, Amazon Web Services(亚马逊网络服务生成人工智能创新中心)
;
Amazon(亚马逊)
GuideMe: Multi-Domain Task Guidance and Intervention in Streaming Video
GuideMe:流视频中的多域任务指导与干预
Fang Liu, Jinpeng Chen, Ke Xu, Yuhao Liu, Huankang Guan, Xudong Lu, Bo Yang, Gerhard Hancke, Rui Liu, Rynson W. H. Lau
机构
*
City University of Hong Kong(香港城市大学)
;
Huawei Research(华为研究院)
;
University of Science and Technology of China(中国科学技术大学)
;
Chinese University of Hong Kong(香港中文大学)
;
City University of Hong Kong (Dongguan)(香港城市大学(东莞))
机构
*
Zhejiang University(浙江大学)
;
DAMO Academy, Alibaba Group(阿里巴巴达摩院)
;
The Hong Kong University of Science and Technology(香港科技大学)
;
Monash University(莫纳什大学)
;
TRE, Alibaba Group(阿里巴巴TRE)
机构
*
IMT Nord Europe, University of Lille, CNRS UMR 9189 - CRIStAL(IMT 北欧洲、里尔大学、法国国家科学研究中心 UMR 9189 - CRIStAL)
;
University of Lille, CNRS UMR 9189 - CRIStAL(里尔大学、法国国家科学研究中心 UMR 9189 - CRIStAL)
;
Explain
CDER-SME: A Cross-Device Event-RGB Micro-Expression Dataset under Multi-Level Stress Induction
CDER-SME:多级压力诱导下的跨设备事件-RGB微表情数据集
Jingting Li, Hui Sha, Su-Jing Wang
机构
*
State Key Laboratory of Cognitive Science and Mental Health, Institute of Psychology, Chinese Academy of Sciences(中国科学院心理研究所认知科学与心理健康国家重点实验室)
;
Department of Psychology, University of the Chinese Academy of Sciences(中国科学院大学心理学系)
;
School of Computer Science, Jiangsu University of Science and Technology(江苏科技大学计算机科学学院)
Decoupling Semantics and Logic: A Training-Free Coarse-to-Fine Pipeline for Video Retrieval-Augmented Generation
解耦语义与逻辑:一种无需训练的从粗到精的视频检索增强生成流水线
Jiaxin Dai, Zehang Wei, Jiamin Yan, Xiang Xiang
机构
*
School of Computer Science & Tech, Huazhong University of Science and Technology(华中科技大学计算机科学与技术学院)
;
School of AI and Automation, Huazhong University of Science and Technology(华中科技大学人工智能与自动化学院)
Imagine Before You Predict: Interleaved Latent Visual Reasoning for Video Event Prediction
在预测之前想象:用于视频事件预测的交错潜在视觉推理
Tianxiang Jiang, Linquan Wu, Sheng Xia, Songze Li, Ziang Yan, Haoyu Yang, Yu Qiao, Yi Wang
机构
*
University of Science and Technology of China(中国科学技术大学)
;
Shanghai AI Laboratory(上海人工智能实验室)
;
City University of Hong Kong(香港城市大学)
;
Nanjing University(南京大学)
;
Fudan University(复旦大学)
;
Zhejiang University(浙江大学)
;
University of Electronic Science and Technology of China(电子科技大学)
机构
*
MoE Key Laboratory of Brain-Machine Intelligence Technology, College of Artificial Intelligence, Nanjing University of Aeronautics(脑机智能技术关键实验室、人工智能学院、南京航空航天大学)
;
Dalian University of Technology(大连理工大学)
;
Nanjing University(南京大学)
;
National University of Singapore(新加坡国立大学)