OmniTraffic: A Controllable Generation Pipeline and Benchmark for Spatio-Temporal Traffic Reasoning
OmniTraffic:面向时空交通推理的可控生成流水线与基准
Maonan Wang, Zhengyan Huang, Kemou Jiang, Yuhang Fu, Jiayue Zhu, Yuxin Cai, Xingchen Zou, Qiaosheng Zhang, Yi Yu, Ding Wang, Xi Chen, Ben M. Chen, Yuxuan Liang, Zhiyong Cui, Man On Pun, Yirong Chen
机构
*
The Chinese University of Hong Kong, Shenzhen(香港中文大学(深圳))
;
Shanghai AI Lab(上海人工智能实验室)
;
Beihang University(北京航空航天大学)
;
Nanyang Technological University(南洋理工大学)
;
The Hong Kong University of Science and Technology (Guangzhou)(香港科技大学(广州))
;
The Chinese University of Hong Kong(香港中文大学)
机构
*
School of Intelligence Science and Technology, Peking University(北京大学智能科学与技术学院)
;
State Key Laboratory of General Artificial Intelligence, Peking University(北京大学通用人工智能国家重点实验室)
;
Yuanpei College, Peking University(北京大学元培学院)
;
Institute for Artificial Intelligence, Peking University(北京大学人工智能研究院)
;
School of Psychological and Cognitive Sciences, Peking University(北京大学心理学与认知科学学院)
;
University of Wisconsin-Madison(威斯康星大学麦迪逊分校)
SagaQA: A Multi-hop Reasoning Benchmark for Long-form Narrative Understanding in TV Series
SagaQA:面向电视剧长篇叙事理解的多跳推理基准
Galann Pennec, Zhengyuan Liu, Nicholas Asher, Philippe Muller, Nancy F. Chen
机构
*
IRIT, University of Toulouse, France(法国图卢兹大学IRIT中心)
;
Agency for Science, Technology and Research (A*STAR), Singapore(新加坡科技研究局)
;
CNRS, IRIT, France(法国CNRS与IRIT)
CoCoVideo: The High-Quality Commercial-Model-Based Contrastive Benchmark for AI-Generated Video Detection
CoCoVideo: 基于商业模型的高质量对比基准用于AI生成视频检测
Huidong Feng, Wentao Chen, Jie Chen, Xinqi Cai, Ruolong Ma, Yinglin Zheng, Yuxin Lin, Ming Zeng
机构
*
School of Informatics, Xiamen University(厦门大学信息学院)
;
China Academy of Information and Communications Technology(中国信息通信技术研究院)
;
AI Transcend Pte. Ltd.(AI Transcend有限公司)
机构
*
Shanghai Jiao Tong University(上海交通大学)
;
Eastern Institute of Technology(东部技术研究院)
;
School of Electronic Information and Electrical Engineering(电子信息与电气工程学院)
;
Li Auto(力汽车)
;
National University of Singapore(新加坡国立大学)
;
Tsinghua University(清华大学)
;
Ningbo Key Laboratory of Spatial Intelligence and Digital Derivative(宁波空间智能与数字衍生实验室)
;
Ningbo Institute of Digital Twin(宁波数字孪生研究院)
机构
*
University of California, Los Angeles(加州大学洛杉矶分校)
;
University of Pittsburgh(匹兹堡大学)
;
Fudan University(复旦大学)
;
University of California, Riverside(加州大学河滨分校)
;
Hong Kong University of Science(香港科学大学)
;
Maharishi International University(玛希拉国际大学)
Hi-GaTA: Hierarchical Gated Temporal Aggregation Adapter for Surgical Video Report Generation
Hi-GaTA:用于外科视频报告生成的分层门控时间聚合适配器
Kedi Sun, Chaohui Dang, Yue Feng, James Glasbey, Theodoros N. Arvanitis, Le Zhang
机构
*
School of Engineering, College of Engineering and Physical Sciences, University of Birmingham, Birmingham, UK(英国伯明翰大学工程学院)
;
School of Computer Science, University of Birmingham, Birmingham, UK(英国伯明翰大学计算机科学学院)
;
Department of Applied Health Sciences, University of Birmingham, Birmingham, UK(英国伯明翰大学应用健康科学系)
BARISTA: A Multi-Task Egocentric Benchmark for Compositional Visual Understanding
BARISTA:一种多任务第一人称视角基准,用于组合视觉理解
Patrick Knab, Orgest Xhelili, Inis Buzi, Drago Andres Guggiana Nilo, Mohd Saquib Khan, Lorenz Kolb, Manuel Scherzer, Kerem Yildirir, Christian Bartelt, Philipp Johannes Schubert
机构
*
Ramblr.ai Research(Ramblr.ai 研究院)
;
Technical University of Clausthal(Clausthal 技术大学)
机构
*
University of Electronic Science and Technology of China(电子科学与技术大学)
;
Peking University(北京大学)
;
Institute of Automation, Chinese Academy of Sciences(中国科学院自动化研究所)
;
Zhongguancun Academy(中关村学院)
机构
*
The University of Hong Kong(香港大学)
;
Fudan University(复旦大学)
;
Zhejiang University(浙江大学)
;
Hong Kong University of Science and Technology(香港科学与技术大学)
;
University of Sydney(悉尼大学)
;
Alibaba Group(阿里巴巴集团)
OmniVTG: A Large-Scale Dataset and Training Paradigm for Open-World Video Temporal Grounding
OmniVTG:一种大规模数据集和开放世界视频时间定位的训练范式
Minghang Zheng, Zihao Yin, Yi Yang, Yuxin Peng, Yang Liu
机构
*
Wangxuan Institute of Computer Technology, Peking University(北京大学王轩计算机技术研究所)
;
State Key Laboratory of General Artificial Intelligence, Peking University(北京大学通用人工智能国家重点实验室)
;
Central Media Technology Institute, Huawei Technologies Ltd.(华为技术有限公司中央媒体技术研究所)
;
PKU-WUHAN Institute for Artificial Intelligence, Peking University(北京大学武汉人工智能研究所)
机构
*
Hangzhou Institute for Advanced Study, UCAS(浙江大学杭州高等研究院)
;
Computer Network Information Center, CAS(中国科学院计算机网络信息中心)
;
Department of AI Infrastructure, Bilibili Inc.(B站人工智能基础设施部)
Eevee: Towards Close-up High-resolution Video-based Virtual Try-on
Eevee:迈向基于视频的高分辨率虚拟试衣
Jianhao Zeng, Yancheng Bai, Ruidong Chen, Xuanpu Zhang, Lei Sun, Dongyang Jin, Ryan Xu, Nannan Zhang, Dan Song, Xiangxiang Chu
机构
*
Amap, Alibaba Group(高德,阿里巴巴集团)
;
Tianjin University(天津大学)
;
Shenzhen Institutes of Advanced Technology, Chinese Academy of Sciences(中国科学院深圳先进技术研究院)