EVE: Towards End-to-End Video Subtitle Extraction with Vision-Language Models
EVE:基于视觉-语言模型的端到端视频字幕提取
Haiyang Yu, Mengyang Zhao, Jinghui Lu, Ke Niu, Yanjie Wang, Weijie Yin, Weitao Jia, Teng Fu, Yang Liu, Jun Liu, Hong Chen
机构
*
College of Computer Science and Artificial Intelligence, Fudan University(复旦大学计算机科学与人工智能学院)
;
ByteDance Inc.(字节跳动公司)
;
College of Electronic and Information Engineering, Tongji University(同济大学电子与信息工程学院)
;
School of Computing and Communications, Lancaster University(兰卡斯特大学计算机与通讯学院)
The Autonomy-Alignment Problem in Open-Ended Learning Robots: Formalising the Purpose Framework
开放性学习机器人中的自主性-对齐问题:目的框架的正式化
Gianluca Baldassarre, Richard J. Duro, Emilio Cartoni, Mehdi Khamassi, Alejandro Romero, Vieri Giuliano Santucci
机构
*
Institute of Cognitive Sciences and Technologies, National Research Council(认知科学与技术研究所,国家研究理事会)
;
Integrated Group for Engineering Research, CITIC, Universidade da Coruña(工程研究联合组,CITIC,科鲁纳大学)
;
Institute of Intelligent Systems and Robotics, Sorbonne University / CNRS(智能系统与机器人研究所,索邦大学 / CNRS)