机构
*
Hangzhou City University(杭州城市大学)
;
University of Wisconsin–Madison(威斯康星大学麦迪逊分校)
;
Hangzhou Normal University(杭州师范大学)
;
Zhejiang Provincial Engineering Research Center for Real-Time SmartTech in Urban Security Governance(浙江省实时智能安防技术工程研究中心)
OpenFace 3.0: A Lightweight Multitask System for Comprehensive Facial Behavior Analysis
Jiewen Hu, Leena Mathur, Paul Pu Liang, Louis-Philippe Morency
机构
*
Carnegie Mellon University(卡内基梅隆大学)
;
Massachusetts Institute of Technology(麻省理工学院)
专题命中
视频多模态
:multimodal(abstract);分类 cs.CV
CommentsIEEE FG 2025, \c{opyright} 2025 IEEE. Personal use of this material is permitted. Permission from IEEE must be obtained for all other uses, in any current or future media, including reprinting/republishing this material for advertising or promotional purposes, creating new collective works, for resale or redistribution to servers or lists, or reuse of any copyrighted component of this work
机构
*
School of Computer Science and Technology, Chongqing University of Posts and Telecommunications(重庆邮电大学计算机科学与技术学院)
;
Chongqing Institute for Brain and Intelligence(重庆脑科学与智能技术研究院)
;
Guangyang Bay Laboratory(广阳湾实验室)
机构
*
Qingdao Institute of Software, College of Computer Science and Technology, China University of Petroleum (East China)(青岛软件研究所,计算机科学与技术学院,中国石油大学(华东))
;
Thrust of AI, Hong Kong University of Science and Technology (GuangZhou)(人工智能研究组,香港科学与技术大学(广州))
;
Shandong Key Laboratory of Intelligent Oil & Gas Industrial Software(山东省智能油气工业软件重点实验室)
;
Wangxuan Institute of Computer Technology, Peking University(王轩计算机技术研究所,北京大学)
;
School of Automation, Southeast University(自动化学院,东南大学)
;
School of Computer Science, Beijing Institute of Technology(计算机科学学院,北京理工大学)
;
School of Artificial Intelligence and Computer Science, Nantong University(人工智能与计算机科学学院,南通大学)
;
Thrust of Intelligent Transportation, Hong Kong University of Science and Technology (GuangZhou)(智能交通研究组,香港科学与技术大学(广州))
;
Department of Electrical Engineering and Electronics, University of Liverpool(电子工程与电子学院,利物浦大学)
SpaceR: Reinforcing MLLMs in Video Spatial Reasoning
Kun Ouyang, Yuanxin Liu, Haoning Wu, Yi Liu, Hao Zhou, Jie Zhou, Fandong Meng, Xu Sun
机构
*
National Key Laboratory for Multimedia Information Processing, School of Computer Science, Peking University(国家多媒体信息处理重点实验室,计算机学院,北京大学)
;
Nanyang Technological University(南洋理工大学)
;
WeChat AI, Tencent Inc., China(微信AI,腾讯公司,中国)
GTransPDM: A Graph-embedded Transformer with Positional Decoupling for Pedestrian Crossing Intention Prediction
Chen Xie, Ciyun Lin, Xiaoyu Zheng, Bowen Gong, Antonio M. López
机构
*
Department of Traffic Information and Control Engineering, Jilin University(交通信息与控制工程系,吉林大学)
;
BIT-Barcelona Innovative Transportation Research Group, Civil Engineering School, UPC Barcelona Tech(巴塞罗那创新交通研究组,土木工程学院,UPC巴塞罗那技术学院)
;
Computer Vision Center (CVC), Computer Science Department, Universitat Autònoma de Barcelona (UAB)(计算机视觉中心(CVC),计算机科学系,巴塞罗那自治大学(UAB))