Video Event Reasoning and Prediction by Fusing World Knowledge from LLMs with Vision Foundation Models
通过融合来自LLM的世界知识与视觉基础模型进行视频事件推理与预测
L'ea Dubois, Klaus Schmidt, Chengyu Wang, Ji-Hoon Park, Lin Wang, Santiago Munoz
机构
*
INRIA(法国国家信息与自动化研究所)
;
Max Planck Institute for Intelligent Systems(人工智能研究所)
;
San Francisco State University(旧金山州立大学)
;
Seoul AI Institute (SAII)(首尔人工智能研究所)
;
Vision & Robotics Center, Tsinghua University(清华大学视觉与机器人中心)
;
Polytechnic University of Madrid(马德里理工大学)
Hongzhe Bi, Hengkai Tan, Shenghao Xie, Zeyuan Wang, Shuhe Huang, Haitian Liu, Ruowen Zhao, Yao Feng, Chendong Xiang, Yinze Rong, Hongyan Zhao, Hanyu Liu, Zhizhong Su, Lei Ma, Hang Su, Jun Zhu
机构
*
Dept. of Comp. Sci. and Tech., Institute for AI, BNRist Center, THBI Lab, Tsinghua-Bosch Joint ML Center, Tsinghua University(计算机科学与技术系,人工智能研究院,BNRist中心,THBI实验室,清华-博世联合机器学习中心,清华大学)
;
Peking University(北京大学)
;
Horizon Robotics
Physically Realistic Sequence-Level Adversarial Clothing for Robust Human-Detection Evasion
物理真实序列级对抗性服装用于鲁棒的人体检测规避
Dingkun Zhou, Patrick P. K. Chan, Hengxu Wu, Shikang Zheng, Ruiqi Huang, Yuanjie Zhao
机构
*
School of Future Technology, South China University of Technology(未来技术学院,华南理工大学)
;
Institute for Interdisciplinary Information Sciences, Tsinghua University(交叉信息研究院,清华大学)
PHI: Bridging Domain Shift in Long-Term Action Quality Assessment via Progressive Hierarchical Instruction
Kanglei Zhou, Hubert P. H. Shum, Frederick W. B. Li, Xingxing Zhang, Xiaohui Liang
机构
*
State Key Laboratory of Virtual Reality Technology and Systems, Beihang University(虚拟现实技术与系统国家重点实验室,北京航空航天大学)
;
Department of Computer Science, Durham University(杜伦大学计算机科学系)
;
Department of Computer Science and Technology, Institute for AI, BNRist Center, Tsinghua-Bosch Joint ML Center, THBI Lab, Tsinghua University(人工智能研究院,北京人工智能研究院,清华大学-博世联合机器学习中心,THBI实验室,清华大学计算机科学与技术系)
;
Zhongguancun Laboratory, Beijing(中关村实验室,北京)
专题命中
动作与事件理解
:long video(abstract);分类 cs.CV
CommentsAccepted by IEEE Transactions on Image Processing
Physical Autoregressive Model for Robotic Manipulation without Action Pretraining
Zijian Song, Sihan Qin, Tianshui Chen, Liang Lin, Guangrun Wang
机构
*
Sun Yat-sen University(中山大学)
;
Guangdong Key Laboratory of Big Data Analysis and Processing(广东大数据分析与处理重点实验室)
;
X-Era AI Lab(X-Era人工智能实验室)
;
Guangdong University of Technology(广东工业大学)