Advancing Semantic Future Prediction through Multimodal Visual Sequence Transformers
通过多模态视觉序列变压器推进语义未来预测
机构 * Archimedes, Athena Research Center(阿基米德研究中心) ; National Technical University of Athens(希腊国家技术大学) ; University of Crete(克里特大学) ; IACM-Forth(第四研究机构(IACM-Forth))
AI总结 FUTURIST通过多模态视觉序列变压器架构实现高效的多模态未来语义预测,提升预测精度并简化训练流程。
Comments CVPR 2025