HiST-VLA: A Hierarchical Spatio-Temporal Vision-Language-Action Model for End-to-End Autonomous Driving
HiST-VLA:一种用于端到端自动驾驶的分层时空视觉-语言-动作模型
机构 * Bosch Corporate Research(博世企业研究) ; School of Communication and Information Engineering(信息工程学院)
专题命中 端到端驾驶 :autonomous driving(title,abstract);分类 cs.RO、cs.CV、cs.AI
AI总结 HiST-VLA通过分层时空视觉-语言-动作模型提升自动驾驶轨迹生成的精度与效率,实现端到端的自动驾驶系统。