A Survey on Vision-Language-Action Models for Embodied AI
具身人工智能中视觉-语言-动作模型的综述
机构 * Chinese University of Hong Kong(香港中文大学) ; University of Bristol(布里斯托大学) ; Huawei Noah’s Ark Laboratory(华为诺亚实验室)
AI总结 本文综述了具身人工智能中视觉-语言-动作模型的发展,探讨了其在机器人任务中的应用,分类了三种主要研究方向,并讨论了面临的挑战与未来方向。
Comments Project page: https://github.com/yueen-ma/Awesome-VLA
Journal ref IEEE Transactions on Neural Networks and Learning Systems (Early Access), 2026