VAT: Vision Action Transformer by Unlocking Full Representation of ViT
通过解锁ViT的完整表示构建视觉动作变换器
机构 * Hong Kong University of Science and Technology (Guangzhou)(香港科技大学(广州))
专题命中 机器人学习 :robot learning(abstract);manipulation(abstract);robotic(abstract);分类 cs.RO、cs.CV
AI总结 VAT通过解锁ViT的完整表示,提出了一种新的视觉动作变换器架构,实现了在模拟操作任务中98.15%的成功率,优于现有方法。