Depth-Wise Probing and Pruning of the Planning Token in a Driving Vision-Language-Action Model
驾驶视觉-语言-动作模型中规划 token 的深度探测与剪枝
机构 * Robert Bosch GmbH(罗伯特·博世有限公司) ; Karlsruhe Institute of Technology(卡尔斯鲁厄理工学院)
AI总结 本文针对驾驶 VLA 模型的规划 token,通过探测其 32 个解码器层的信号并剪枝部分层,在误差小幅增加下实现 1.33 倍解码器加速,验证了规划信息早期存在但格式适配问题。
Comments Accepted at the 6th DriveX Workshop (Foundation Models for Autonomous Driving), ECCV 2026. 14 pages, 8 figures, 4 tables