GeoVLA: Empowering 3D Representations in Vision-Language-Action Models
机构 * Tianjin University(天津大学) ; Dexmal ; Tsinghua University(清华大学)
专题命中 VLA模型 :vision-language-action(title,abstract);action model(title);VLA(abstract);分类 cs.RO
Comments The project is visible at https://linsun449.github.io/GeoVLA/