Spatial Traces: Enhancing VLA Models with Spatial-Temporal Understanding
机构 * MIPT(莫斯科物理技术学院) ; AIRI(人工智能研究院)
专题命中 VLA模型 :VLA(title,abstract);vision-language-action(abstract);action model(abstract);分类 cs.RO、cs.CV、cs.AI
视觉与机器人
视觉-语言-动作模型、机器人基础模型和语言条件机器人控制。
机构 * MIPT(莫斯科物理技术学院) ; AIRI(人工智能研究院)
专题命中 VLA模型 :VLA(title,abstract);vision-language-action(abstract);action model(abstract);分类 cs.RO、cs.CV、cs.AI
机构 * AI for Science Institute(人工智能科学研究院) ; DP Technology(DP技术) ; College of Chemistry and Molecular Engineering(化学与分子工程学院) ; School of Mathematical Sciences(数学科学学院) ; Center for Machine Learning Research(机器学习研究中心)
专题命中 VLA模型 :action model(title,abstract);分类 cs.LG
机构 * Max Planck Institute for Intelligent Systems(马克斯·普朗克智能系统研究所) ; S-Lab, Nanyang Technological University(南洋理工大学S实验室) ; Stanford University(斯坦福大学)
专题命中 VLA模型 :action model(abstract);分类 cs.CV
专题命中 VLA模型 :action model(abstract)
Comments Presented at the 39th International Cosmic Ray Conference (ICRC2025). 7 pages, 4 figures
专题命中 VLA模型 :VLA(abstract)
Comments 18 pages, 15 figures. Accepted for Publication in A&A
Journal ref A&A 700, A113 (2025)
专题命中 数据集与评测 :action model(abstract)
Comments 37 pages, 8 figures. This is the author's version of the article published in Nature Communications under CC BY 4.0. The final published version is available at https://doi.org/10.1038/s41467-025-61949-x
Journal ref Nat Commun 16, 6749 (2025)