Can LLMs Do Rocket Science? Exploring the Limits of Complex Reasoning with GTOC 12
大语言模型能做火箭科学吗?通过GTOC 12探索复杂推理的极限
专题命中 复杂问题求解 :reasoning(title,abstract);planning(abstract);分类 cs.AI
AI总结 本研究通过GTOC 12竞赛评估大语言模型在复杂轨道优化任务中的表现,发现其在战略规划上能力显著提升,但在实际执行中存在物理一致性等关键问题。
Comments Extended version of the paper presented at AIAA SciTech 2026 Forum. Includes futher experiments, corrections and new appendix
Journal ref Proceedings of the AIAA SciTech 2026 Forum, January 2026