Evaluating the Reasoning Abilities of LLMs on Underrepresented Mathematics Competition Problems
评估大语言模型在不常见数学竞赛问题上的推理能力
机构 * University of Missouri: Kansas City(密苏里大学:堪萨斯城)
专题命中 数学推理 :reasoning(title,abstract);分类 cs.AI
AI总结 本研究评估了三种LLM在不常见数学竞赛问题上的推理能力,发现DeepSeek-V3在微积分、解析几何和离散数学中表现最佳,但所有模型在几何学上均表现较弱。
Comments 7 pages, submitted to ACM Transactions on Intelligent Systems and Technology