arXivDaily arXiv每日学术速递 周一至周五更新
arXiv周末暂无论文更新,休息一下吧,周末愉快~~
arXiv 2608.21106cs.CYcs.AI

原子学习模型(ALM):一间真实课堂如何被标记化

Atom Learning Model (ALM): how a real classroom got tokenised

发表机构伦敦帝国学院 · 计算机系
查看机构详情
  • Imperial College London(伦敦帝国学院)
  • Department of Computing(计算机系)

机构由 AI 辅助整理,请以论文原文为准。

Philipp Bogdan

首次发表
浏览论文内容

中文总结 AI 辅助

本研究提出原子学习模型(ALM),将中学数学教材标记化为原子与前置条件链接,构建系统生成题目供学生使用,发现多项与预期不符的结果,核心前提未被证伪。

中文摘要 AI 辅助

原子学习模型(ALM)将学校课程标记化。机器读取两本中学数学教材,得到1934个原子,每个原子代表学习者可在单一步骤完成的一项任务,这些原子由4616条机器生成的前置条件链接排序。课程的正反两面均以该结构表达:一道题目是一组原子及其下属所有内容;学生的能力是在同一图谱上每个原子对应的0到1之间的分数;题目是否适合学生是基于一个索引的算术运算,无需为双方拟合难度参数。无人编写过原子、链接或题目。读取757页内容花费55英镑,构建整个结构花费615至1230英镑;系统在七周内为英国两所中学的373名学生生成了6648道题目,每道生成题成本为26便士。有四项测量结果与预期不符:成本产生于链接而非页面;生成器自身的难度标签与实测难度的秩相关系数为-0.0123,即看到题目的语言模型无法判断其难度;当标记耗时为7秒而非3秒时,学生便会停止作答;部署阶段从未使用超过两个前置条件步骤的题目,而这正是核心前提可检验的位置,导致该前提未被证伪也未被证实。

英文摘要

The Atom Learning Model (ALM) tokenises a school curriculum. 757 pages of GCSE and Further Mathematics material were read by machine into 1,934 atoms, each one thing a learner can do in a single step, ordered by 4,616 machine-written prerequisite links. Both sides of a lesson are then expressed in that one structure: a question is a set of atoms plus everything beneath them, a child's ability is a score between 0 and 1 on every atom of the same graph, and whether a question suits a child is arithmetic over one index, with no difficulty parameter fitted for either side. Nobody wrote an atom, a link or a question. Reading the 757 pages cost £55, building the whole structure cost between £615 and £1,230, and against it the system composed 6,648 questions for 373 children in two English secondary schools over seven weeks, at 26p per composed question. Four measurements went against expectation. The cost is in the links, not the pages. The composer's own difficulty label has a rank correlation of -0.0123 with measured facility, so a language model shown a question cannot say how hard it is. Children stop working when a mark takes seven seconds instead of three. And the deployment never served a question deeper than two prerequisite steps, which is exactly where the central premise becomes testable, leaving it unfalsified rather than confirmed.

补充信息

↑