从生成到匹配:个性化中文手写体的开发报告
From Generation to Matching: A Development Report on Personalized Chinese Handwriting
浏览论文内容
中文总结 AI 辅助
该项目最初尝试个性化中文手写体生成,后转为逐字符匹配,所建系统可生成接近用户真实手写体的稳定实用类B质量内容。
中文摘要 AI 辅助
本文记录了一个已冻结的个性化中文手写体工程项目。该项目从一名用户的约200张真实手写图像入手,涵盖197个独特汉字,最初被设定为未见字符的少样本生成任务。一系列以规范为中心的个性化路径反复暴露出同一矛盾:结构压力增大使输出更具规范性,而个性化程度提升可能破坏定义身份的笔画。因此项目围绕真实人类字符等价类重置,多书写者CASIA候选池显示,与用户兼容的实现往往已存在于有效人类样本中。任务随之从合成转变为逐字符匹配,再跨书写者组合成虚拟书写者。该冻结系统使用真实墨水特征、字符特定人类人口百分位数、前20候选修剪及贪心最难优先整行选择。在覆盖的目标集上,197个用户字符均有真实人类候选,100字符评估子集的覆盖率为100/100。已知字符保留对比中,有一行经视觉判断几乎与用户真实手写体无法区分。60集稳定性审计将每一集都置于预定义的类A机器代理区域,但这些并非独立的人类类A水平判断。最终证据支持稳定的实用类B质量,根据用户定义的标准,许多输出接近类A水平。该报告记录了为何此案例中生成变得不必要,并未声称可进行无限制或通用的手写体合成。
英文摘要
This paper documents a frozen engineering project on personalized Chinese handwriting. The project started from approximately 200 real handwriting images from one user, covering 197 unique Chinese characters, and was initially formulated as few-shot generation of unseen characters. A sequence of canonical-centered personalization routes repeatedly exposed the same conflict: increasing structural pressure made outputs more canonical, while increasing personalization could damage identity-defining strokes. The project was therefore reset around real-human character equivalence classes. A multi-writer CASIA candidate pool showed that a USER-compatible realization often already existed among valid human samples. The task consequently changed from synthesis to character-wise matching, followed by cross-writer composition into a virtual writer. The frozen system uses real-ink features, character-specific human population percentiles, top-20 candidate pruning, and greedy hardest-first whole-row selection. On the covered target set, all 197 USER characters had real-human candidates, and the 100-character evaluation subset was covered 100/100. Knowncharacter held-out comparisons included a row judged visually almost indistinguishable from genuine USER handwriting. A 60- episode stability audit placed every episode in a predefined A-like machine-proxy region, but these were not independent human A-level judgments. The final evidence supports stable practical B-level quality, with many outputs approaching A-level under the USER-defined criterion. The report records why generation became unnecessary for this case without claiming unrestricted or universal handwriting synthesis.
发表机构
- The Chinese University of Hong Kong, Shenzhen(香港中文大学(深圳))
机构由 AI 辅助整理,请以论文原文为准。