arXivDaily arXiv每日学术速递 周一至周五更新
arXiv周末暂无论文更新,休息一下吧,周末愉快~~
arXiv 2609.03577cs.CL

语言、语言模型与我们所谈论的内容

Language, Language Models, and What We're Talking About

Malvina Nissim

首次发表
浏览论文内容

中文总结 AI 辅助

该研究以意大利语语言模型为例,探讨语言模型的本质、测试数据的意义等问题,提出需区分技术产品型与研究工具型语言模型,明确其应产出的语言类型。

中文摘要 AI 辅助

语言模型通常被视为技术产物,但显然会受到训练数据所传递的语言世界的塑造。我以意大利语语言模型为例,希望让人们关注以下问题:在对翻译数据、合成数据进行训练和专业化处理并进一步筛选后得到的系统的本质是什么?在同样不自然的数据上对其进行测试又有何意义?这些模型最终是意大利语的模型吗?是语言的模型吗?自然语言处理(NLP)是否仍关注语言本身?这些问题引出了另一个更具体的问题:我们真正希望语言模型产出的是何种语言?我认为,若不先明确区分作为技术产品设计的语言模型与作为研究语言本身的工具设计的语言模型,就无法回答这个问题。之后答案可能是多样的,我们所谈论的语言也可能是多样的,情况或许并不像我们担心的那样悲观。

英文摘要

Language models are commonly discussed as technical artefacts, but they are obviously shaped by the linguistic worlds conveyed by data during their training. Using Italian language models as evidence, I want to bring attention to the nature of the systems which result from training and specialising models on translated and synthetic data, and further curating them, and to the meaning of testing them on equally unnatural data. Are these eventually models of Italian? Are they models of language? Does NLP still care about language? These questions yield another, more concrete question: what language do we actually want language models to produce? I argue that this question cannot be answered if we do not first consider a clearer distinction between language models designed as technical products and language models designed as tools for studying language itself. The answers then might be diverse, the languages we are talking about might be diverse, and the picture might not be as pessimistic as we fear.

发表机构

  • University of Groningen(格罗宁根大学)

机构由 AI 辅助整理,请以论文原文为准。

补充信息

↑