AuraTracer智迹闻
中文

EVENT DOSSIER

How Do Language Models Represent and Use Phonological Information for Allomorph Selection?

2026-09-07 12:00 Science 🔥 42.2 heat score
1sources
1days unfolding
42.2heat score
1mentions
SummaryAI generated

The latest research confirms that despite masking the phonetic structure of words during training, language models can still reliably generate morpheme forms constrained by phonetic conditions. Regarding the selection mechanism for the indefinite articles ‘a’ and ‘an’ in English, the study found that these phonetic conditions are encoded in a single linear direction of the trigger word embedding, and this direction因果ly drives the article selection in the token-level ‘wug’ test. When predicting the position of articles, the model predicts the upcoming trigger words and uses the phonetic features of these predicted words to decide whether to use ‘a’ or ‘an’. These results indicate that language models possess the ability to make cross-linguistic and explicit phonetic judgments based on rules, providing a mechanistic explanation for the selection of morphemes constrained by phonetic conditions, and distinguishing this ability during generation from metalinguistic judgment.

Related eventsRELATED EVENTS
Key entitiesKEY ENTITIES
arXiv

SignalsSIGNALS

Keyword heat
  • arXiv1

All reports (1)SOURCES

A arXiv cs.CL en 2026-09-07 12:00

How Do Language Models Represent and Use Phonological Information for Allomorph Selection?

语言模型在训练时虽掩盖了单词的语音结构,却能可靠地生成受语音条件制约的词素形式。针对英语不定冠词 a/an,研究证实该语音条件被编码在触发词嵌入的单一线性方向中,且该方向因果驱动了 token 级 wug 测试中的文章选择。模型在预测文章位置时,会预报即将出现的触发词,并利用预报词的语音特征来选择文章。这些结果表明,语言模型具备基于规则进行跨语言及显式语音判断的泛化能力,为语音条件制约的词素选择提供了机制性解释,并将生成时的这种能力与元语言判断区分开来。