跳到正文
arXiv cs.CL· Jing Chen, Giulia Loca, Simona Amenta, Marco Marelli·· 10 小时前AI 评分22

伪词探针研究:LLM 缺乏人类伪词处理所依赖的亚词汇敏感性

Pseudowords as probes: Large Language Models show little of the sublexical sensitivity that governs human pseudoword processing

AI 导读

研究用两个意大利语二选一伪词实验测试五款 LLM,并与人类行为基线对比。当选项含真实词提供的词汇熟悉度线索时,LLM 与人类的一致性更高;在纯伪词条件下则明显低于字符 n-gram 模型 fastText。驱动人类与 fastText 一致的亚词汇余弦相似度线索未能稳定迁移到人类与 LLM 的对齐上,推理 token 消耗也与人类处理难度无稳定关系。

来源:arXiv cs.CL · arxiv.org