跳到正文
原文
arXiv · Computation and Language· Yusheng Zhou, Eleanor Lin, David Jurgens·· 6 小时前AI 评分54

Lost in Translation:测量非母语英语对大语言模型终端用户表现的影响

Lost in Translation: Measuring the Effect of Non-Native English on End User Performance of Large Language Models

AI 导读

论文提出FABLE数据集,包含190,911个英语提示变体,源自174K条真实用户写作类任务提示。评估34个开源权重大语言模型后发现,模型不会传播拼写等表面错误,但会镜像用户提示中的高阶修辞与词汇特征;提示词流利度越低,模型回复质量与流利度均显著下降。

来源:arXiv · Computation and Language · arxiv.org