arXiv · Statistics Machine Learning· Etienne Casanova, Rafal Kocielnik, R. Michael Alvarez·· 16 小时前AI 评分49
LLM 适配性的边界:模型内化先验如何影响标注任务表现
On the Limits of LLM Adaptability: Impact of Model-Internalized Priors on Annotation Task Performance
AI 导读
研究提出 Definition-Specific Familiarity(DSF)指标,衡量模型概念与目标定义的对齐程度,在九个 LLM 与六个毒性数据集上预测标注表现(控制数据集身份后 partial r=+0.41),而三种常见文本记忆指标无正相关。
来源:arXiv · Statistics Machine Learning · arxiv.org