跳到正文
原文
arXiv · Computation and Language· Jiachen Zhao, Zhengxuan Wu, David Bau, Weiyan Shi·· 3 小时前AI 评分48

Persona Hierarchy Model:理解 LLM 微调中的上下文泛化

The Persona Hierarchy Model: Understanding Contextual Generalization in Fine-Tuning LLMs

AI 导读

论文提出 Persona Hierarchy Model 解释 LLM 微调行为为何有时局限于训练上下文、有时能泛化到未见上下文:共享默认 persona 影响跨上下文行为,修改共享 persona 促进迁移,修改局部 persona 则更局限。

来源:arXiv · Computation and Language · arxiv.org