Skip to content
arXiv · Computation and Language· Jiachen Zhao, Zhengxuan Wu, David Bau, Weiyan Shi·· 4 hr agoAI score48

Persona Hierarchy Model:理解 LLM 微调中的上下文泛化

The Persona Hierarchy Model: Understanding Contextual Generalization in Fine-Tuning LLMs

AI brief

论文提出 Persona Hierarchy Model 解释 LLM 微调行为为何有时局限于训练上下文、有时能泛化到未见上下文:共享默认 persona 影响跨上下文行为,修改共享 persona 促进迁移,修改局部 persona 则更局限。

Source: arXiv · Computation and Language · arxiv.org