arXiv · Computation and Language· Jiachen Zhao, Zhengxuan Wu, David Bau, Weiyan Shi·· 4 hr agoAI score48
Persona Hierarchy Model:理解 LLM 微调中的上下文泛化
The Persona Hierarchy Model: Understanding Contextual Generalization in Fine-Tuning LLMs
AI brief
论文提出 Persona Hierarchy Model 解释 LLM 微调行为为何有时局限于训练上下文、有时能泛化到未见上下文:共享默认 persona 影响跨上下文行为,修改共享 persona 促进迁移,修改局部 persona 则更局限。
Source: arXiv · Computation and Language · arxiv.org