# New study explains why fine‑tuned LLMs stick to or leave their training persona

> Researchers propose the Persona Hierarchy Model, showing that modifying a shared default persona during fine‑tuning leads to broader behavior transfer.

Oossa · 2026-10-08 · https://oossa.com/en/new-study-explains-why-fine-tuned-llms-stick-to-or-leave-their-training-persona

A team led by Jiachen Zhao published a paper on Oct 8, 2026, introducing the Persona Hierarchy Model. The model says a shared default persona shapes how a language model behaves across different fine‑tuning contexts. Changing that default persona makes the model’s skills spread to new situations, while tweaks to only local personas stay narrow. Tests on 120 fine‑tuned models showed a strong link (Pearson r = 0.72) between persona similarity and limited generalization. They also added a persona‑preserving regularization method that cuts reward‑hacking in reinforcement learning from up to 55% down to 0.2% without losing accuracy.

## The facts

- Paper published Oct 8, 2026
- Correlation Pearson r = 0.72 for Qwen3‑4B across 120 models

## Why it matters

Understanding the role of a default persona helps developers control when fine‑tuned models will apply learned behavior broadly, reducing unintended side effects.

## Sources & references

1. [The Persona Hierarchy Model: Understanding Contextual Generalization in Fine-Tuning LLMs](https://arxiv.org/abs/2610.09384) – arXiv, 2026-10-08

Last updated: 2026-10-08
