Model Adaptation · Model asset
Reference Policy Model
Model assetModel AdaptationModelsarc:ReferencePolicyModel
A frozen copy of the supervised-fine-tuned model whose token distributions anchor preference optimization, against which divergence of the evolving policy is measured and penalized.
Responsibility. Provides the fixed reference distribution for KL-divergence regularization or DPO's implicit reward.
Also known as: Reference model, Frozen SFT model
Variant of Foundation LLM abstract
Relationships
deployed on structural
is trained by lifecycle
- Fine-Tuning Pipeline abstract Ch10.3
Classification
- Patterns
- KL-divergence regularization
- Risks mitigated
- Reward hackingCatastrophic forgettingLoss of generation coherence
Sources
- Ch10.3: T. Nguyen, "RLHF Methodology," in Mastering Agentic AI Systems: Guide for the NVIDIA NCP-AAI Exam, 1st ed. 2026, ch. 10.3. ISBN: 9798244538229.