Model Adaptation · Software component

Behavior Cloning Trainer

Software componentModel AdaptationModelsarc:BehaviorCloningTrainer

An imitation learner that trains a policy by supervised learning to predict the expert's action for each demonstrated state.

Responsibility. Fits a policy to expert state-action pairs by supervised learning.

Also known as: Behavior cloning

Variant of Imitation Learner abstract

When to choose. Choose when demonstrations comprehensively cover deployment situations, the expert is consistent, and decision sequences are short enough that errors do not compound.

readsalternative toalternative tospecializesAgent Trajectory Dataset: readsAgent Trajectory DatasetDAgger Trainer: alternative toDAgger TrainerInverse RL Reward Learner: alternative toInverse RL Reward LearnerImitation Learner: specializesImitation Learner
Direct neighbourhood (hover for relationship types)

Relationships

reads dependency

alternative to variability

Design guidance

Quantitative guidance

As stated by the sources; verify before use.

Classification

Patterns
Supervised imitation

Sources

  1. Ch5.12: T. Nguyen, "Learning-Based Decision Making Fundamentals," in Mastering Agentic AI Systems: Guide for the NVIDIA NCP-AAI Exam, 1st ed. 2026, ch. 5.12. ISBN: 9798244538229.