Observability & Evaluation · Software component

Shadow Test Runner

Software componentObservability & EvaluationObservability & Evaluationarc:ShadowTestRunner

A version experimenter that runs a new agent version silently alongside production, logging what it would have done without executing those actions.

Responsibility. Executes candidate versions side-effect-free on mirrored production inputs.

Also known as: Shadow testing, Shadow deployment

Variant of Agent Version Experimenter abstract

When to choose. Choose when a new version must be assessed on production inputs before committing to deployment, without executing its actions.

invokeswritesis target of alternativeTospecializesAgent Controller: invokesAgent ControllerTrace Store: writesTrace StoreA/B Test Traffic Splitter: is target of alternativeToA/B Test Traffic SplitterAgent Version Experimenter: specializesAgent Version Experimenter
Direct neighbourhood (hover for relationship types)

Relationships

invokes dependency

writes dependency

alternative to variability

Design guidance

Classification

Patterns
Shadow mode
Quality attributes
Safety (ISO/IEC 25010 | NIST AI RMF: safe)
Risks mitigated
Side effects from untested versions

Sources

  1. Ch3.3: T. Nguyen, "Web Navigation and Interaction Benchmarks," in Mastering Agentic AI Systems: Guide for the NVIDIA NCP-AAI Exam, 1st ed. 2026, ch. 3.3. ISBN: 9798244538229.