Observability & Evaluation · Software component
Simulated User Agent
Software componentObservability & EvaluationObservability & Evaluationarc:SimulatedUserAgent
An evaluation component that plays the user role in multi-turn benchmark episodes, issuing requests grounded in natural-language scenario instructions to the agent under test.
Responsibility. Drives the agent under test with realistic simulated user turns.
Also known as: User simulator, Simulated user
Relationships
invokes dependency
- Agent Controller abstract Ch3.2 Ch3.3
is invoked by dependency
Classification
- Patterns
- Stateful multi-turn evaluation
- Technologies
- tau-Benchtau-bench
- Quality attributes
- Functional suitability: correctness and validity (ISO/IEC 25010 | NIST AI RMF: valid)
- Risks mitigated
- Single-turn evaluation failing to predict multi-turn behaviour
Sources
- Ch3.2: T. Nguyen, "Compare Agent Performance Across Tasks and Datasets," in Mastering Agentic AI Systems: Guide for the NVIDIA NCP-AAI Exam, 1st ed. 2026, ch. 3.2. ISBN: 9798244538229.
- Ch3.3: T. Nguyen, "Web Navigation and Interaction Benchmarks," in Mastering Agentic AI Systems: Guide for the NVIDIA NCP-AAI Exam, 1st ed. 2026, ch. 3.3. ISBN: 9798244538229.