Observability & Evaluation · Software component
Model Health Prober
Software componentObservability & EvaluationObservability & Evaluationarc:ModelHealthProber
A monitoring component that periodically sends a synthetic inference request to a served model and checks that it is loaded, answers within a timeout and maintains its quality baseline.
Responsibility. Verifies end-to-end model inference health with synthetic requests.
Also known as: Inference health check
Relationships
invokes dependency
is invoked by dependency
triggers dynamic
monitors assurance
Design guidance
- SHOULD run regularly and check model loading, response time and quality baseline, not only process liveness.
Quantitative guidance
As stated by the sources; verify before use.
- Example health request uses a 1000ms timeout (Ref7.16).
- Model quality health passes when > 80% of canned test prompts match expected patterns (Ref8.07).
Classification
- Patterns
- Synthetic probing
- Quality attributes
- Reliability (ISO/IEC 25010 | NIST AI RMF: valid and reliable)
- Risks mitigated
- Model failing to loadSilent inference slowdowns or quality loss
Sources
- Ref7.16: "Production Monitoring and Operations for Agentic AI," unpublished reference note (16-Production-Monitoring-Operations.md), Mastering Agentic AI Systems: Guide for the NVIDIA NCP-AAI Exam supplementary materials, 2026. unpublished note
- Ref8.07: "Agent Health Checks and Diagnostics," unpublished reference note (07-Agent-Health-Checks-Diagnostics.md), Mastering Agentic AI Systems: Guide for the NVIDIA NCP-AAI Exam supplementary materials, 2026. unpublished note