Cognition · Model asset

Tabular Value Function

Model assetCognitionCognition & Memoryarc:TabularValueFunction

A learned policy model stored as a table holding one value estimate per state-action pair (Q-table).

Responsibility. Stores a Q-value for every state-action pair.

Also known as: Q-table

Variant of Learned Decision Policy abstract

When to choose. Choose when the state-action space is small and discrete enough to enumerate and visit.

is target of alternativeTospecializesPolicy Network: is target of alternativeToPolicy NetworkLearned Decision Policy: specializesLearned Decision Policy
Direct neighbourhood (hover for relationship types)

Relationships

alternative to variability

Quantitative guidance

As stated by the sources; verify before use.

Classification

Patterns
Tabular Q-learning

Sources

  1. Ch5.12: T. Nguyen, "Learning-Based Decision Making Fundamentals," in Mastering Agentic AI Systems: Guide for the NVIDIA NCP-AAI Exam, 1st ed. 2026, ch. 5.12. ISBN: 9798244538229.