Governance & Compliance · Data artifact

Constitution

Data artifactGovernance & ComplianceSafety, Security & GovernanceVariation point (abstract)arc:Constitution

A versioned, human-readable set of explicit natural-language principles (e.g., helpful, harmless, honest; domain rules) that guides model training and runtime behaviour and can be publicly inspected and debated.

Responsibility. States the explicit values an AI system must adhere to.

Also known as: Constitutional principles, Principle set, Explicit value specification, Value documentation, Global value specification, Value specification

configures; is read byconfigures; is read byreceives data from; is audited byconstrainsconfiguresconfiguresconfiguresreceives data fromconfiguresis audited byis specialized byconstrainsconfiguresis specialized byis specialized byis read byAI-Feedback Preference Labeler: configures; is read byAI-Feedback Preference L…Critique-Revision Generator: configures; is read byCritique-Revision Genera…Constitution Authoring Board: receives data from; is audited byConstitution Authoring B…Agent Controller: constrainsAgent ControllerGuardrail Policy: configuresGuardrail PolicySelf-Reflection Critic: configuresSelf-Reflection CriticEscalation Threshold Policy: configuresEscalation Threshold Pol…Decision Stakeholder: receives data fromDecision StakeholderOperational Norm Set: configuresOperational Norm SetExternal Auditor: is audited byExternal AuditorRegional Constitution: is specialized byRegional ConstitutionCross-System Record Conflict Resolver: constrainsCross-System Record Conf…Principle Adherence Test Suite: configuresPrinciple Adherence Test…Unified Constitution: is specialized byUnified ConstitutionUser-Configurable Constitution: is specialized byUser-Configurable Consti…Value Sentinel Agent: is read byValue Sentinel Agent
Direct neighbourhood (hover for relationship types)

Variants

VariantWhen to choose
Regional ConstitutionChoose for globally deployed systems whose jurisdictions differ in law and cultural norms (e.g., political speech, religious expression), accepting higher design, evaluation and maintenance cost.
Unified ConstitutionChoose when one consistent value set can serve all users, explicitly acknowledging limited scope (particular value commitments) rather than claiming universal validity.
User-Configurable ConstitutionChoose when users legitimately differ in how they weight values such as privacy versus convenience, and customization can be bounded by non-negotiable principles.

Relationships

configures structural

is read by dependency

receives data from dynamic

constrains control

is audited by assurance

Design guidance

Classification

Patterns
Constitutional AIExplicit value encodingMulti-source principle derivationDomain-specific principle extensionCore vs. culturally-adapted principlesTop-down value alignment
Quality attributes
Transparency and accountability (NIST AI RMF: accountable and transparent)Maintainability (ISO/IEC 25010)Reliability (ISO/IEC 25010 | NIST AI RMF: valid and reliable)
Risks mitigated
Opaque implicit values learned from preference dataUntraceable alignment failuresUnauthorized policy exceptionsHarmful, biased or deceptive outputsUnstated designer assumptions about values
Frameworks & regulations
EU AI Act: transparency requirementsUN Declaration of Human Rights (principle source)Equal Credit Opportunity ActFair Housing ActGDPRSecurities regulations / fiduciary dutyEU AI Act (high-risk requirements)IEEE Ethically Aligned Design

Sources

  1. Ch9.5: T. Nguyen, "Constitutional AI," in Mastering Agentic AI Systems: Guide for the NVIDIA NCP-AAI Exam, 1st ed. 2026, ch. 9.5. ISBN: 9798244538229.
  2. Ch9.6: T. Nguyen, "Value Alignment Frameworks," in Mastering Agentic AI Systems: Guide for the NVIDIA NCP-AAI Exam, 1st ed. 2026, ch. 9.6. ISBN: 9798244538229.
  3. Ch10.3: T. Nguyen, "RLHF Methodology," in Mastering Agentic AI Systems: Guide for the NVIDIA NCP-AAI Exam, 1st ed. 2026, ch. 10.3. ISBN: 9798244538229.