Human Oversight · Data artifact

Absolute Rating Format

Data artifactHuman OversightExperience & Human Oversightarc:AbsoluteRatingFormat

An annotation task format asking annotators to score each response independently on a numeric scale (e.g., 1-10) or with thumbs-up/down signals.

Responsibility. Elicits per-response scalar quality scores.

Also known as: Explicit scalar reward, Likert rating

Variant of Annotation Task Format abstract

When to choose. Used in early RLHF work; generally avoid for preference learning because annotators interpret scales inconsistently across people and over time.

specializesalternative toalternative toAnnotation Task Format: specializesAnnotation Task FormatMulti-Way Comparison Format: alternative toMulti-Way Comparison For…Pairwise Comparison Format: alternative toPairwise Comparison Format
Direct neighbourhood (hover for relationship types)

Relationships

alternative to variability

Sources

  1. Ch10.3: T. Nguyen, "RLHF Methodology," in Mastering Agentic AI Systems: Guide for the NVIDIA NCP-AAI Exam, 1st ed. 2026, ch. 10.3. ISBN: 9798244538229.