Human Oversight · Data artifact
Absolute Rating Format
Data artifactHuman OversightExperience & Human Oversightarc:AbsoluteRatingFormat
An annotation task format asking annotators to score each response independently on a numeric scale (e.g., 1-10) or with thumbs-up/down signals.
Responsibility. Elicits per-response scalar quality scores.
Also known as: Explicit scalar reward, Likert rating
Variant of Annotation Task Format abstract
When to choose. Used in early RLHF work; generally avoid for preference learning because annotators interpret scales inconsistently across people and over time.
Relationships
alternative to variability
Sources
- Ch10.3: T. Nguyen, "RLHF Methodology," in Mastering Agentic AI Systems: Guide for the NVIDIA NCP-AAI Exam, 1st ed. 2026, ch. 10.3. ISBN: 9798244538229.