Model Adaptation · Data artifact

Preference Dataset

Data artifactModel AdaptationModelsarc:PreferenceDataset

A dataset of prompts with pairs of candidate responses labeled by which is preferred, per criterion, used to train reward models or directly optimize policies.

Responsibility. Records comparative human (or model) preference judgments for training.

Also known as: Preference comparisons, Pairwise preference data, AI-generated preference data, RLAIF preference labels, Preference comparison dataset

is written byis written byis evaluated byreceives data fromreceives data fromis written byis written byis written byis read byis read byis read byis written byLLM Judge: is written byLLM JudgeFeedback Collector: is written byFeedback CollectorDomain Expert Annotator: is evaluated byDomain Expert AnnotatorRed Team Tester: receives data fromRed Team TesterPreference Annotator: receives data fromPreference AnnotatorAI-Feedback Preference Labeler: is written byAI-Feedback Preference L…Preference Annotation Console: is written byPreference Annotation Co…Trajectory Harvester: is written byTrajectory HarvesterReward Model Trainer: is read byReward Model TrainerAnnotation Quality Monitor: is read byAnnotation Quality MonitorDirect Preference Optimizer: is read byDirect Preference Optimi…Preference Agreement Filter: is written byPreference Agreement Fil…
Direct neighbourhood (hover for relationship types)

Relationships

is read by dependency

is written by dependency

receives data from dynamic

is evaluated by assurance

Design guidance

Quantitative guidance

As stated by the sources; verify before use.

Classification

Quality attributes
Functional suitability: correctness and validity (ISO/IEC 25010 | NIST AI RMF: valid)

Sources

  1. Ch3.5: T. Nguyen, "Prompt Optimization, Few-Shot Learning, Fine-Tuning," in Mastering Agentic AI Systems: Guide for the NVIDIA NCP-AAI Exam, 1st ed. 2026, ch. 3.5. ISBN: 9798244538229.
  2. Ch9.5: T. Nguyen, "Constitutional AI," in Mastering Agentic AI Systems: Guide for the NVIDIA NCP-AAI Exam, 1st ed. 2026, ch. 9.5. ISBN: 9798244538229.
  3. Ch10.3: T. Nguyen, "RLHF Methodology," in Mastering Agentic AI Systems: Guide for the NVIDIA NCP-AAI Exam, 1st ed. 2026, ch. 10.3. ISBN: 9798244538229.
  4. Ch10.5: T. Nguyen, "Human-over-the-Loop," in Mastering Agentic AI Systems: Guide for the NVIDIA NCP-AAI Exam, 1st ed. 2026, ch. 10.5. ISBN: 9798244538229.