Safety & Security · Model asset
Bias Classifier Model
Model assetSafety & SecuritySafety, Security & Governancearc:BiasClassifierModel
Classifier weights trained on thousands of labeled examples to distinguish biased (gender, racial, cultural stereotyping) from neutral text.
Responsibility. Scores the likelihood that a text contains biased language.
Relationships
deployed on structural
Classification
- Technologies
- Weave BiasScorerLlamaGuard
Sources
- Ch9.4: T. Nguyen, "Fairness and Bias Mitigation," in Mastering Agentic AI Systems: Guide for the NVIDIA NCP-AAI Exam, 1st ed. 2026, ch. 9.4. ISBN: 9798244538229.