A social media company wants to use an LLM for content moderation and evaluate outputs for bias or discrimination against groups or individuals with the least administrative effort. Which data source should they use?
Choose an answer
Tap an option to check your answer.
Correct answer: Benchmark datasets.
Why this is the answer
Benchmark datasets are specifically designed to evaluate LLM performance against predefined criteria, including bias and fairness. They contain diverse examples and expected outputs, allowing for systematic and quantifiable assessment with minimal manual effort. User-generated content is too unstructured and voluminous for efficient bias evaluation. Moderation logs record past decisions but don't provide a structured way to test for new or subtle biases in LLM outputs. Content moderation guidelines define rules but are not a data source for evaluating an LLM's adherence to those rules in practice.
Pass your exam — without the endless answer hunt
Get every verified question and explanation for this exam in one place, and save hours of prep. 1,000+ certifications · 20+ languages · free to start.
Pass your exam faster → No card needed