BBQ (Bias Benchmark for QA)
BBQ (the Bias Benchmark for QA) is a hand-built evaluation dataset that measures whether a question answering (QA) language model relies on social stereotypes when it answers.
Explore AI Safety through related topics and the articles other pages reference most.
Articles that also belong to these categories. Counts cover all of AI Safety.
Showing 1-6 of 6 articles
BBQ (the Bias Benchmark for QA) is a hand-built evaluation dataset that measures whether a question answering (QA) language model relies on social stereotypes when it answers.
Grounding in artificial intelligence is the process of anchoring an AI system's outputs to verifiable
Hallucination in generative AI is the production of content that is unsupported, contradicted by an applicable source, factually wrong, internally inconsistent, or otherwise presented without an adequate basis.
SimpleQA is a factuality benchmark released by OpenAI on October 30, 2024 that measures whether large language models can answer short, fact-seeking questions correctly instead of producing hallucinations.
ToxiGen is a large-scale, machine-generated dataset designed for adversarial and implicit hate speech detection.
TruthfulQA is a benchmark designed to measure whether large language models (LLMs) generate truthful answers to questions.