BBQ (Bias Benchmark for QA)
BBQ (the Bias Benchmark for QA) is a hand-built evaluation dataset that measures whether a question answering (QA) language model relies on social stereotypes when it answers.
Explore Natural Language Processing through related topics and the articles other pages reference most.
Articles that also belong to these categories. Counts cover all of Natural Language Processing.
Showing 1-6 of 6 articles
BBQ (the Bias Benchmark for QA) is a hand-built evaluation dataset that measures whether a question answering (QA) language model relies on social stereotypes when it answers.
BookCorpus (also written BooksCorpus, and sometimes called the Toronto Book Corpus) is a text dataset built from free, self-published English-language ebooks scraped from the distribution platform Smashwords.
Emily M. Bender is an American linguist and a professor in the Department of Linguistics at the University of Washington, where she directs the Computational Linguistics Laboratory.
Machine-generated text detection is the problem of deciding whether a given passage of text was written by a person or produced by a large language model.
Reporting bias is a type of data bias in machine learning that occurs when the frequency of events, properties, or outcomes captured in a dataset does not reflect their real-world frequency, because people…
ToxiGen is a large-scale, machine-generated dataset designed for adversarial and implicit hate speech detection.