Interpretability
Interpretability in artificial intelligence concerns what people can learn about a system's behavior, predictions, or internal computations, and whether that understanding is reliable enough for a stated…
Explore AI Ethics through related topics and the articles other pages reference most.
Articles that also belong to these categories. Counts cover all of AI Ethics.
Showing 61-92 of 92 articles
Interpretability in artificial intelligence concerns what people can learn about a system's behavior, predictions, or internal computations, and whether that understanding is reliable enough for a stated…
Isaac Asimov (officially born January 2, 1920, died April 6, 1992) was an American writer and biochemist whose science fiction gave robotics and AI ethics two of their most durable pieces of vocabulary: the…
Joy Buolamwini is a Canadian-American computer scientist and digital activist known for research exposing racial and gender bias in commercial facial recognition and facial-analysis systems
The Leiden Declaration on Artificial Intelligence and Mathematics is a statement on the use of artificial intelligence in mathematical research, published on 2 June 2026 at leidendeclaration.ai and deposited…
MACHIAVELLI is a benchmark for evaluating the ethical behavior of AI agents in text-based interactive environments.
The key machine learning fairness terms are the formal criteria used to define and measure when a model treats demographic groups equitably, together with the named biases that make models unfair.
Machine-generated text detection is the problem of deciding whether a given passage of text was written by a person or produced by a large language model.
Margaret Mitchell is an American computer scientist who works on AI ethics, fairness in machine learning, and the documentation of AI systems.
Max Tegmark is a Swedish-American physicist and artificial intelligence researcher who is a professor of physics at the Massachusetts Institute of Technology (MIT) and the co-founder and president of the…
Meredith Whittaker is an American technologist, researcher, and privacy advocate who serves as president of the Signal Foundation, the nonprofit behind the encrypted messaging app Signal
A model card is a short, standardized document that accompanies a trained machine learning model and reports its intended use, training data, evaluation results across different population subgroups, ethical…
Model welfare is the research area that investigates whether advanced AI systems might have morally relevant experiences or interests, such as suffering or wellbeing, and what (if anything) their developers…
On September 8, 2026 the National Security Agency (NSA), the Cybersecurity and Infrastructure Security Agency (CISA) and the Federal Bureau of Investigation (FBI) released a joint cybersecurity advisory, alert…
Nick Bostrom (born Niklas Boström, 10 March 1973) is a Swedish-born philosopher best known for the 2014 book Superintelligence: Paths, Dangers, Strategies, the 2003 simulation argument, and the…
Out-group homogeneity bias, also called the out-group homogeneity effect, is the cognitive bias in which people perceive members of an out-group as more similar to one another than members of their own…
Predictive parity is a group fairness metric in machine learning that holds when a classifier's positive predictive value (PPV), also called precision
Predictive rate parity (PRP), also called predictive parity, predictive value parity, or the sufficiency criterion, is a group fairness metric in machine learning that requires a classifier's positive…
A proxy for a sensitive attribute is an ordinary input feature that is statistically correlated with a protected characteristic (such as race, gender, age, religion, or disability) and therefore leaks…
Replika is an artificial intelligence companion chatbot developed and operated by Luka Inc., a San Francisco company founded in 2014 by Eugenia Kuyda and Phil Dudchuk.
Reporting bias is a type of data bias in machine learning that occurs when the frequency of events, properties, or outcomes captured in a dataset does not reflect their real-world frequency, because people…
Responsible AI (RAI) is a framework for developing, deploying, and governing artificial intelligence systems in ways that are ethical, transparent, accountable, and aligned with human values.
SAG-AFTRA (the Screen Actors Guild-American Federation of Television and Radio Artists) is an American labor union representing performers in film, scripted television, streaming, video games, commercials, and…
Sampling bias is a systematic error in statistics and machine learning that occurs when a sample is collected so that some members of the intended population have a higher or lower probability of being…
Selection bias is a systematic error that occurs when the data used for analysis, training, or evaluation does not accurately represent the population or domain it is intended to describe
A sensitive attribute (also called a protected attribute or protected characteristic) is any feature in a dataset that corresponds to a legally or ethically protected personal trait, such as race, sex or…
Camera-equipped smart glasses pose a privacy problem that phones do not, because the person being recorded usually cannot tell.
Timnit Gebru is an Ethiopian-born computer scientist and a leading researcher in AI ethics, best known for co-authoring the 2018 "Gender Shades" study on bias in facial recognition, co-leading Google's Ethical…
Toby Ord is an Australian moral philosopher at the University of Oxford who founded the effective-altruism organisation Giving What We Can in 2009 and wrote the 2020 book The Precipice: Existential Risk and…
ToxiGen is a large-scale, machine-generated dataset designed for adversarial and implicit hate speech detection.
Transhumanism is an intellectual and cultural movement that holds that the human condition can and should be fundamentally improved through science and technology, in particular through technologies that…
Unawareness to a sensitive attribute, more commonly called fairness through unawareness (FTU), is a machine learning fairness approach that tries to make a model fair by simply not giving it the sensitive or…
William David MacAskill (born William Crouch; 24 March 1987) is a Scottish moral philosopher, author, and a co-founder of the effective altruism movement, best known as the leading public proponent of…