AI Safety

Explore AI Safety through related topics and the articles other pages reference most.

Explore articles

Reset filters
Browse subtopics: AI Ethics

Articles that also belong to these categories. Counts cover all of AI Safety.

Showing 1-26 of 26 articles

AI Alignment

AI alignment is the study and practice of making artificial intelligence systems behave in ways that accord with intended goals, preferences, constraints, or institutions. The term is used at several levels.

AI EthicsMachine Learning

AI bias

AI bias (also called algorithmic bias) is systematic, repeatable error in artificial intelligence systems that produces unfair, discriminatory, or skewed outcomes, typically disadvantaging groups defined by…

AI EthicsArtificial Intelligence

AI ethics

AI ethics is the field that studies the moral principles, values, and frameworks governing how artificial intelligence systems are designed, built, deployed, and used, and the obligations that developers and…

AI EthicsArtificial Intelligence

AI regulation

AI regulation is the body of laws, binding rules, technical standards, and government enforcement mechanisms that oversee how artificial intelligence systems are built, sold, and used.

AI EthicsArtificial Intelligence

Algorithmic fairness

Algorithmic fairness is the study of how automated decision systems can be made to produce decisions that are equitable across protected attributes such as race, gender, age, religion, and disability.

AI EthicsMachine Learning

Amanda Askell

Amanda Askell is a Scottish philosopher and artificial intelligence researcher who works on fine-tuning and alignment at Anthropic, where she leads the team responsible for the character, persona, and values…

AI EthicsPeople

Autonomous weapons

Autonomous weapons, usually discussed under the label lethal autonomous weapon systems (LAWS), are weapon systems that, once activated, can select and engage targets without further intervention by a human…

AI EthicsAI Policy & Regulation

Confirmation Bias

Confirmation bias is the tendency to search for, interpret, favor, and recall information in ways that confirm one's preexisting beliefs, and in artificial intelligence it appears in three main forms: human…

AI EthicsData Science

Effective Altruism

Effective altruism (often abbreviated EA) is a philosophical and social movement that uses evidence and careful reasoning to identify the most effective ways to benefit others

AI Ethics

Max Tegmark

Max Tegmark is a Swedish-American physicist and artificial intelligence researcher who is a professor of physics at the Massachusetts Institute of Technology (MIT) and the co-founder and president of the…

AI EthicsPeople

Meredith Whittaker

Meredith Whittaker is an American technologist, researcher, and privacy advocate who serves as president of the Signal Foundation, the nonprofit behind the encrypted messaging app Signal

AI EthicsPeople

Model welfare

Model welfare is the research area that investigates whether advanced AI systems might have morally relevant experiences or interests, such as suffering or wellbeing, and what (if anything) their developers…

AI Ethics

Nick Bostrom

Nick Bostrom (born Niklas Boström, 10 March 1973) is a Swedish-born philosopher best known for the 2014 book Superintelligence: Paths, Dangers, Strategies, the 2003 simulation argument, and the…

AI EthicsPeople

Responsible AI

Responsible AI (RAI) is a framework for developing, deploying, and governing artificial intelligence systems in ways that are ethical, transparent, accountable, and aligned with human values.

AI EthicsArtificial Intelligence

Timnit Gebru

Timnit Gebru is an Ethiopian-born computer scientist and a leading researcher in AI ethics, best known for co-authoring the 2018 "Gender Shades" study on bias in facial recognition, co-leading Google's Ethical…

AI EthicsPeople

Toby Ord

Toby Ord is an Australian moral philosopher at the University of Oxford who founded the effective-altruism organisation Giving What We Can in 2009 and wrote the 2020 book The Precipice: Existential Risk and…

AI EthicsPeople

Transhumanism

Transhumanism is an intellectual and cultural movement that holds that the human condition can and should be fundamentally improved through science and technology, in particular through technologies that…

AI EthicsAI History

William MacAskill

William David MacAskill (born William Crouch; 24 March 1987) is a Scottish moral philosopher, author, and a co-founder of the effective altruism movement, best known as the leading public proponent of…

AI EthicsPeople