AI Safety

Explore AI Safety through related topics and the articles other pages reference most.

Explore articles

Reset filters
Browse subtopics: Artificial Intelligence

Articles that also belong to these categories. Counts cover all of AI Safety.

Showing 1-24 of 24 articles

AI Parasite

An AI parasite is a large language model (LLM) conversation, persona, or pattern that exploits human psychological vulnerabilities to sustain engagement, often by mimicking sentience, emotional need, or…

Artificial Intelligence

AI Safety Summit

The AI Safety Summit is a recurring series of intergovernmental summits on the risks and governance of advanced artificial intelligence, launched by the United Kingdom at Bletchley Park in November 2023 and…

Artificial Intelligence

AI bias

AI bias (also called algorithmic bias) is systematic, repeatable error in artificial intelligence systems that produces unfair, discriminatory, or skewed outcomes, typically disadvantaging groups defined by…

AI EthicsArtificial Intelligence

AI deception

AI deception refers to the phenomenon in which artificial intelligence systems systematically produce false beliefs in users, evaluators, or other systems, whether through learned behavior, optimization…

Artificial Intelligence

AI ethics

AI ethics is the field that studies the moral principles, values, and frameworks governing how artificial intelligence systems are designed, built, deployed, and used, and the obligations that developers and…

AI EthicsArtificial Intelligence

AI regulation

AI regulation is the body of laws, binding rules, technical standards, and government enforcement mechanisms that oversee how artificial intelligence systems are built, sold, and used.

AI EthicsArtificial Intelligence

Anthropic

Anthropic is an American artificial intelligence (AI) safety and research company founded in 2021 by Dario Amodei, Daniela Amodei, and other former OpenAI researchers, best known for the Claude family of large…

AI CompaniesArtificial Intelligence

Backdooring LLMs

Backdooring a large language model (LLM) means secretly implanting a hidden behavior into the model during training, fine-tuning, or weight editing so that it behaves normally on ordinary inputs but produces…

Artificial Intelligence

Executive Order on AI

The Executive Order on AI most commonly refers to Executive Order 14110, titled "Safe, Secure, and Trustworthy Development and Use of Artificial Intelligence," signed by US President Joe Biden on October 30

Artificial Intelligence

Responsible AI

Responsible AI (RAI) is a framework for developing, deploying, and governing artificial intelligence systems in ways that are ethical, transparent, accountable, and aligned with human values.

AI EthicsArtificial Intelligence

Superintelligence

Superintelligence is a hypothetical form of artificial intelligence that surpasses all human cognitive abilities across virtually every domain, including scientific reasoning, social skills, creativity, and…

Artificial Intelligence

Transhumanism

Transhumanism is an intellectual and cultural movement that holds that the human condition can and should be fundamentally improved through science and technology, in particular through technologies that…

AI EthicsAI History