AI Safety Institutes
AI Safety Institutes are government-established organizations that test, evaluate, and research the safety and security risks of advanced artificial intelligence systems
Explore AI Safety through related topics and the articles other pages reference most.
Articles that also belong to these categories. Counts cover all of AI Safety.
Showing 1-31 of 31 articles
AI Safety Institutes are government-established organizations that test, evaluate, and research the safety and security risks of advanced artificial intelligence systems
The AI Seoul Summit was an international meeting on artificial intelligence governance held on 21-22 May 2024 and co-hosted by the Republic of Korea and the United Kingdom.
AI governance is the collection of frameworks, norms, standards, policies, and institutional arrangements that guide the development, deployment, and use of artificial intelligence systems so that they are…
AI hallucinations in court filings are fabricated legal citations, quotations, case names, and facts that generative AI tools invent and that lawyers or self-represented litigants then submit to a court as if…
AI watermarking is a family of techniques for embedding an imperceptible, machine-detectable signal in content produced by generative artificial intelligence systems so that the content can later be identified…
Autonomous weapons, usually discussed under the label lethal autonomous weapon systems (LAWS), are weapon systems that, once activated, can select and engage targets without further intervention by a human…
The Bletchley Declaration is an international political statement on the safety of frontier AI systems, signed on 1 November 2023 by 28 countries and the European Union at the AI Safety Summit hosted by the…
California Senate Bill 53, formally titled the Transparency in Frontier Artificial Intelligence Act (TFAIA), is a 2025 California state law that requires the largest developers of frontier artificial…
Capability overhang is a term used in ai safety and AI policy discourse to describe a situation in which the latent capabilities of a deployed AI system
Compute governance is a policy framework that uses regulation of the computational resources used to train and run artificial intelligence systems as a primary lever for governing advanced AI.
The EU Action Plan on Cybersecurity and Artificial Intelligence is a European Commission Communication, published as COM(2026) 577 final and adopted in Strasbourg on 7 July 2026
Executive Order 14110, formally titled "Safe, Secure, and Trustworthy Development and Use of Artificial Intelligence," was signed by President Joe Biden on October 30, 2023 and published in the Federal…
Frontier models are artificial intelligence models at or near a selected boundary of capability, scale, or risk.
The Frontier Security Institute (FSI) is a Washington, D.C. organization launched by the Center for AI Safety (CAIS) to connect frontier artificial intelligence developers with the United States national…
The General-Purpose AI Code of Practice (abbreviated GPAI Code of Practice or CoP) is a voluntary compliance framework published by the European Commission through its AI Office on 10 July 2025 to help…
Garcia v. Character Technologies is a wrongful death and product liability lawsuit filed in late October 2024 in the United States District Court for the Middle District of Florida (Orlando Division)
The Grok child safety controversy was a multinational regulatory and legal crisis that began in late December 2025, when xAI's Grok chatbot and its image and video generator, Grok Imagine
Human-in-the-loop (HITL) describes any arrangement in which a person is a required participant in an automated system's operating cycle rather than a bystander to it.
The International AI Safety Report is the first independent, government-mandated scientific assessment of the capabilities, risks, and safety of general-purpose artificial intelligence, written by an…
Joseph Robinette Biden Jr. is an American politician who served as the 46th President of the United States, sworn in on January 20, 2021 and leaving office on January 20, 2025.
Miles Brundage is an American AI policy and governance researcher best known for leading policy research at OpenAI and serving as the company's senior advisor for AGI Readiness before leaving in October 2024.
NIST ARIA (Assessing Risks and Impacts of AI) is a testing, evaluation, validation, and verification (TEVV) program operated by the United States National Institute of Standards and Technology (NIST) to…
The Open Secure AI Alliance is an industry coalition announced on July 27, 2026 to develop and share open technologies, techniques, and tools for securing software and AI agents.
A Responsible Scaling Policy (RSP) is a self-imposed governance framework in which a frontier AI developer commits in advance to safety practices, capability evaluations, deployment restrictions, and security…
Robot safety is the discipline concerned with minimizing the risk of physical harm, property damage, and other hazards arising from the operation of robotic systems.
SB 1047, officially the Safe and Secure Innovation for Frontier Artificial Intelligence Models Act
The Seoul Declaration is the short name for the Seoul Declaration for Safe, Innovative and Inclusive AI, a non-binding international statement adopted on 21 May 2024 by 10 countries plus the European Union at…
Situational Awareness: The Decade Ahead is a 165-page essay series published on June 4, 2024 by Leopold Aschenbrenner, a former member of OpenAI's Superalignment team.
The Anthropic Institute is a research organization within Anthropic dedicated to studying the societal challenges posed by increasingly powerful artificial intelligence systems.
The UK AI Security Institute (AISI) is a research organization within the United Kingdom's Department for Science, Innovation and Technology (DSIT) that conducts pre-deployment evaluations of frontier AI…
The US AI Safety Institute (USAISI or US AISI), since June 2025 the Center for AI Standards and Innovation (CAISI), is a research and evaluation body housed within the National Institute of Standards and…