OpenAI

Explore OpenAI through related topics and the articles other pages reference most.

Explore articles

Reset filters
Browse subtopics: AI Alignment

Articles that also belong to these categories. Counts cover all of OpenAI.

Showing 1-4 of 4 articles

InstructGPT

InstructGPT is a family of language models released by OpenAI in January 2022 that take the base GPT-3 and fine-tune it to follow user instructions more helpfully, truthfully, and with less toxic output, using…

AI AlignmentLarge Language Models

Model Spec

The Model Spec is a public document published by openai that defines the intended behavior of the company's language models: how they should follow instructions, when they should refuse a request, how to…

AI AlignmentAI Safety

Rule-Based Rewards (RBR)

Rule-Based Rewards (RBR) is a safety-alignment technique introduced by OpenAI in July 2024 that replaces large quantities of human-labeled safety preference data with an explicit collection of natural-language…

AI AlignmentAI Safety