01.AI
01.AI (Chinese: 零一万物, Língyi Wànwù) is a Chinese artificial intelligence company founded in 2023 by Kai-Fu Lee that developed the Yi family of large language models.
Explore language models, how they work, and the techniques used to build applications with them.
Understand language models and a common application architecture. Read in order or jump to a topic.
Ranked by links from other AI Wiki pages.
Articles that also belong to these categories. Counts cover all of Large Language Models.
Showing 1-60 of 545 articles
01.AI (Chinese: 零一万物, Língyi Wànwù) is a Chinese artificial intelligence company founded in 2023 by Kai-Fu Lee that developed the Yi family of large language models.
An AI agent is a software system that selects and performs actions in an environment in pursuit of an objective.
The pace of frontier AI model releases went from a single landmark launch in late 2022 to a new flagship roughly every one to two weeks by 2026.
AI pricing refers to the cost structures and economic models used by providers of artificial intelligence services, particularly large language model (LLM) APIs.
An AI browser agent is a class of AI agent that uses artificial intelligence to autonomously navigate, interpret, and interact with web browsers to complete tasks on behalf of a user.
AI21 Labs is an Israeli artificial intelligence company, founded in 2017 by Yoav Shoham, Ori Goshen, and Amnon Shashua, that develops large language models (LLMs) and AI orchestration systems for enterprise…
Activation-aware Weight Quantization (AWQ) is a post-training quantization method for large language models that compresses weights to 4-bit (and optionally 3-bit) integers while keeping near-FP16 task…
Abliterated Model Large V2 is a hosted, text-only reasoning large language model offered by Abliteration.ai under the API identifier abliterated-model-large-v2.
Activation steering is a family of inference-time techniques in mechanistic interpretability and AI safety that modify a neural network's internal activations to influence its behavior, without retraining the…
Adept AI (also known as Adept AI Labs) is an American artificial intelligence company, founded on January 5, 2022 in San Francisco, that pioneered "action models": AI agents trained to operate existing…
AdvBench (Adversarial Behavior Benchmark) is a red-teaming benchmark dataset for measuring how easily an aligned large language model can be pushed into producing harmful or objectionable content
AgentBench is a multi-dimensional benchmark for evaluating large language models (LLMs) as autonomous agents across eight distinct interactive environments
Agentic Context Engineering (ACE) is a framework for scalable and efficient context adaptation in large language models (LLMs) that lets an AI system improve itself by treating its own context as an evolving
An agentic workflow is a multi-step process in which one or more AI agents independently plan a sequence of actions, select and use tools, evaluate intermediate results, and iterate until a goal is reached.
Aleph Alpha is a German artificial intelligence company, founded in 2019 and headquartered in Heidelberg, Germany, that builds sovereign AI software for European enterprises and governments.
AlpacaEval is an automatic evaluation framework for instruction-following large language models (LLMs) developed by Stanford University's Tatsu Lab
AlphaCode is an artificial intelligence system developed by Google DeepMind that generates computer programs capable of solving competitive programming problems at a human-competitive level
Amazon Nova is a family of foundation models developed by Amazon and offered through Amazon Bedrock, announced on December 3, 2024, at the AWS re:Invent conference in Las Vegas.
Amazon Nova Act is an agentic AI model and developer software development kit (SDK) from Amazon for building AI agents that reliably perform multi-step actions inside a web browser.
Anthropic is an American artificial intelligence (AI) safety and research company founded in 2021 by Dario Amodei, Daniela Amodei, and other former OpenAI researchers, best known for the Claude family of large…
The Anthropic API is the developer interface for Anthropic's Claude family of large language models, a Messages-based HTTP service hosted at for building conversational AI applications, AI agents, and…
Apertus is a fully open, multilingual large language model developed in Switzerland and released on September 2, 2025, by EPFL, ETH Zurich, and the Swiss National Supercomputing Centre (CSCS) under the Swiss…
Apple Foundation Models (AFM) are the family of large language models Apple built to power Apple Intelligence, the company's generative AI system introduced at WWDC 2024 in June 2024.
Apriel is a family of open-weight small language models developed by ServiceNow, the enterprise-software company, through its in-house AI research organization.
Aquila (Chinese: 悟道·天鹰, Wudao Tianying) is a series of open bilingual (Chinese and English) large language models developed by the Beijing Academy of Artificial Intelligence (BAAI).
Artificial Analysis is an independent benchmarking and analytics platform that evaluates artificial intelligence models and API providers across intelligence, speed, price, and latency, and it is best known…
Atlas is a retrieval-augmented language model developed by researchers at Meta AI (the group then known as Facebook AI Research, or FAIR).
Auto-CoT (Automatic Chain of Thought) is an automated prompting method that builds few-shot Chain-of-Thought demonstrations for large language models without any human-written exemplars, introduced by…
An autoregressive model predicts each element of a sequence from the elements that precede it, feeding its own earlier outputs back in as context for every later prediction.
Aya is an open multilingual model family and global research initiative from Cohere Labs (formerly Cohere For AI, or C4AI), the nonprofit research arm of the Canadian artificial intelligence company Cohere.
BABILong is a benchmark for testing how well a large language model can reason over facts scattered through very long text.
BERT, short for Bidirectional Encoder Representations from Transformers, is a pretrained language model introduced by researchers at Google Research in 2018.
BIG-Bench (Beyond the Imitation Game Benchmark) is a large-scale, collaborative benchmark of 204 tasks, contributed by 450 authors across 132 institutions, built to measure and extrapolate the capabilities of…
BLOOM (BigScience Large Open-science Open-access Multilingual Language Model) is a 176-billion-parameter open-access large language model released on July 12, 2022, and was the first language model with more…
A backdoor attack on a large language model (LLM) is an adversarial training-time attack in which an attacker manipulates training data, fine-tuning data, preference labels, or model weights so that the…
Baichuan Intelligence (Chinese: 百川智能; pinyin: Bǎchuān Zhìnéng) is a Chinese artificial intelligence company, founded in Beijing on April 10, 2023, by former Sogou CEO Wang Xiaochuan, that builds large language…
Baidu AI is an umbrella term for the artificial intelligence research, infrastructure, models, cloud services, and applications developed by Baidu.
Baidu ERNIE (Enhanced Representation through Knowledge Integration) is the family of large language and multimodal foundation models built by the Chinese technology company Baidu, spanning the original 2019…
Bard was the conversational AI chatbot developed by Google, launched as an experimental service on February 6, 2023, opened to a public waitlist on March 21, 2023, and rebranded to Gemini on February 8, 2024.
The Berkeley Function Calling Leaderboard (BFCL) is the standard benchmark for measuring how accurately large language models (LLMs) invoke functions, APIs, and tools, created by the Gorilla project at UC…
PyTorch, TensorFlow, JAX, Rust, Core ML, Safetensors, Transformers
As of July 2026, the strongest general reasoning models are Anthropic's Claude Opus 4.8 and Claude Fable 5, OpenAI's GPT-5.5, and Google's Gemini 3.1 Pro and Gemini 3 Pro in its Deep Think mode
As of July 2026, the best local LLM to run offline depends almost entirely on how much memory you can give it, so this guide ranks picks by hardware tier.
As of July 2026, the strongest open-weight large language model overall is GLM-5.2 from Z.ai (Zhipu AI), which tops the Artificial Analysis open-weight Intelligence Index at 51 and leads all open models on…
As of July 2026, the two strongest small language models (open models under 15 billion parameters) are Google DeepMind's Gemma 4 12B and Alibaba's Qwen3.5-9B.
A bidirectional language model is a language model that, when computing a representation for a token, conditions on both the tokens that come before it (the left context) and the tokens that come after it (the…
BigBang-V1 is an open-weights large language model released on August 2, 2026 by The Endless Frontier, a Shanghai-based research team drawn from Shanghai Jiao Tong University's School of Artificial…
BioBERT (Bidirectional Encoder Representations from Transformers for Biomedical Text Mining) is a domain-specific language model that adapts BERT to biomedicine by continuing its pre-training on large…
BioGPT is a domain-specific generative pre-trained Transformer language model for biomedical text generation and mining, developed by Microsoft Research.
BitNet is a family of large language model architectures developed by Microsoft Research Asia that constrain the weights of a transformer to extremely low bit-widths: initially a single bit ({-1, +1}) and…
BitNet b1.58 is a ternary-weight large language model architecture from Microsoft Research in which every weight is constrained to one of three values, -1, 0, or +1
A Browser-use agent is an artificial intelligence software agent that operates a standard web browser through its normal user interface, clicking, typing, scrolling, and navigating
The Byte Latent Transformer (BLT) is a tokenizer-free large language model architecture introduced by researchers at Meta AI's Fundamental AI Research (FAIR) group in December 2024.
Byte-pair encoding (BPE) is a subword tokenization algorithm that splits text into tokens by starting from individual characters or bytes and iteratively merging the most frequent adjacent pair into a new token
ByteDance AI refers to the artificial intelligence research, products, and infrastructure developed by ByteDance, the Chinese technology company best known as the parent of TikTok.
CamemBERT is a French monolingual language model based on the RoBERTa architecture, released in late 2019 by researchers at Inria, Facebook AI Research, and Sorbonne Université.
Chain of Verification (CoVe) is a prompting technique that reduces factual hallucinations in large language models by having the model fact-check its own draft response through a structured four-step…
ChatGLM is a series of open, bilingual (Chinese and English) conversational large language models developed by Zhipu AI together with the Knowledge Engineering Group (KEG) lab at Tsinghua University.
ChatGPT is a conversational artificial intelligence chatbot developed by OpenAI and built on the company's GPT series of large language models.
Comparison date: July 28, 2026. Product access, model routing, prices, and features can change. This article separates the consumer products from the models and APIs available through them.