Alpaca (model)
Alpaca is an instruction-following language model released on March 13, 2023 by Stanford University's Center for Research on Foundation Models (CRFM)
Explore AI Models through related topics and the articles other pages reference most.
Articles that also belong to these categories. Counts cover all of AI Models.
Showing 1-60 of 120 articles
Alpaca is an instruction-following language model released on March 13, 2023 by Stanford University's Center for Research on Foundation Models (CRFM)
Apertus is a fully open, multilingual large language model developed in Switzerland and released on September 2, 2025, by EPFL, ETH Zurich, and the Swiss National Supercomputing Centre (CSCS) under the Swiss…
Apriel is a family of open-weight small language models developed by ServiceNow, the enterprise-software company, through its in-house AI research organization.
Aya is an open multilingual model family and global research initiative from Cohere Labs (formerly Cohere For AI, or C4AI), the nonprofit research arm of the Canadian artificial intelligence company Cohere.
As of July 2026, the strongest open-weight large language model overall is GLM-5.2 from Z.ai (Zhipu AI), which tops the Artificial Analysis open-weight Intelligence Index at 51 and leads all open models on…
BigBang-V1 is an open-weights large language model released on August 2, 2026 by The Endless Frontier, a Shanghai-based research team drawn from Shanghai Jiao Tong University's School of Artificial…
Boltz is a family of open-source biomolecular structure prediction models developed primarily at MIT's Computer Science and Artificial Intelligence Laboratory (CSAIL) and the MIT Jameel Clinic.
Boltz-2 is an open-source biomolecular foundation model that jointly predicts the 3D structure of biological complexes and the binding affinity between small molecules and proteins.
Codestral is a family of code-specialized large language models developed by Mistral AI, beginning with Codestral 22B, released on May 29, 2024
Cohere Command A is a 111 billion parameter dense large language model released by Cohere on March 13, 2025, built for enterprise agents, Retrieval-Augmented Generation, and tool use across 23 languages.
Cohere Transcribe is a family of automatic speech recognition models developed by Cohere and Cohere Labs.
Command R+ is a large language model developed by Cohere, the Toronto- and San Francisco-based enterprise artificial intelligence company, and released on April 4, 2024.
DeepSeek-V3 is a 671-billion-parameter open-weights Mixture of Experts large language model from Chinese AI lab DeepSeek, released on December 26, 2024, that activates only 37 billion parameters per token and…
DeepSeek V3.1 is a large language model developed by DeepSeek, released on August 19, 2025 and made broadly available via the official API on August 21, 2025.
DeepSeek V4 is a family of open-weight Mixture of Experts large language models developed by DeepSeek, a Hangzhou-based AI research lab.
DeepSeek V4-Pro is the flagship model of the DeepSeek V4 family: a Mixture of Experts large language model with 1.6 trillion total parameters, 49 billion of them activated per token, and a one-million-token…
DeepSeek V4.1-Flash is an open-weight multimodal mixture-of-experts model released by DeepSeek on September 10, 2026. It accepts text and images and generates text.
DeepSeek, Llama, and Qwen are families of large language models, not single systems. DeepSeek announced V4 as a preview on April 24, 2026 .
DeepSeek-R1-Distill is a family of six open-weight reasoning language models released by DeepSeek on January 20, 2025, alongside the flagship DeepSeek-R1 reasoning model.
Devstral is a family of open-weight and API large language models specialized for agentic software engineering, developed by Mistral AI in collaboration with All Hands AI
Donut (Document understanding transformer) is an OCR-free visual document understanding model introduced by researchers at NAVER CLOVA in the paper "OCR-free Document Understanding Transformer," first posted…
ERNIE X1 is a deep-reasoning large language model developed by Baidu, the Chinese search and artificial-intelligence company, as part of its ERNIE (Wenxin) family.
ESM3 (Evolutionary Scale Modeling 3) is a frontier multimodal generative language model for biology, released by EvolutionaryScale on June 25, 2024, that was the first model to reason jointly over the…
Evo 2 is a genomic foundation model built by the Arc Institute together with NVIDIA, Stanford University, and collaborators, and first released in February 2025.
Z1T is a family of sparse, transformer-like language models that Extropic designed to run partly on its Z1 probabilistic chip, described in a research post dated September 4, 2026 by Guillaume Verdon…
F5-TTS (short for "A Fairytaler that Fakes Fluent and Faithful Speech with Flow Matching") is an open-source text-to-speech and zero-shot voice cloning model released in October 2024 by researchers from…
FLUX.1 is a family of text-to-image generation models developed by Black Forest Labs, released on August 1, 2024.
Falcon 3 is a family of open-weight large language models released on December 17, 2024 by the Technology Innovation Institute (TII), an applied research center based in Abu Dhabi, United Arab Emirates .
Falcon-H1 is a family of open-weight large language models released in 2025 by the Technology Innovation Institute (TII)
GLM-4.5 is an open-weights large language model released by Zhipu AI (operating internationally as Z.ai) on July 28, 2025, built on a 355-billion-parameter Mixture of Experts architecture that activates 32…
GLM-4.6 is a flagship open-weight large language model released by Zhipu AI under its international brand Z.ai on September 30, 2025, built on a sparse Mixture of Experts (MoE) architecture with roughly 357…
GLM-5 is an open-weight flagship large language model released by the Chinese AI company Zhipu AI, under its international brand Z.ai, on February 11, 2026.
GLM-5.3-Flash is an open-weight, natively multimodal mixture-of-experts large language model released by Z.ai on August 26, 2026.
Gemma 2 is a family of open-weights large language models developed by Google DeepMind and released starting June 27, 2024, in three parameter sizes: 2 billion (2B), 9 billion (9B), and 27 billion (27B).
Gemma 3 is a family of open-weight large language models developed by Google DeepMind and released on March 12, 2025.
Hermes 4 is a family of open-weight large language models released by Nous Research in late August 2025.
Hunyuan is Tencent's brand for its foundation models, an umbrella covering large language models, video generation, image synthesis, 3D asset creation, and multimodal reasoning.
Hunyuan 3D is a family of open weight generative artificial intelligence models from Tencent that turn text prompts, single images, sketches, and other inputs into ready to use three dimensional assets…
Hy4 Preview is an open-weight large language model released by the Tencent Hy Team on August 28, 2026.
IBM Granite is a family of open foundation models built by IBM for enterprise use, released under the permissive Apache 2.0 license and delivered through IBM's watsonx platform.
IBM Granite 4.0 is the fourth generation of IBM's Granite family of open-weight enterprise large language models, released on October 2, 2025.
Inkling is a multimodal large language model released on 2026-07-15 by Thinking Machines Lab, the startup founded by former OpenAI chief technology officer Mira Murati.
Jamba Reasoning 3B is an open-weight small reasoning model released by the Israeli artificial-intelligence company AI21 Labs on October 8, 2025 .
Jamba2 is the second generation of hybrid State Space Model and Transformer language models released by AI21 Labs on January 8, 2026.
Jet-Nemotron is a family of small hybrid-architecture language models released by NVIDIA Research in August 2025.
Jina Embeddings v3 is a multilingual text embedding model released by Jina AI on September 18, 2024, with 570 million parameters, support for 89 languages, an 8,192 token context window, and a stack of…
Kimi K2 is an open-weights Mixture of Experts language model from Moonshot AI, a Beijing startup, released on July 11, 2025 with 1.04 trillion total parameters and 32.6 billion activated per token
Kimi K2.5 is an open-weights, natively multimodal large language model developed by Moonshot AI and released on January 27, 2026 .
Kimi K2.7-Code is an open-weight, coding-focused agentic model released by Moonshot AI on 2026-06-12.
Kimi K3 is an open-weight multimodal reasoning model developed by Moonshot AI. Moonshot made the model available through its hosted products on July 16, 2026 and published the weights, code, configuration…
Kimi Linear is a hybrid linear attention architecture published by Moonshot AI on October 30, 2025, together with a 48-billion-parameter mixture-of-experts model that activates 3 billion parameters per token.
As of July 2026, OpenAI has never officially disclosed how many parameters GPT-5 has, and no GPT-5.x version ships a public size
Leanstral 1.5 is an open-weight model from Mistral AI for formal proof engineering in Lean 4.
Ling-3.0-flash is an open-weight mixture-of-experts language model from inclusionAI, the open-source AI initiative of Ant Group, and the first model of the Ling 3.0 generation.
LingBot-VLA 2.0 is a 6-billion-parameter vision-language-action model developed by Robbyant, the embodied-intelligence unit of Ant Group.
Llama 3 is a family of open-weight large language models developed by Meta. Meta released the original Llama 3 checkpoints on April 18, 2024, in 8-billion-parameter and 70-billion-parameter sizes.
Llama 3.1 is a family of open-weight large language models released by Meta on July 23, 2024, in three sizes, 8 billion, 70 billion, and 405 billion parameters, each shipped in both a pre-trained base form and…
Llama 3.2 is a family of four open-weight large language models released by Meta on September 25, 2024, comprising lightweight 1 billion and 3 billion parameter text-only models for on-device AI and the 11…
Llama 3.3 is an instruction-tuned, text-only large language model with 70 billion parameters that Meta released on December 6, 2024
Llama 4 Behemoth is the announced but never publicly released flagship model in the Llama 4 family from Meta AI.