01.AI
01.AI (Chinese: 零一万物, Língyi Wànwù) is a Chinese artificial intelligence company founded in 2023 by Kai-Fu Lee that developed the Yi family of large language models.
Explore Open Source AI through related topics and the articles other pages reference most.
Articles that also belong to these categories. Counts cover all of Open Source AI.
Showing 1-60 of 165 articles
01.AI (Chinese: 零一万物, Língyi Wànwù) is a Chinese artificial intelligence company founded in 2023 by Kai-Fu Lee that developed the Yi family of large language models.
Apertus is a fully open, multilingual large language model developed in Switzerland and released on September 2, 2025, by EPFL, ETH Zurich, and the Swiss National Supercomputing Centre (CSCS) under the Swiss…
Apriel is a family of open-weight small language models developed by ServiceNow, the enterprise-software company, through its in-house AI research organization.
Aquila (Chinese: 悟道·天鹰, Wudao Tianying) is a series of open bilingual (Chinese and English) large language models developed by the Beijing Academy of Artificial Intelligence (BAAI).
Aya is an open multilingual model family and global research initiative from Cohere Labs (formerly Cohere For AI, or C4AI), the nonprofit research arm of the Canadian artificial intelligence company Cohere.
As of July 2026, the best local LLM to run offline depends almost entirely on how much memory you can give it, so this guide ranks picks by hardware tier.
As of July 2026, the strongest open-weight large language model overall is GLM-5.2 from Z.ai (Zhipu AI), which tops the Artificial Analysis open-weight Intelligence Index at 51 and leads all open models on…
As of July 2026, the two strongest small language models (open models under 15 billion parameters) are Google DeepMind's Gemma 4 12B and Alibaba's Qwen3.5-9B.
BigBang-V1 is an open-weights large language model released on August 2, 2026 by The Endless Frontier, a Shanghai-based research team drawn from Shanghai Jiao Tong University's School of Artificial…
Code Llama is a family of open-weight large language models specialized for code generation and understanding, released by Meta AI on August 24, 2023.
Codestral is a family of code-specialized large language models developed by Mistral AI, beginning with Codestral 22B, released on May 29, 2024
Cohere Command A is a 111 billion parameter dense large language model released by Cohere on March 13, 2025, built for enterprise agents, Retrieval-Augmented Generation, and tool use across 23 languages.
Command R+ is a large language model developed by Cohere, the Toronto- and San Francisco-based enterprise artificial intelligence company, and released on April 4, 2024.
Cosmopedia is an open synthetic pretraining dataset released by Hugging Face in February 2024, made up of textbooks, blog posts, stories, and WikiHow-style articles written entirely by a large language model.
CrewAI is an open-source multi-agent orchestration framework that enables developers to build teams of AI agents that collaborate to accomplish complex tasks.
DBRX is an open-weight mixture of experts large language model developed by Databricks and its Mosaic AI research team, released on March 27, 2024.
DeepSeek is a Chinese artificial intelligence company based in Hangzhou. It was founded in 2023 by Liang Wenfeng, who had previously co-founded the quantitative investment firm High-Flyer.
DeepSeek-V3 is a 671-billion-parameter open-weights Mixture of Experts large language model from Chinese AI lab DeepSeek, released on December 26, 2024, that activates only 37 billion parameters per token and…
DeepSeek V3.1 is a large language model developed by DeepSeek, released on August 19, 2025 and made broadly available via the official API on August 21, 2025.
DeepSeek-V3.2 is an open-weight Mixture of Experts large language model family developed by DeepSeek that introduces DeepSeek Sparse Attention (DSA)
DeepSeek V4 is a family of open-weight Mixture of Experts large language models developed by DeepSeek, a Hangzhou-based AI research lab.
DeepSeek V4-Pro is the flagship model of the DeepSeek V4 family: a Mixture of Experts large language model with 1.6 trillion total parameters, 49 billion of them activated per token, and a one-million-token…
DeepSeek V4.1-Flash is an open-weight multimodal mixture-of-experts model released by DeepSeek on September 10, 2026. It accepts text and images and generates text.
DeepSeek, Llama, and Qwen are families of large language models, not single systems. DeepSeek announced V4 as a preview on April 24, 2026 .
DeepSeek-R1-Distill is a family of six open-weight reasoning language models released by DeepSeek on January 20, 2025, alongside the flagship DeepSeek-R1 reasoning model.
Devstral is a family of open-weight and API large language models specialized for agentic software engineering, developed by Mistral AI in collaboration with All Hands AI
DiffusionGemma is an experimental open-weight large language model developed by Google DeepMind for multimodal text generation through discrete diffusion.
Dolma is an open three-trillion-token English pretraining corpus released by the Allen Institute for AI (AI2) to power its fully open OLMo language models and to let researchers study how training data shapes…
ERNIE 4.5 is a family of large language models released by the Chinese technology company Baidu, open-sourced on June 30, 2025 under the Apache 2.0 license .
ERNIE X1 is a deep-reasoning large language model developed by Baidu, the Chinese search and artificial-intelligence company, as part of its ERNIE (Wenxin) family.
EXAONE (an acronym for EXpert AI for EveryONE) is the family of large language models and foundation models developed by LG AI Research
EleutherAI is a non-profit artificial intelligence research institute that builds and openly releases large language models, datasets, and evaluation tools, and studies their interpretability and alignment.
EmbeddingGemma is an open text embedding model from Google, released in September 2025, that turns text into dense numeric vectors for search, retrieval, classification, and clustering.
Falcon is a family of open-source large language models built by the Technology Innovation Institute (TII)
Falcon 3 is a family of open-weight large language models released on December 17, 2024 by the Technology Innovation Institute (TII), an applied research center based in Abu Dhabi, United Arab Emirates .
Falcon-H1 is a family of open-weight large language models released in 2025 by the Technology Innovation Institute (TII)
GLM-130B is a 130-billion-parameter bilingual (English and Chinese) large language model released in August 2022 by the Knowledge Engineering Group (KEG) and the Data Mining research group at Tsinghua…
GLM-4.5 is an open-weights large language model released by Zhipu AI (operating internationally as Z.ai) on July 28, 2025, built on a 355-billion-parameter Mixture of Experts architecture that activates 32…
GLM-4.6 is a flagship open-weight large language model released by Zhipu AI under its international brand Z.ai on September 30, 2025, built on a sparse Mixture of Experts (MoE) architecture with roughly 357…
GLM-5 is an open-weight flagship large language model released by the Chinese AI company Zhipu AI, under its international brand Z.ai, on February 11, 2026.
GLM-5.1 is an open-weight large language model developed by the Chinese AI company Zhipu AI, which markets its products internationally under the brand Z.ai.
GLM-5.2 is an open-weight large language model developed by the Chinese company Zhipu AI, which sells its products internationally under the brand Z.ai.
GLM-5.3-Flash is an open-weight, natively multimodal mixture-of-experts large language model released by Z.ai on August 26, 2026.
GPT-J (full release name GPT-J-6B) is a 6-billion-parameter autoregressive transformer language model released by the EleutherAI collective on June 9, 2021, and first announced in a blog post dated June 4
Gemma is a family of open-weight models developed by Google DeepMind. The family began in February 2024 with text-only decoder models derived from research used for Google's Gemini systems, then expanded to…
Gemma 2 is a family of open-weights large language models developed by Google DeepMind and released starting June 27, 2024, in three parameter sizes: 2 billion (2B), 9 billion (9B), and 27 billion (27B).
Gemma 3 is a family of open-weight large language models developed by Google DeepMind and released on March 12, 2025.
Gemma 4 is a family of open-weight multimodal models developed by Google DeepMind and released on April 2, 2026.
Goedel-Prover is an open-source large language model designed for automated formal theorem proving in Lean 4.
Hermes 4 is a family of open-weight large language models released by Nous Research in late August 2025.
Hunyuan is Tencent's brand for its foundation models, an umbrella covering large language models, video generation, image synthesis, 3D asset creation, and multimodal reasoning.
Hunyuan-A13B is an open-weight mixture-of-experts large language model released by Tencent in late June 2025.
Hy4 Preview is an open-weight large language model released by the Tencent Hy Team on August 28, 2026.
IBM Granite is a family of open foundation models built by IBM for enterprise use, released under the permissive Apache 2.0 license and delivered through IBM's watsonx platform.
IBM Granite 4.0 is the fourth generation of IBM's Granite family of open-weight enterprise large language models, released on October 2, 2025.
IBM watsonx is an enterprise artificial intelligence and data platform built by IBM and announced on May 9, 2023, at IBM's Think conference by CEO Arvind Krishna.
Inkling is a multimodal large language model released on 2026-07-15 by Thinking Machines Lab, the startup founded by former OpenAI chief technology officer Mira Murati.
Instructor is an open-source Python library that returns type-safe, Pydantic-validated structured outputs from large language model APIs.
InternLM (Chinese name Shusheng Puyu, 书生·浦语) is a family of open-weight large language models developed primarily by Shanghai AI Laboratory, with partners that include SenseTime, the Chinese University of Hong…
Jamba Reasoning 3B is an open-weight small reasoning model released by the Israeli artificial-intelligence company AI21 Labs on October 8, 2025 .