Common Pile
Common Pile v0.1 is an 8 terabyte corpus of openly licensed and public domain text, released on June 5, 2025, by EleutherAI and a consortium of more than two dozen academic and industry collaborators.
Explore Open Source AI through related topics and the articles other pages reference most.
Articles that also belong to these categories. Counts cover all of Open Source AI.
Showing 61-120 of 444 articles
Common Pile v0.1 is an 8 terabyte corpus of openly licensed and public domain text, released on June 5, 2025, by EleutherAI and a consortium of more than two dozen academic and industry collaborators.
Composio is an open-source AI agent tool integration platform that provides infrastructure for connecting large language models (LLMs) and AI agents to over 1,000 external applications and services.
Continue is an open-source AI code assistant that integrates directly into code editors, letting developers connect any large language model (LLM) and customize AI-powered coding features including…
Cosmopedia is an open synthetic pretraining dataset released by Hugging Face in February 2024, made up of textbooks, blog posts, stories, and WikiHow-style articles written entirely by a large language model.
CrewAI is an open-source multi-agent orchestration framework that enables developers to build teams of AI agents that collaborate to accomplish complex tasks.
DBRX is an open-weight mixture of experts large language model developed by Databricks and its Mosaic AI research team, released on March 27, 2024.
DCLM, short for DataComp for Language Models (also styled DataComp-LM), is an open benchmark, dataset, and software framework, released in June 2024
Deep Agents is an open-source agent harness developed by LangChain for long-running, complex, multi-step, non-deterministic tasks such as research, coding, and go-to-market automation.
Deep Cogito is a San Francisco artificial intelligence research lab that develops open-weight large language models under the Cogito name.
DeepEP is an open-source GPU communication library built by DeepSeek for Mixture-of-Experts (MoE) models.
DeepGEMM is an open-source library from DeepSeek that provides fast FP8 general matrix multiplication (GEMM) kernels for NVIDIA Hopper GPUs.
DeepSeek is a Chinese artificial intelligence company based in Hangzhou. It was founded in 2023 by Liang Wenfeng, who had previously co-founded the quantitative investment firm High-Flyer.
DeepSeek Harness is an open-source AI agent harness developed by DeepSeek. Also called dsh, it supplies the software around a language model: the agent loop, tools, sessions, filesystems, permission controls…
DeepSeek-V3 is a 671-billion-parameter open-weights Mixture of Experts large language model from Chinese AI lab DeepSeek, released on December 26, 2024, that activates only 37 billion parameters per token and…
DeepSeek V3.1 is a large language model developed by DeepSeek, released on August 19, 2025 and made broadly available via the official API on August 21, 2025.
DeepSeek-V3.2 is an open-weight Mixture of Experts large language model family developed by DeepSeek that introduces DeepSeek Sparse Attention (DSA)
DeepSeek V4 is a family of open-weight Mixture of Experts large language models developed by DeepSeek, a Hangzhou-based AI research lab.
DeepSeek V4-Pro is the flagship model of the DeepSeek V4 family: a Mixture of Experts large language model with 1.6 trillion total parameters, 49 billion of them activated per token, and a one-million-token…
DeepSeek V4.1-Flash is an open-weight multimodal mixture-of-experts model released by DeepSeek on September 10, 2026. It accepts text and images and generates text.
DeepSeek, Llama, and Qwen are families of large language models, not single systems. DeepSeek announced V4 as a preview on April 24, 2026 .
DeepSeek-OCR is an open-source optical character recognition (OCR) and document-understanding system released by DeepSeek on 20 October 2025 that pioneers a contexts optical compression paradigm: it encodes…
DeepSeek-R1-Distill is a family of six open-weight reasoning language models released by DeepSeek on January 20, 2025, alongside the flagship DeepSeek-R1 reasoning model.
Detectron2 is an open-source software library for object detection and image segmentation, built on PyTorch and developed by Facebook AI Research (FAIR), the research group now part of Meta AI.
Devstral is a family of open-weight and API large language models specialized for agentic software engineering, developed by Mistral AI in collaboration with All Hands AI
DiffusionGemma is an experimental open-weight large language model developed by Google DeepMind for multimodal text generation through discrete diffusion.
Dify (stylized Dify.AI) is an open-source platform for building, deploying, and operating applications and agents powered by large language models.
Dolma is an open three-trillion-token English pretraining corpus released by the Allen Institute for AI (AI2) to power its fully open OLMo language models and to let researchers study how training data shapes…
Donut (Document understanding transformer) is an OCR-free visual document understanding model introduced by researchers at NAVER CLOVA in the paper "OCR-free Document Understanding Transformer," first posted…
ERNIE 4.5 is a family of large language models released by the Chinese technology company Baidu, open-sourced on June 30, 2025 under the Apache 2.0 license .
ERNIE X1 is a deep-reasoning large language model developed by Baidu, the Chinese search and artificial-intelligence company, as part of its ERNIE (Wenxin) family.
ESM3 (Evolutionary Scale Modeling 3) is a frontier multimodal generative language model for biology, released by EvolutionaryScale on June 25, 2024, that was the first model to reason jointly over the…
EXAONE (an acronym for EXpert AI for EveryONE) is the family of large language models and foundation models developed by LG AI Research
EleutherAI is a non-profit artificial intelligence research institute that builds and openly releases large language models, datasets, and evaluation tools, and studies their interpretability and alignment.
EmbeddingGemma is an open text embedding model from Google, released in September 2025, that turns text into dense numeric vectors for search, retrieval, classification, and clustering.
Evo 2 is a genomic foundation model built by the Arc Institute together with NVIDIA, Stanford University, and collaborators, and first released in February 2025.
EvolutionaryScale is an American artificial intelligence company that builds frontier generative models for biology, best known for ESM3, a multimodal generative protein language model that can reason over and…
EXL2 (ExLlamaV2 format) is an open-source, mixed-bit weight-quantization format for compressing large language models so they run fast on a single consumer-class NVIDIA GPU.
Z1T is a family of sparse, transformer-like language models that Extropic designed to run partly on its Z1 probabilistic chip, described in a research post dated September 4, 2026 by Guillaume Verdon…
F5-TTS (short for "A Fairytaler that Fakes Fluent and Faithful Speech with Flow Matching") is an open-source text-to-speech and zero-shot voice cloning model released in October 2024 by researchers from…
FAISS (Facebook AI Similarity Search) is an open-source library from Meta for efficient similarity search and clustering of dense vectors
FLUX.1 is a family of text-to-image generation models developed by Black Forest Labs, released on August 1, 2024.
Fairlearn is an open-source Python toolkit for assessing and improving the fairness of machine-learning models with respect to sensitive attributes such as race, gender, or age.
Falcon is a family of open-source large language models built by the Technology Innovation Institute (TII)
Falcon 3 is a family of open-weight large language models released on December 17, 2024 by the Technology Innovation Institute (TII), an applied research center based in Abu Dhabi, United Arab Emirates .
Falcon-H1 is a family of open-weight large language models released in 2025 by the Technology Innovation Institute (TII)
FineWeb is a large-scale, open pretraining dataset for large language models (LLMs) created by Hugging Face.
FlashInfer is an open-source GPU kernel library and code-generation system for large language model inference.
FlashMLA is an open-source GPU kernel from DeepSeek that accelerates the decoding step of Multi-head Latent Attention (MLA)
Flowise is an open-source, low-code platform for building large language model applications, chatbots, and AI agents by dragging and dropping nodes onto a visual canvas.
Flux is a family of text-to-image generative models developed by Black Forest Labs (BFL), the German-American startup founded by the original creators of Stable Diffusion.
GGML is an open-source tensor library written in pure C that runs machine learning inference efficiently on consumer hardware
GLM-130B is a 130-billion-parameter bilingual (English and Chinese) large language model released in August 2022 by the Knowledge Engineering Group (KEG) and the Data Mining research group at Tsinghua…
GLM-4.5 is an open-weights large language model released by Zhipu AI (operating internationally as Z.ai) on July 28, 2025, built on a 355-billion-parameter Mixture of Experts architecture that activates 32…
GLM-4.6 is a flagship open-weight large language model released by Zhipu AI under its international brand Z.ai on September 30, 2025, built on a sparse Mixture of Experts (MoE) architecture with roughly 357…
GLM-5 is an open-weight flagship large language model released by the Chinese AI company Zhipu AI, under its international brand Z.ai, on February 11, 2026.
GLM-5.1 is an open-weight large language model developed by the Chinese AI company Zhipu AI, which markets its products internationally under the brand Z.ai.
GLM-5.2 is an open-weight large language model developed by the Chinese company Zhipu AI, which sells its products internationally under the brand Z.ai.
GLM-5.3-Flash is an open-weight, natively multimodal mixture-of-experts large language model released by Z.ai on August 26, 2026.
GPT-J (full release name GPT-J-6B) is a 6-billion-parameter autoregressive transformer language model released by the EleutherAI collective on June 9, 2021, and first announced in a blog post dated June 4
GPT4All is an open-source ecosystem from Nomic AI for running large language models locally and privately on consumer laptops and desktops, with no internet connection, GPU, or API key required.