Black Forest Labs
Black Forest Labs (BFL) is a German-American artificial intelligence company founded in 2024 by Robin Rombach, Andreas Blattmann, Patrick Esser, and Dominik Lorenz
Explore Diffusion Models through related topics and the articles other pages reference most.
Articles that also belong to these categories. Counts cover all of Diffusion Models.
Showing 1-18 of 18 articles
Black Forest Labs (BFL) is a German-American artificial intelligence company founded in 2024 by Robin Rombach, Andreas Blattmann, Patrick Esser, and Dominik Lorenz
DALL-E is a family of text-to-image systems developed by OpenAI that generates images from natural-language descriptions.
A Diffusion Transformer (DiT) is a transformer-based neural network backbone for diffusion models that replaces the U-Net with a Vision Transformer operating on patches of an image latent.
FLUX.1 is a family of text-to-image generation models developed by Black Forest Labs, released on August 1, 2024.
FLUX.2 is the second-generation image generation and editing model family developed by Black Forest Labs, released on November 25, 2025.
Flux is a family of text-to-image generative models developed by Black Forest Labs (BFL), the German-American startup founded by the original creators of Stable Diffusion.
GLIDE (Guided Language to Image Diffusion for Generation and Editing) is a text-conditional diffusion model for text-to-image synthesis and editing released by OpenAI in December 2021.
Imagen is a family of text-to-image diffusion models developed by Google, first introduced in May 2022 and as of 2026 in its fourth generation (Imagen 4).
Imagen 2 is the second generation of Google's text-to-image diffusion model, developed by Google DeepMind and first announced for developers and enterprises on December 13, 2023.
Latent Consistency Models (LCMs) are a family of accelerated text-to-image generative models that apply the consistency-models framework of Song et al.
MMDiT (Multimodal Diffusion Transformer, sometimes written MM-DiT) is a transformer architecture for text-conditioned image generation that gives image tokens and text tokens their own separate weights but…
Midjourney is an artificial intelligence image generation service and independent research lab headquartered in San Francisco, California
Runwayml/stable-diffusion-v1-5 is the Hugging Face repository name of the Stable Diffusion v1.5 checkpoint, a text-to-image latent diffusion model published on October 20
SDXL, short for Stable Diffusion XL, is an open-weights latent text-to-image diffusion model released by Stability AI on 26 July 2023, built around a 2.6 billion parameter U-Net backbone, two text encoders…
Stable Diffusion is a family of generative image models that can synthesize and edit images from text and other conditions.
Stable Diffusion 3 (SD3) is a family of text-to-image diffusion models developed by Stability AI, first announced as an early preview on February 22, 2024, and built on a new architecture called the Multimodal…
Stable Diffusion 3.5 (SD 3.5) is a family of open-weights text-to-image diffusion models released by Stability AI on October 22, 2024, comprising three variants: Stable Diffusion 3.5 Large (8.1 billion…
Würstchen is an efficient three-stage cascaded latent diffusion architecture for text-to-image synthesis introduced by Pablo Pernias, Dominic Rampas, Mats L. Richter, Christopher J. Pal