AI art
AI art is artwork created with the assistance of artificial intelligence systems, spanning visual images, music, video, and other creative outputs generated or co-created by algorithms, neural networks, and…
Explore Image Generation through related topics and the articles other pages reference most.
Articles that also belong to these categories. Counts cover all of Image Generation.
Showing 1-51 of 51 articles
AI art is artwork created with the assistance of artificial intelligence systems, spanning visual images, music, video, and other creative outputs generated or co-created by algorithms, neural networks, and…
AUTOMATIC1111 Stable Diffusion Web UI (commonly called A1111, SD WebUI, or simply Automatic1111) is the open-source
Adobe Firefly is a family of creative generative AI models developed by Adobe and integrated across its product ecosystem.
AI art is visual, audio, or other creative work produced with the help of artificial intelligence, most commonly images generated from text prompts by text-to-image models such as DALL-E, Midjourney, and…
As of July 2026, the best all-around AI image generator is GPT Image 2 from OpenAI.
Black Forest Labs (BFL) is a German-American artificial intelligence company founded in 2024 by Robin Rombach, Andreas Blattmann, Patrick Esser, and Dominik Lorenz
ComfyUI is a free, open-source node-based graphical user interface and backend for generative AI workflows, primarily focused on image generation using diffusion models such as Stable Diffusion and FLUX.
ControlNet is a neural network architecture that adds spatial and structural control to large pretrained text-to-image diffusion models.
CycleGAN (Cycle-Consistent Generative Adversarial Network) is a deep learning architecture for unpaired image-to-image translation.
DALL-E is a family of text-to-image systems developed by OpenAI that generates images from natural-language descriptions.
A Diffusion Transformer (DiT) is a transformer-based neural network backbone for diffusion models that replaces the U-Net with a Vision Transformer operating on patches of an image latent.
FLUX.1 is a family of text-to-image generation models developed by Black Forest Labs, released on August 1, 2024.
FLUX.2 is the second-generation image generation and editing model family developed by Black Forest Labs, released on November 25, 2025.
Flux is a family of text-to-image generative models developed by Black Forest Labs (BFL), the German-American startup founded by the original creators of Stable Diffusion.
The Frechet Inception Distance (FID) is the standard metric for measuring the quality of images produced by generative models: it computes the Frechet distance between two multivariate Gaussian distributions…
GPT Image 1 (API identifier gpt-image-1) is a natively multimodal image generation model developed by OpenAI, integrated into ChatGPT on March 25, 2025, and released as a standalone API on April 23, 2025.
Grok Imagine is a generative media product from xAI, the company founded by Elon Musk.
Higgsfield AI is an American artificial intelligence company that develops video and image generation software for social media creators, marketers, and enterprise content teams.
Hunyuan Image 3.0 (styled HunyuanImage 3.0) is a text-to-image generation model released by Tencent as part of its Hunyuan family of foundation models.
Ideogram is a Toronto-based artificial intelligence company, founded in 2022 by four former Google Brain researchers who built Google's Imagen system
Ideogram 3.0 is a text-to-image generation model released by Ideogram on March 26, 2025.
Imagen is a family of text-to-image diffusion models developed by Google, first introduced in May 2022 and as of 2026 in its fourth generation (Imagen 4).
Imagen 3 is a text-to-image generation model developed by Google DeepMind, announced at Google I/O on May 14, 2024 and progressively rolled out to users through mid-2024 and into 2025.
Imagen 4 is the fourth-generation text-to-image model developed by Google DeepMind, announced on May 20, 2025, at Google I/O 2025.
Krea AI (legally Krea, krea.ai) is a San Francisco generative AI company, founded in 2022, that builds a browser-based creative suite for image generation, video creation, three-dimensional asset production…
Latent Consistency Models (LCMs) are a family of accelerated text-to-image generative models that apply the consistency-models framework of Song et al.
Leonardo.AI (stylized as Leonardo.Ai) is a Sydney-based generative AI platform for image generation, video creation, and creative design, founded in December 2022 by CEO JJ Fiasson and five co-founders and…
Magnific AI is an AI-powered image upscaling and enhancement platform founded in November 2023 by Javi Lopez and Emilio Nicolas in Murcia, Spain.
Make-A-Scene is a text-to-image generation model published by Meta AI (then Meta AI Research) in 2022.
Midjourney is an artificial intelligence image generation service and independent research lab headquartered in San Francisco, California
Midjourney V7 is the seventh major text-to-image generation model developed by Midjourney Inc. It launched in alpha on April 3, 2025, and became the platform's default model on June 17, 2025.
Muse Image is a proprietary image-generation and editing system developed by Meta Superintelligence Labs, a division of Meta.
NVIDIA Picasso is a cloud-based generative AI foundry from NVIDIA for building, training, and deploying visual generative models that produce images, video, and 3D content from text prompts.
Nano Banana is the codename, later turned official brand, for Google's native image generation and editing models built into the Gemini ecosystem and developed by Google DeepMind.
Nano Banana 2 is the public nickname for Gemini 3.1 Flash Image, an image generation and editing model released by Google DeepMind on 26 February 2026 .
Nano Banana 2 Lite is a text-to-image generation and image editing model released by Google DeepMind on June 30, 2026.
Nano Banana Pro is a professional grade image generation and editing model from Google DeepMind, released on November 20, 2025.
Parti (Pathways Autoregressive Text-to-Image) is a text-to-image generation model from Google Research that produces images from natural-language descriptions by treating the task as a sequence-to-sequence…
Qwen-Image-3.0 is a text-to-image foundation model announced by Alibaba's Qwen team on July 21, 2026, as the third generation of the Qwen-Image series .
Recraft AI is an AI image generation platform built for professional designers, creative teams, and brand-focused workflows.
Recraft V3 is a text-to-image generation model developed by Recraft AI and released on October 30, 2024, that became the first model to reach the number-one position on the Artificial Analysis Text-to-Image…
Reve Image is a family of text-to-image generative models developed by Reve AI, Inc., a Palo Alto, California startup, whose current flagship, Reve 2.0 (released 3 June 2026)
Seedream is a series of text-to-image and image-editing foundation models built by the Seed research team at ByteDance, the company behind TikTok and Douyin.
Seedream 4.0 is a unified image generation and editing model built by the Seed team at ByteDance.
Stability AI is a generative artificial intelligence company best known for helping fund and release the Stable Diffusion family of image models.
Stable Diffusion is a family of generative image models that can synthesize and edit images from text and other conditions.
Stable Diffusion 3 (SD3) is a family of text-to-image diffusion models developed by Stability AI, first announced as an early preview on February 22, 2024, and built on a new architecture called the Multimodal…
StyleGAN is a family of style-based generative adversarial network (GAN) architectures developed by NVIDIA Research for high-quality unconditional image synthesis
Training AI to Paint with Code is an experimental AI art project published by designer and researcher Surya Narreddi in March 2026.
Unstable Diffusion is a Discord community and affiliated commercial platform, operated by the company Equilibrium AI
pix2pix is a supervised image-to-image translation method that learns a mapping between two visual domains from aligned input-output image pairs.