AI art
AI art is artwork created with the assistance of artificial intelligence systems, spanning visual images, music, video, and other creative outputs generated or co-created by algorithms, neural networks, and…
Explore Image Generation through related topics and the articles other pages reference most.
Ranked by links from other AI Wiki pages.
Articles that also belong to these categories. Counts cover all of Image Generation.
Showing 1-60 of 84 articles
AI art is artwork created with the assistance of artificial intelligence systems, spanning visual images, music, video, and other creative outputs generated or co-created by algorithms, neural networks, and…
AUTOMATIC1111 Stable Diffusion Web UI (commonly called A1111, SD WebUI, or simply Automatic1111) is the open-source
Adobe Firefly is a family of creative generative AI models developed by Adobe and integrated across its product ecosystem.
AI art is visual, audio, or other creative work produced with the help of artificial intelligence, most commonly images generated from text prompts by text-to-image models such as DALL-E, Midjourney, and…
Art ChatGPT Plugins were a small category of third-party extensions inside ChatGPT that connected the chatbot to fine art collections, image generation backends, and prompt-crafting helpers oriented toward…
As of July 2026, the best all-around AI image generator is GPT Image 2 from OpenAI.
Black Forest Labs (BFL) is a German-American artificial intelligence company founded in 2024 by Robin Rombach, Andreas Blattmann, Patrick Esser, and Dominik Lorenz
CLIP Score (also written CLIPScore or CLIP-S) is a reference-free automatic evaluation metric that measures how well a text caption matches an image, computed as the rescaled cosine similarity of the image and…
CM3leon (pronounced "chameleon") is a multimodal generative model from Meta AI, introduced in July 2023, that handles both text-to-image and image-to-text generation in a single architecture.
Civitai (pronounced "siv-it-eye") is the largest community hub and marketplace for sharing open-source image generation models, hosting hundreds of thousands of Stable Diffusion and Flux checkpoints, LoRAs…
Clipdrop is an AI-powered image editing and creation platform developed by the French startup Init ML.
ComfyUI is a free, open-source node-based graphical user interface and backend for generative AI workflows, primarily focused on image generation using diffusion models such as Stable Diffusion and FLUX.
ControlNet is a neural network architecture that adds spatial and structural control to large pretrained text-to-image diffusion models.
CycleGAN (Cycle-Consistent Generative Adversarial Network) is a deep learning architecture for unpaired image-to-image translation.
DALL-E is a family of text-to-image systems developed by OpenAI that generates images from natural-language descriptions.
DALL-E (Agent) refers to the official DALL·E GPT published by OpenAI inside ChatGPT, a first-party agent variant of ChatGPT that exposes DALL·E 3 image generation through a dedicated conversational interface.
DALL-E 2 (stylized DALL·E 2) is a text-to-image generation system that OpenAI announced on April 6, 2022, capable of producing photorealistic 1024 by 1024 pixel images from a written prompt
DALL-E 3 (stylized by OpenAI as DALL·E 3) is the third generation of OpenAI's text-to-image system, announced on September 20, 2023 and released to ChatGPT Plus and Enterprise subscribers in October 2023
A Diffusion Transformer (DiT) is a transformer-based neural network backbone for diffusion models that replaces the U-Net with a Vision Transformer operating on patches of an image latent.
Doubao-Seedream is the family of text-to-image generation foundation models developed by the ByteDance Seed team and shipped through ByteDance's Doubao product line and the company's Volcano Engine cloud…
Emu is a text-to-image generation foundation model developed by Meta AI and unveiled at the Meta Connect conference in September 2023.
Emu Edit is an instruction-based image editing model from Meta AI, announced on November 16, 2023 alongside the text-to-video model Emu Video.
FLUX.1 is a family of text-to-image generation models developed by Black Forest Labs, released on August 1, 2024.
FLUX.2 is the second-generation image generation and editing model family developed by Black Forest Labs, released on November 25, 2025.
Flux is a family of text-to-image generative models developed by Black Forest Labs (BFL), the German-American startup founded by the original creators of Stable Diffusion.
The Frechet Inception Distance (FID) is the standard metric for measuring the quality of images produced by generative models: it computes the Frechet distance between two multivariate Gaussian distributions…
GLIDE (Guided Language to Image Diffusion for Generation and Editing) is a text-conditional diffusion model for text-to-image synthesis and editing released by OpenAI in December 2021.
GPT Image 1 (API identifier gpt-image-1) is a natively multimodal image generation model developed by OpenAI, integrated into ChatGPT on March 25, 2025, and released as a standalone API on April 23, 2025.
GPT Image 2 (API model gpt-image-2, marketed inside the product as ChatGPT Images 2.0) is an image-generation model released by OpenAI on April 21, 2026.
GPT Image 2.5 is a family of image generation and editing models released by OpenAI on September 8, 2026.
GenEval is an object-focused benchmark for evaluating how well text-to-image models follow the content of a prompt.
Getty Images v. Stability AI is a pair of parallel intellectual property lawsuits brought by the visual content licensing company Getty Images against Stability AI
Grok Imagine is a generative media product from xAI, the company founded by Elon Musk.
HiDream most commonly refers to HiDream-I1, an open-source text-to-image generative foundation model released in April 2025 by the Chinese company HiDream.ai (Chinese: 智象未来).
Higgsfield AI is an American artificial intelligence company that develops video and image generation software for social media creators, marketers, and enterprise content teams.
Hunyuan Image 3.0 (styled HunyuanImage 3.0) is a text-to-image generation model released by Tencent as part of its Hunyuan family of foundation models.
Ideogram is a Toronto-based artificial intelligence company, founded in 2022 by four former Google Brain researchers who built Google's Imagen system
Ideogram 3.0 is a text-to-image generation model released by Ideogram on March 26, 2025.
Imagen is a family of text-to-image diffusion models developed by Google, first introduced in May 2022 and as of 2026 in its fourth generation (Imagen 4).
Imagen 2 is the second generation of Google's text-to-image diffusion model, developed by Google DeepMind and first announced for developers and enterprises on December 13, 2023.
Imagen 3 is a text-to-image generation model developed by Google DeepMind, announced at Google I/O on May 14, 2024 and progressively rolled out to users through mid-2024 and into 2025.
Imagen 4 is the fourth-generation text-to-image model developed by Google DeepMind, announced on May 20, 2025, at Google I/O 2025.
Jimeng (Chinese: 即梦, "instant dream"), branded internationally as Dreamina, is a consumer AI image and video generation platform developed by ByteDance that turns text prompts, reference images, or both into…
Kaiber AI is a creative technology company that builds AI video generation tools for musicians, visual artists, and content creators.
Krea AI (legally Krea, krea.ai) is a San Francisco generative AI company, founded in 2022, that builds a browser-based creative suite for image generation, video creation, three-dimensional asset production…
Latent Consistency Models (LCMs) are a family of accelerated text-to-image generative models that apply the consistency-models framework of Song et al.
Leonardo.AI (stylized as Leonardo.Ai) is a Sydney-based generative AI platform for image generation, video creation, and creative design, founded in December 2022 by CEO JJ Fiasson and five co-founders and…
MAI-Image-2.6 is a text-to-image and image-editing model built in-house by Microsoft AI, the division led by Mustafa Suleyman.
MMDiT (Multimodal Diffusion Transformer, sometimes written MM-DiT) is a transformer architecture for text-conditioned image generation that gives image tokens and text tokens their own separate weights but…
Magnific AI is an AI-powered image upscaling and enhancement platform founded in November 2023 by Javi Lopez and Emilio Nicolas in Murcia, Spain.
Make-A-Scene is a text-to-image generation model published by Meta AI (then Meta AI Research) in 2022.
MidJourney Prompt Generator is a generic category name that refers to the broad class of tools, custom GPTs, and prompt templates that help users compose effective text prompts for Midjourney
Midjourney is an artificial intelligence image generation service and independent research lab headquartered in San Francisco, California
Midjourney V7 is the seventh major text-to-image generation model developed by Midjourney Inc. It launched in alpha on April 3, 2025, and became the platform's default model on June 17, 2025.
Muse Image is a proprietary image-generation and editing system developed by Meta Superintelligence Labs, a division of Meta.
NVIDIA Picasso is a cloud-based generative AI foundry from NVIDIA for building, training, and deploying visual generative models that produce images, video, and 3D content from text prompts.
Nano Banana is the codename, later turned official brand, for Google's native image generation and editing models built into the Gemini ecosystem and developed by Google DeepMind.
Nano Banana 2 is the public nickname for Gemini 3.1 Flash Image, an image generation and editing model released by Google DeepMind on 26 February 2026 .
Nano Banana 2 Lite is a text-to-image generation and image editing model released by Google DeepMind on June 30, 2026.
Nano Banana Pro is a professional grade image generation and editing model from Google DeepMind, released on November 20, 2025.