Decktopus
Decktopus (also styled Decktopus AI) is an online presentation builder developed by Decktopus, Inc. that generates fully designed slide decks from a short text prompt or topic.
Explore Generative AI through related topics and the articles other pages reference most.
Articles that also belong to these categories. Counts cover all of Generative AI.
Showing 61-120 of 264 articles
Decktopus (also styled Decktopus AI) is an online presentation builder developed by Decktopus, Inc. that generates fully designed slide decks from a short text prompt or topic.
Design is one of the creative fields most visibly transformed by artificial intelligence.
Dia is an AI-first web browser developed by The Browser Company of New York, the studio behind the Arc browser, and owned since October 2025 by Atlassian.
Diffusion Forcing is a training paradigm for sequence generative modeling introduced in 2024 that assigns each token in a sequence its own independent, randomly sampled noise level during training .
A Diffusion Transformer (DiT) is a transformer-based neural network backbone for diffusion models that replaces the U-Net with a Vision Transformer operating on patches of an image latent.
A diffusion model is a generative model that learns to transform samples from a simple reference distribution into samples resembling a data distribution by reversing a gradual corruption process.
A discrete diffusion language model is a class of generative model for text that produces tokens by iteratively denoising a corrupted sequence, rather than by predicting one token at a time from left to right.
A discriminator is the neural network in a generative adversarial network (GAN) that is trained to tell real data apart from data produced by the generator
DreamBooth is a subject-driven fine-tuning method for text-to-image diffusion models that personalizes a pretrained model to a specific subject, for example a particular dog, toy, or person, from just 3 to 5…
EDM is the common shorthand for the paper "Elucidating the Design Space of Diffusion-Based Generative Models" by Tero Karras, Miika Aittala, Timo Aila, and Samuli Laine of NVIDIA, presented at NeurIPS 2022.
ElevenLabs is a voice and audio artificial intelligence company that builds text-to-speech AI, voice cloning, AI dubbing, generative sound effects, music synthesis, and conversational voice agent technology.
ElevenLabs Music (also marketed as Eleven Music) is an AI music generation product developed by ElevenLabs, the voice AI company founded in 2022 by Piotr Dabkowski and Mati Staniszewski.
Eleven v3, marketed by ElevenLabs as Eleven v3 (alpha), is a third-generation text-to-speech model that ElevenLabs released in public alpha on June 5, 2025 and described as "the most expressive Text to Speech…
Emad Mostaque is a British entrepreneur and former hedge-fund manager who founded Stability AI, the company that released and popularized the open-weight image model Stable Diffusion.
FLUX 3 Video is a video generation model by Black Forest Labs (BFL), released into general availability on August 4, 2026.
FLUX.1 is a family of text-to-image generation models developed by Black Forest Labs, released on August 1, 2024.
FLUX.2 is the second-generation image generation and editing model family developed by Black Forest Labs, released on November 25, 2025.
Flow Matching is a simulation-free training framework for generative models that fits a time-dependent velocity field to transport samples from a source distribution (typically a standard Gaussian) to a data…
Flux is a family of text-to-image generative models developed by Black Forest Labs (BFL), the German-American startup founded by the original creators of Stable Diffusion.
The Frechet Inception Distance (FID) is the standard metric for measuring the quality of images produced by generative models: it computes the Frechet distance between two multivariate Gaussian distributions…
GAIA-2 (Generative AI for Autonomy 2) is a controllable, multi-camera generative world model for autonomous driving, announced by the British self-driving company Wayve on 26 March 2025.
GAIA-3 is a 15-billion-parameter generative world model for autonomous driving released by Wayve on 2 December 2025, the third generation in the company's GAIA family.
GAIA-4 is a multimodal generative world model for closed-loop autonomous driving simulation, announced by the British self-driving company Wayve on 3 August 2026 as the latest generation of its GAIA family.
GAN stands for generative adversarial network, a class of deep learning generative models in which two neural networks are trained against each other: a generator that fabricates synthetic data and a…
GPT Image 1 (API identifier gpt-image-1) is a natively multimodal image generation model developed by OpenAI, integrated into ChatGPT on March 25, 2025, and released as a standalone API on April 23, 2025.
OpenAI's GPT lineage runs from GPT-1 (June 2018, 117 million parameters, a 512-token context window) to the GPT-5.6 "Sol, Terra, Luna" family that entered limited preview on June 26, 2026.
GPT-4 (Generative Pre-trained Transformer 4) is a large language model developed by OpenAI and released on March 14, 2023.
Galileo AI (at the domain usegalileo.ai) was a generative AI design tool that turned plain-text prompts into editable, high-fidelity user interface designs, a workflow it popularized as "text to UI." It was…
AI in gaming covers the artificial intelligence systems used inside games, the AI tools used to make games, and the AI agents that have learned to beat humans at games as research milestones.
Gemini is a family of natively multimodal large language models developed by Google DeepMind, first announced on December 6, 2023, that can reason across text, images, audio, video, and code within a single…
Gemini 3.6 Flash is a proprietary, multimodal large language model released by Google on July 21, 2026. It belongs to the Gemini 3 series and uses the stable API identifier gemini-3.6-flash.
Gemini 3.7 Flash is a proprietary, multimodal large language model released by Google on August 13, 2026.
Gemini 3.8 Flash is a multimodal model in Google's Gemini family. Google DeepMind released it on September 2, 2026 as a generally available model for software engineering, tool-using agents, and knowledge work.
Gemini Omni is a family of proprietary multimodal AI models from Google DeepMind for generating and editing media.
Generative AI is a class of artificial intelligence systems that produces new data instances, such as text, software code, images, audio, video, molecular structures, or other representations, by learning…
A generative model is a class of statistical and machine learning model that learns the joint probability distribution P(X) of the observed data, or the joint distribution P(X, Y) of inputs and labels
Generative adversarial networks (GANs) are a family of generative models trained through competition between two learned functions.
A generator is a neural network within a generative adversarial network (GAN) that learns to produce synthetic data samples from random noise.
Genie is an 11-billion-parameter generative interactive environment from Google DeepMind, described by its creators as the first foundation world model: it turns a single image, photo, or sketch into a…
Genie 2 is a foundation world model developed by Google DeepMind, unveiled on December 4, 2024.
Genie 3 is a general-purpose foundation world model developed by Google DeepMind and announced on August 5, 2025, that generates interactive, navigable 3D environments from a single text prompt and runs in…
Genmo is a San Francisco artificial intelligence company that builds open video generation models, best known for Mochi 1
GitHub Spark is a natural-language application builder developed by GitHub that lets users describe a web app in plain English and watch it be generated, deployed, and hosted without manually writing or…
Google AI Studio is Google's free, web-based developer tool for prototyping with the Gemini family of models and getting a Gemini API key in seconds.
Grok is a family of generative artificial intelligence chatbots and large language models (LLMs) built by xAI, the company founded by Elon Musk.
Grok 4.5 is a proprietary multimodal large language model and reasoning model in the Grok family. It was developed by SpaceXAI in collaboration with Cursor and released through the xAI API on July 8, 2026.
Grok 4.6 is a proprietary large language model and reasoning model in the Grok family, developed by SpaceXAI and released jointly with Cursor through the xAI API on August 12, 2026.
Grok Imagine is a generative media product from xAI, the company founded by Elon Musk.
Grokipedia is an online encyclopedia created by xAI, the artificial intelligence company founded by Elon Musk.
Hailuo AI (海螺AI) is the consumer AI video generation service operated by MiniMax, a Shanghai-based foundation model developer that has traded on the Hong Kong Stock Exchange as MiniMax Group Inc. (stock code…
Haiper was a London-based generative AI startup that built consumer video-generation models, operating the haiper.ai web app between 2024 and early 2025.
Hedra Character is a family of generative video foundation models that turn a single image plus an audio clip (and, in later versions, a text prompt) into a video in which the pictured person or character…
Hell Grind is a 2026 action fantasy feature film produced by the AI video generation company Higgsfield, which says every character, setting, and prop in the 95-minute film was generated with AI tools on its…
HeyGen is an American artificial intelligence company headquartered in Los Angeles, California, that develops AI-powered video generation software specializing in digital avatars, voice cloning, and…
Avatar IV is the fourth generation of the AI avatar engine from HeyGen, the AI video company co-founded in 2020 by Joshua Xu and Wayne Liang.
Higgsfield AI is an American artificial intelligence company that develops video and image generation software for social media creators, marketers, and enterprise content teams.
Hume Octave 2 is a multilingual emotional text-to-speech model released by Hume AI on October 1, 2025.
Hunyuan 3D is a family of open weight generative artificial intelligence models from Tencent that turn text prompts, single images, sketches, and other inputs into ready to use three dimensional assets…
Hunyuan Image 3.0 (styled HunyuanImage 3.0) is a text-to-image generation model released by Tencent as part of its Hunyuan family of foundation models.
HunyuanVideo is an open-source video generation model developed by Tencent and released on December 3, 2024, with over 13 billion parameters