Generative AI

Explore Generative AI through related topics and the articles other pages reference most.

Explore articles

Browse subtopics (61)

Articles that also belong to these categories. Counts cover all of Generative AI.

Showing 181-240 of 264 articles

Photography

Artificial intelligence in photography covers a wide span of techniques, from the computational pipelines baked into modern smartphones to the generative editing tools now built into Photoshop and Lightroom…

AI Tools & ProductsComputer Vision

Pika (video generation)

Pika is an artificial intelligence video generation platform developed by Pika Labs, Inc. that lets users create and edit short videos from text prompts, images, and existing clips, and it is best known for…

AI CompaniesAI Models

Pika 2.5

Pika 2.5 is a generative video model developed by Pika Labs, the San Francisco based AI video startup co-founded in April 2023 by Stanford AI Lab dropouts Demi Guo and Chenlin Meng.

AI ModelsComputer Vision

Pika Labs

Pika Labs (legally incorporated as Mellis, Inc. and doing business as Pika) is an American generative AI company that builds consumer software for creating short AI videos from text, images, or clips.

AI CompaniesVideo Generation

PlayHT

PlayHT, later rebranded PlayAI (and reachable at play.ht and play.ai), was an American generative AI voice company that built text-to-speech models, voice cloning tools, and a platform for conversational AI…

AI CompaniesSpeech & Audio AI

Point-E

Point-E is a text-to-3D generative system developed by OpenAI that produces colored 3D point clouds from natural-language prompts.

OpenAI

Prompt-to-Prompt

Prompt-to-Prompt is a training-free image editing technique for text-conditioned diffusion models that edits a generated image by manipulating the model's cross-attention maps when the text prompt is changed .

Deep Learning

Qwen-Image

Qwen-Image is an open-weight image-generation foundation model released by Alibaba's Qwen team in August 2025.

AI Models

Qwen-Image-3.0

Qwen-Image-3.0 is a text-to-image foundation model announced by Alibaba's Qwen team on July 21, 2026, as the third generation of the Qwen-Image series .

AI ModelsChinese AI

Recraft V3

Recraft V3 is a text-to-image generation model developed by Recraft AI and released on October 30, 2024, that became the first model to reach the number-one position on the Artificial Analysis Text-to-Image…

AI ModelsImage Generation

Rectified Flow

Rectified Flow is a generative modeling framework that learns a transport ordinary differential equation between two probability distributions by regressing a velocity field along straight-line interpolations…

Diffusion Models

Replit Agent

Replit Agent is an AI software development agent built by Replit that turns a natural language description into a deployed, full-stack web application without the user writing code.

AI AgentsAI Companies

Reve Image

Reve Image is a family of text-to-image generative models developed by Reve AI, Inc., a Palo Alto, California startup, whose current flagship, Reve 2.0 (released 3 June 2026)

AI CompaniesImage Generation

Rime (company)

Rime (also styled Rime Labs, and reachable at rime.ai) is an American artificial intelligence company that builds text-to-speech and spoken-language models tuned specifically for business voice agents…

AI CompaniesSpeech & Audio AI

Rodin Gen-2

Rodin Gen-2 is a generative artificial intelligence model for producing three dimensional assets from text prompts or reference images, developed by the Shanghai based company Deemos and delivered through the…

AI ModelsChinese AI

Runway Aleph

Runway Aleph is an in-context AI video editing model from Runway that transforms and edits an existing video clip from a plain-text instruction, performing a wide range of tasks in a single model: adding…

AI ModelsComputer Vision

Runway Gen-3 Alpha

Runway Gen-3 Alpha is a generative AI text to video model developed by Runway (the trade name of the New York based company Runway AI, Inc.) and unveiled on June 17, 2024 as the successor to Gen-2 and the…

Video Generation

SDEdit

SDEdit (Stochastic Differential Editing) is a method for guided image synthesis and editing that turns a rough user guide, such as a stroke painting, a coarse collage, or a real photograph with edits pasted…

Deep Learning

Sand.ai

Sand.ai is an artificial-intelligence company that develops video-generation models, research software, and commercial creation tools.

AI CompaniesChinese AI

Seedance

Seedance is the family of foundation video generation models built by the Seed team at ByteDance, the Chinese internet company that owns TikTok and Douyin.

AI ModelsChinese AI

Seedance 2.5

Seedance 2.5 is a proprietary AI video generation model released by ByteDance on July 31, 2026. It jointly generates audio and video from combinations of text, image, video, and audio inputs.

AI ModelsChinese AI

Seedream

Seedream is a series of text-to-image and image-editing foundation models built by the Seed research team at ByteDance, the company behind TikTok and Douyin.

AI ModelsChinese AI

Sesame CSM

Sesame CSM (Conversational Speech Model) is an open weights speech generation model from Sesame AI, a San Francisco startup co-founded by former Oculus chief executive Brendan Iribe.

AI ModelsOpen Source AI

Shap-E

Shap-E is a conditional generative model for 3D assets developed by OpenAI, introduced in the paper "Shap-E: Generating Conditional 3D Implicit Functions" by Heewoo Jun and Alex Nichol, submitted to arXiv on…

AI ModelsOpenAI

Show-o

Show-o is a unified multimodal model, introduced in 2024, that handles both multimodal understanding and visual generation inside a single Transformer.

Deep Learning

Sora

Sora was a text-to-video generation model developed by OpenAI. OpenAI revealed the research model on February 15, 2024 and granted early access to red teamers, visual artists, designers, and filmmakers.

Diffusion ModelsOpenAI

Sora 2

Sora 2 is a text-to-video and audio generation model developed by OpenAI, released on September 30, 2025, that OpenAI called "the GPT-3.5 moment for video." It succeeded the original Sora research preview from…

AI ModelsMultimodal AI

Spatial intelligence

Spatial intelligence is the ability of an AI system to perceive, understand, reason about, generate, and interact with three-dimensional space rather than just text or two-dimensional pixels.

Computer VisionEmbodied AI

Stable Diffusion 3

Stable Diffusion 3 (SD3) is a family of text-to-image diffusion models developed by Stability AI, first announced as an early preview on February 22, 2024, and built on a new architecture called the Multimodal…

AI CompaniesDiffusion Models

Stefano Ermon

Stefano Ermon is an Italian computer scientist and an associate professor of computer science at Stanford University, best known for foundational work on score-based generative models

Machine LearningPeople

StyleGAN

StyleGAN is a family of style-based generative adversarial network (GAN) architectures developed by NVIDIA Research for high-quality unconditional image synthesis

Computer VisionImage Generation

Suno

Suno is a generative artificial intelligence company that develops a text-to-music platform capable of producing complete songs, including vocals, instrumentals, and lyrics, from simple text prompts.

AI CompaniesMusic & Audio Generation

Suno v5

Suno v5 is the fifth-generation AI music generation model from Suno Inc., the Cambridge, Massachusetts startup, released on September 23, 2025 to Pro and Premier subscribers as what Suno called "the world's…

AI ModelsMusic & Audio Generation

SynthID

SynthID is a family of digital watermarking technologies developed by Google DeepMind for marking and identifying content generated by generative AI systems.

AI SafetyGoogle DeepMind

Synthesia

Synthesia is a British artificial intelligence company, founded in 2017 and headquartered in London, that builds a text-to-video platform letting users generate professional videos featuring AI-generated…

AI CompaniesVideo Generation

Synthesia 3.0

Synthesia 3.0 is a major release of the AI video generation platform from Synthesia, the London-based company co-founded in 2017 by Victor Riparbelli, Steffen Tjerrild, Lourdes Agapito and Matthias Niessner.

AI ModelsComputer Vision

Tavus

Tavus is an American generative AI research company headquartered in San Francisco, California, that develops video generation models and real-time conversational video technology.

AI CompaniesConversational AI

Textual Inversion

Textual Inversion is a technique for personalizing text-to-image diffusion models that teaches a frozen model a new visual concept from only three to five example images by learning a single new "pseudo-word"…

Deep Learning

Thinking Machines Lab

Thinking Machines Lab is an American artificial intelligence research company, headquartered in San Francisco, California, that was founded in February 2025 by Mira Murati, the former chief technology officer…

AI Companies

Transfusion

Transfusion is a recipe, introduced by Meta in 2024, for training a single Transformer over a mixture of discrete text and continuous image data using two training objectives simultaneously: a…

Deep Learning

Tripo P1

Tripo P1, marketed in full as Tripo Smart Mesh P1.0, is a production grade native 3D diffusion model that generates clean, engine ready 3D meshes from text or image prompts in as little as two seconds.

AI ModelsChinese AI