Photography
Artificial intelligence in photography covers a wide span of techniques, from the computational pipelines baked into modern smartphones to the generative editing tools now built into Photoshop and Lightroom…
Explore Generative AI through related topics and the articles other pages reference most.
Articles that also belong to these categories. Counts cover all of Generative AI.
Showing 181-240 of 264 articles
Artificial intelligence in photography covers a wide span of techniques, from the computational pipelines baked into modern smartphones to the generative editing tools now built into Photoshop and Lightroom…
Pika is an artificial intelligence video generation platform developed by Pika Labs, Inc. that lets users create and edit short videos from text prompts, images, and existing clips, and it is best known for…
Pika 2.5 is a generative video model developed by Pika Labs, the San Francisco based AI video startup co-founded in April 2023 by Stanford AI Lab dropouts Demi Guo and Chenlin Meng.
Pika Labs (legally incorporated as Mellis, Inc. and doing business as Pika) is an American generative AI company that builds consumer software for creating short AI videos from text, images, or clips.
PlayHT, later rebranded PlayAI (and reachable at play.ht and play.ai), was an American generative AI voice company that built text-to-speech models, voice cloning tools, and a platform for conversational AI…
Point-E is a text-to-3D generative system developed by OpenAI that produces colored 3D point clouds from natural-language prompts.
Prompt-to-Prompt is a training-free image editing technique for text-conditioned diffusion models that edits a generated image by manipulating the model's cross-attention maps when the text prompt is changed .
Proto is an open-source programming framework for generative biology developed by researchers at Stanford University and the Arc Institute.
Qwen-Image is an open-weight image-generation foundation model released by Alibaba's Qwen team in August 2025.
Qwen-Image-3.0 is a text-to-image foundation model announced by Alibaba's Qwen team on July 21, 2026, as the third generation of the Qwen-Image series .
Recraft AI is an AI image generation platform built for professional designers, creative teams, and brand-focused workflows.
Recraft V3 is a text-to-image generation model developed by Recraft AI and released on October 30, 2024, that became the first model to reach the number-one position on the Artificial Analysis Text-to-Image…
Rectified Flow is a generative modeling framework that learns a transport ordinary differential equation between two probability distributions by regressing a velocity field along straight-line interpolations…
Replit Agent is an AI software development agent built by Replit that turns a natural language description into a deployed, full-stack web application without the user writing code.
Resemble AI is a generative voice and AI security company based in San Francisco, California, and originally founded in Toronto, Canada.
Reve Image is a family of text-to-image generative models developed by Reve AI, Inc., a Palo Alto, California startup, whose current flagship, Reve 2.0 (released 3 June 2026)
Rime (also styled Rime Labs, and reachable at rime.ai) is an American artificial intelligence company that builds text-to-speech and spoken-language models tuned specifically for business voice agents…
Rodin Gen-2 is a generative artificial intelligence model for producing three dimensional assets from text prompts or reference images, developed by the Shanghai based company Deemos and delivered through the…
Runway is a New York based artificial intelligence company that builds generative AI tools for video, with its Gen series of models and its GWM-1 world models.
Runway Act-Two is a generative motion capture and character animation model developed by Runway, publicly introduced on July 15, 2025.
Runway Aleph is an in-context AI video editing model from Runway that transforms and edits an existing video clip from a plain-text instruction, performing a wide range of tasks in a single model: adding…
Runway Gen-3 Alpha is a generative AI text to video model developed by Runway (the trade name of the New York based company Runway AI, Inc.) and unveiled on June 17, 2024 as the successor to Gen-2 and the…
Runway Gen-4 is the fourth-generation video generation model developed by Runway (company), announced and released on March 31, 2025.
SDEdit (Stochastic Differential Editing) is a method for guided image synthesis and editing that turns a rough user guide, such as a stroke painting, a coarse collage, or a real photograph with edits pasted…
Sand.ai is an artificial-intelligence company that develops video-generation models, research software, and commercial creation tools.
Seedance is the family of foundation video generation models built by the Seed team at ByteDance, the Chinese internet company that owns TikTok and Douyin.
Seedance 2.5 is a proprietary AI video generation model released by ByteDance on July 31, 2026. It jointly generates audio and video from combinations of text, image, video, and audio inputs.
Seedream is a series of text-to-image and image-editing foundation models built by the Seed research team at ByteDance, the company behind TikTok and Douyin.
Seedream 4.0 is a unified image generation and editing model built by the Seed team at ByteDance.
Sesame CSM (Conversational Speech Model) is an open weights speech generation model from Sesame AI, a San Francisco startup co-founded by former Oculus chief executive Brendan Iribe.
Shap-E is a conditional generative model for 3D assets developed by OpenAI, introduced in the paper "Shap-E: Generating Conditional 3D Implicit Functions" by Heewoo Jun and Alex Nichol, submitted to arXiv on…
Show-o is a unified multimodal model, introduced in 2024, that handles both multimodal understanding and visual generation inside a single Transformer.
SlidesAI (stylized SlidesAI.io) is an artificial intelligence tool that generates presentation slides from text.
Sonauto is a generative artificial intelligence music platform that converts text prompts, lyrics, and melody inputs into complete songs with vocals and instrumentation.
Sora was a text-to-video generation model developed by OpenAI. OpenAI revealed the research model on February 15, 2024 and granted early access to red teamers, visual artists, designers, and filmmakers.
Sora is a short-form video application released by OpenAI on September 30, 2025, alongside the Sora 2 text-to-video model that powers it.
Sora 2 is a text-to-video and audio generation model developed by OpenAI, released on September 30, 2025, that OpenAI called "the GPT-3.5 moment for video." It succeeded the original Sora research preview from…
Spatial intelligence is the ability of an AI system to perceive, understand, reason about, generate, and interact with three-dimensional space rather than just text or two-dimensional pixels.
Stability AI is a generative artificial intelligence company best known for helping fund and release the Stable Diffusion family of image models.
Stable Audio is a family of generative AI models from Stability AI that turn a text prompt into music or sound effects as a stereo audio file.
Stable Audio 2.5 is an enterprise focused text-to-audio generation model released by Stability AI on September 10, 2025.
Stable Diffusion is a family of generative image models that can synthesize and edit images from text and other conditions.
Stable Diffusion 3 (SD3) is a family of text-to-image diffusion models developed by Stability AI, first announced as an early preview on February 22, 2024, and built on a new architecture called the Multimodal…
Stefano Ermon is an Italian computer scientist and an associate professor of computer science at Stanford University, best known for foundational work on score-based generative models
StyleGAN is a family of style-based generative adversarial network (GAN) architectures developed by NVIDIA Research for high-quality unconditional image synthesis
Submagic (at the domain submagic.co) is a French software company that builds an AI-powered tool for editing short-form video.
Suno is a generative artificial intelligence company that develops a text-to-music platform capable of producing complete songs, including vocals, instrumentals, and lyrics, from simple text prompts.
Suno v5 is the fifth-generation AI music generation model from Suno Inc., the Cambridge, Massachusetts startup, released on September 23, 2025 to Pro and Premier subscribers as what Suno called "the world's…
SynthID is a family of digital watermarking technologies developed by Google DeepMind for marking and identifying content generated by generative AI systems.
Synthesia is a British artificial intelligence company, founded in 2017 and headquartered in London, that builds a text-to-video platform letting users generate professional videos featuring AI-generated…
Synthesia 3.0 is a major release of the AI video generation platform from Synthesia, the London-based company co-founded in 2017 by Victor Riparbelli, Steffen Tjerrild, Lourdes Agapito and Matthias Niessner.
Tavus is an American generative AI research company headquartered in San Francisco, California, that develops video generation models and real-time conversational video technology.
Text-to-image models are generative artificial intelligence systems that synthesize a new image from a natural-language description, called a prompt.
Text-to-video (often abbreviated T2V) is the generative AI capability of producing video clips, with or without sound, directly from a written prompt.
Textual Inversion is a technique for personalizing text-to-image diffusion models that teaches a frozen model a new visual concept from only three to five example images by learning a single new "pseudo-word"…
Thinking Machines Lab is an American artificial intelligence research company, headquartered in San Francisco, California, that was founded in February 2025 by Mira Murati, the former chief technology officer…
Training AI to Paint with Code is an experimental AI art project published by designer and researcher Surya Narreddi in March 2026.
Transfusion is a recipe, introduced by Meta in 2024, for training a single Transformer over a mixture of discrete text and continuous image data using two training objectives simultaneously: a…
Tripo is an artificial-intelligence platform for generating three-dimensional models from text prompts, single images, multi-view images, or sketches.
Tripo P1, marketed in full as Tripo Smart Mesh P1.0, is a production grade native 3D diffusion model that generates clean, engine ready 3D meshes from text or image prompts in as little as two seconds.