Computer Vision

Explore Computer Vision through related topics and the articles other pages reference most.

Explore articles

Reset filters
Browse subtopics: Generative AI

Articles that also belong to these categories. Counts cover all of Computer Vision.

Showing 1-37 of 37 articles

BigGAN

BigGAN is a class-conditional generative adversarial network that, when introduced by DeepMind researchers Andrew Brock, Jeff Donahue, and Karen Simonyan in 2018, set a new state of the art for AI image…

Generative AIGoogle DeepMind

Diffusion model

A diffusion model is a generative model that learns to transform samples from a simple reference distribution into samples resembling a data distribution by reversing a gradual corruption process.

Deep LearningGenerative AI

Frechet Inception Distance

The Frechet Inception Distance (FID) is the standard metric for measuring the quality of images produced by generative models: it computes the Frechet distance between two multivariate Gaussian distributions…

AI BenchmarksGenerative AI

GAIA-2 (Wayve)

GAIA-2 (Generative AI for Autonomy 2) is a controllable, multi-camera generative world model for autonomous driving, announced by the British self-driving company Wayve on 26 March 2025.

AI ModelsAutonomous Vehicles

GAIA-4 (Wayve)

GAIA-4 is a multimodal generative world model for closed-loop autonomous driving simulation, announced by the British self-driving company Wayve on 3 August 2026 as the latest generation of its GAIA family.

AI ModelsAutonomous Vehicles

Latent diffusion model

A latent diffusion model (LDM) is a type of diffusion model that runs the denoising diffusion process in a compressed latent space learned by a pretrained autoencoder, rather than directly in pixel space…

Deep LearningGenerative AI

Luma Dream Machine

Luma Dream Machine is a generative AI video and image platform from Luma AI (Luma Labs, Inc.), a San Francisco company, that turns text prompts and still images into short, realistic video clips.

AI ModelsGenerative AI

Lumera

Lumera, short for Light-aware Unified Engine-native Reconstruction and Assembly, is an experimental computer vision benchmark and reference pipeline for turning one RGB image into an editable 3D scene.

Generative AI

Nano Banana

Nano Banana is the codename, later turned official brand, for Google's native image generation and editing models built into the Gemini ecosystem and developed by Google DeepMind.

AI ModelsGenerative AI

Photography

Artificial intelligence in photography covers a wide span of techniques, from the computational pipelines baked into modern smartphones to the generative editing tools now built into Photoshop and Lightroom…

AI Tools & ProductsGenerative AI

Pika 2.5

Pika 2.5 is a generative video model developed by Pika Labs, the San Francisco based AI video startup co-founded in April 2023 by Stanford AI Lab dropouts Demi Guo and Chenlin Meng.

AI ModelsGenerative AI

Runway Aleph

Runway Aleph is an in-context AI video editing model from Runway that transforms and edits an existing video clip from a plain-text instruction, performing a wide range of tasks in a single model: adding…

AI ModelsGenerative AI

Sand.ai

Sand.ai is an artificial-intelligence company that develops video-generation models, research software, and commercial creation tools.

AI CompaniesChinese AI

Seedance

Seedance is the family of foundation video generation models built by the Seed team at ByteDance, the Chinese internet company that owns TikTok and Douyin.

AI ModelsChinese AI

Seedance 2.5

Seedance 2.5 is a proprietary AI video generation model released by ByteDance on July 31, 2026. It jointly generates audio and video from combinations of text, image, video, and audio inputs.

AI ModelsChinese AI

Seedream

Seedream is a series of text-to-image and image-editing foundation models built by the Seed research team at ByteDance, the company behind TikTok and Douyin.

AI ModelsChinese AI

Spatial intelligence

Spatial intelligence is the ability of an AI system to perceive, understand, reason about, generate, and interact with three-dimensional space rather than just text or two-dimensional pixels.

Embodied AIGenerative AI

StyleGAN

StyleGAN is a family of style-based generative adversarial network (GAN) architectures developed by NVIDIA Research for high-quality unconditional image synthesis

Generative AIImage Generation

Synthesia 3.0

Synthesia 3.0 is a major release of the AI video generation platform from Synthesia, the London-based company co-founded in 2017 by Victor Riparbelli, Steffen Tjerrild, Lourdes Agapito and Matthias Niessner.

AI ModelsGenerative AI

Wan 2.1-VACE

Wan 2.1-VACE (also written Wan2.1-VACE) is an open-weights video creation and editing model released by Alibaba's Tongyi Lab on May 14, 2025 .

AI ModelsChinese AI

Wan 2.5

Wan 2.5 is a natively multimodal AI video generation model developed by Alibaba Cloud's Tongyi Lab and previewed at the company's Apsara 2025 conference in Hangzhou on September 24, 2025.

AI ModelsChinese AI

Wayve

Wayve is a British artificial intelligence company, founded in Cambridge in 2017 by Alex Kendall and Amar Shah, that develops end to end "embodied AI" software for autonomous driving and robotics, and in May…

AI CompaniesAutonomous Vehicles

World Labs

World Labs is an American spatial-intelligence company, headquartered in San Francisco, that builds "Large World Models" (LWMs), generative AI systems that perceive, generate, reason about and interact with…

AI CompaniesGenerative AI

pix2pix

pix2pix is a supervised image-to-image translation method that learns a mapping between two visual domains from aligned input-output image pairs.

Generative AIImage Generation