Udio
Udio is an artificial intelligence music generation platform developed by Uncharted Labs, Inc. that creates full songs, complete with vocals, instrumentation, and lyrics, from a single text prompt.
Explore Generative AI through related topics and the articles other pages reference most.
Articles that also belong to these categories. Counts cover all of Generative AI.
Showing 241-264 of 264 articles
Udio is an artificial intelligence music generation platform developed by Uncharted Labs, Inc. that creates full songs, complete with vocals, instrumentation, and lyrics, from a single text prompt.
Uizard is an AI-powered UI/UX design tool that lets people create app and website mockups from text prompts, screenshots, or hand-drawn sketches
Unstable Diffusion is a Discord community and affiliated commercial platform, operated by the company Equilibrium AI
VQ-VAE (Vector Quantized Variational Autoencoder) is a generative neural network that compresses data into a grid or sequence of discrete tokens drawn from a learned codebook
VQGAN (Vector Quantized Generative Adversarial Network) is a two-stage image-synthesis method that first compresses an image into a small grid of discrete codebook tokens with an adversarially trained…
A variational autoencoder (VAE) is a latent-variable generative model that pairs a probabilistic decoder with a learned approximation to posterior inference.
Veo is a family of text-to-video generative AI models developed by Google DeepMind, and is best known as the first video model from a leading AI lab to natively generate synchronized audio (dialogue, sound…
Veo 2 is a text-to-video generative AI model developed by Google DeepMind and announced on December 16, 2024, the second major iteration of the Veo family.
Visual Autoregressive modeling (VAR) is an image generation paradigm, introduced in 2024, that reframes autoregressive image synthesis as coarse-to-fine "next-scale prediction" rather than the conventional…
Vizard (stylized Vizard.ai) is an artificial intelligence video tool that turns long-form videos into short, social-ready clips with automatically generated captions, speaker tracking, and reframing for…
Voice AI is the umbrella term for artificial intelligence systems that listen to, understand, and generate human speech.
Voice cloning is the use of machine learning to generate synthetic speech in the voice of a specific real person (the target speaker) from a sample of their recorded audio.
Voicebox is a non-autoregressive, text-conditioned generative model for speech developed by Meta AI Research and announced on June 16, 2023.
Wan 2.1 (also written Wan2.1, from the Chinese Tongyi Wanxiang or 通义万象) is a family of open-weights text-to-video and image-to-video generation models that Alibaba's Tongyi Wanxiang team released and…
Wan 2.1-VACE (also written Wan2.1-VACE) is an open-weights video creation and editing model released by Alibaba's Tongyi Lab on May 14, 2025 .
Wan 2.5 is a natively multimodal AI video generation model developed by Alibaba Cloud's Tongyi Lab and previewed at the company's Apsara 2025 conference in Hangzhou on September 24, 2025.
A Wasserstein GAN (WGAN) is a generative adversarial network that trains its two networks to minimise the Wasserstein-1 distance, also called the Earth mover's distance, between the real data distribution and…
Wasserstein loss is a loss function for training generative models that measures the distance between two probability distributions as the Wasserstein-1 distance
Wayve is a British artificial intelligence company, founded in Cambridge in 2017 by Alex Kendall and Amar Shah, that develops end to end "embodied AI" software for autonomous driving and robotics, and in May…
World Labs is an American spatial-intelligence company, headquartered in San Francisco, that builds "Large World Models" (LWMs), generative AI systems that perceive, generate, reason about and interact with…
AI is used for writing by applying large language models to draft, edit, rewrite, summarize, and translate text, usually through a chatbot like ChatGPT or through features built into grammar checkers, word…
Z.ai is the international brand of the Chinese artificial intelligence company Zhipu AI (智谱AI), a 2019 spinout from Tsinghua University that builds the open-weight General Language Model (GLM) family and…
fal.ai (legally Features and Labels, Inc., often stylized fal) is a San Francisco-based artificial intelligence company that operates a specialized inference platform for generative-media models, including…
pix2pix is a supervised image-to-image translation method that learns a mapping between two visual domains from aligned input-output image pairs.