BigGAN
BigGAN is a class-conditional generative adversarial network that, when introduced by DeepMind researchers Andrew Brock, Jeff Donahue, and Karen Simonyan in 2018, set a new state of the art for AI image…
Explore Google DeepMind through related topics and the articles other pages reference most.
Articles that also belong to these categories. Counts cover all of Google DeepMind.
Showing 1-17 of 17 articles
BigGAN is a class-conditional generative adversarial network that, when introduced by DeepMind researchers Andrew Brock, Jeff Donahue, and Karen Simonyan in 2018, set a new state of the art for AI image…
Gemini is a family of natively multimodal large language models developed by Google DeepMind, first announced on December 6, 2023, that can reason across text, images, audio, video, and code within a single…
Gemini 3.8 Flash is a multimodal model in Google's Gemini family. Google DeepMind released it on September 2, 2026 as a generally available model for software engineering, tool-using agents, and knowledge work.
Gemini Omni is a family of proprietary multimodal AI models from Google DeepMind for generating and editing media.
Genie is an 11-billion-parameter generative interactive environment from Google DeepMind, described by its creators as the first foundation world model: it turns a single image, photo, or sketch into a…
Genie 2 is a foundation world model developed by Google DeepMind, unveiled on December 4, 2024.
Genie 3 is a general-purpose foundation world model developed by Google DeepMind and announced on August 5, 2025, that generates interactive, navigable 3D environments from a single text prompt and runs in…
Imagen is a family of text-to-image diffusion models developed by Google, first introduced in May 2022 and as of 2026 in its fourth generation (Imagen 4).
Imagen 3 is a text-to-image generation model developed by Google DeepMind, announced at Google I/O on May 14, 2024 and progressively rolled out to users through mid-2024 and into 2025.
Imagen 4 is the fourth-generation text-to-image model developed by Google DeepMind, announced on May 20, 2025, at Google I/O 2025.
Lyria is a family of AI music generation models developed by Google DeepMind, spanning text-to-music synthesis, real-time interactive music performance, and full-length song composition.
Lyria 2 is a high-fidelity, text-to-music generation model built by Google DeepMind that turns text prompts into professional-grade instrumental audio.
Lyria 3.5 is a music generation model from Google DeepMind and the newest member of the Lyria family.
Nano Banana 2 is the public nickname for Gemini 3.1 Flash Image, an image generation and editing model released by Google DeepMind on 26 February 2026 .
SynthID is a family of digital watermarking technologies developed by Google DeepMind for marking and identifying content generated by generative AI systems.
Veo is a family of text-to-video generative AI models developed by Google DeepMind, and is best known as the first video model from a leading AI lab to natively generate synchronized audio (dialogue, sound…
Veo 2 is a text-to-video generative AI model developed by Google DeepMind and announced on December 16, 2024, the second major iteration of the Veo family.