Sunday, June 28, 2026
- Yan Junjiev3Yan Junjie (闫俊杰, born 1989) is a Chinese artificial intelligence researcher and entrepreneur who is the founder, chairman, chief executive officer, and chief...
- Nemotron Nano 2v2Nemotron Nano 2 is a family of small, open-weight reasoning language models released by NVIDIA on August 18, 2025, built on a hybrid Mamba-2 state space model...
- MemGPTv3MemGPT (short for Memory-GPT) is a system and agent design pattern that gives large language model agents long-term memory by managing data between the model's...
- GPT Image 2v2GPT Image 2 (API model gpt-image-2, marketed inside the product as ChatGPT Images 2.0) is an image-generation model released by OpenAI on April 21, 2026. It is...
- NVIDIA Cosmos Reasonv2NVIDIA Cosmos Reason is an open, customizable, 7-billion-parameter reasoning vision-language model (VLM) for physical AI and robotics developed by Nvidia. It...
- Stargate UAEv3Stargate UAE is a large artificial intelligence data center cluster being built in Abu Dhabi, United Arab Emirates, and is the first site of the Stargate...
- SIMPLERv2SIMPLER (Simulated Manipulation Policy Evaluation for Real Robot Setups) is a collection of simulated robot manipulation environments, released in 2024, that...
- Minitronv2Minitron is a family of compact language models from NVIDIA, together with the model-compression method used to build them: take one large, already pretrained...
- DualPipev2DualPipe is a bidirectional pipeline parallelism scheduling algorithm created by DeepSeek-AI to almost fully overlap computation with communication and shrink...
- DPM-Solverv2DPM-Solver is a fast, training-free high-order numerical solver for the ordinary differential equations (ODEs) that arise when sampling from diffusion models,...
- Tabular modelsv3Tabular models are machine learning systems that learn from data arranged in tables, where each row is a sample and each column is a feature. The defining...
- Gemini 1.0v2Gemini 1.0 is the first generation of Gemini, the family of natively multimodal AI models that Google DeepMind announced on 6 December 2023 [1][2]. It shipped...
- Dyna Roboticsv2Dyna Robotics is an American embodied-AI and robotics startup, founded in 2024 and based in Redwood City, California, that builds robot foundation models for...
- Dragoneer Investment Groupv3Dragoneer Investment Group is a San Francisco growth-oriented investment firm, founded in 2012 by Marc Stad, that invests in high-growth technology companies...
- Christian Szegedyv3Christian Szegedy is a Hungarian mathematician and machine learning researcher best known for three foundational deep learning contributions from his decade at...
- Mike Kriegerv2Mike Krieger is a Brazilian-American software engineer and product executive who co-founded Instagram with Kevin Systrom in 2010 and now works at the AI safety...
- Marius Hobbhahnv2Marius Hobbhahn is a German AI safety researcher and the co-founder and chief executive officer of Apollo Research, a London-based organization that studies...
- Ego4Dv2Ego4D is a large-scale egocentric (first-person) video dataset and benchmark suite for computer vision, assembled by Meta AI (then Facebook AI Research)...
- Convictionv2Conviction is an early-stage venture capital firm founded in 2022 by Sarah Guo, a former general partner at Greylock, that invests exclusively in companies...
- Project Rainierv2Project Rainier is a distributed artificial intelligence supercomputer built by Amazon Web Services around its in-house AWS Trainium accelerators, created...
- Jerry Tworekv2Jerry Tworek (full name Jaroslaw Tworek) is a Polish artificial intelligence researcher who served as vice president of research at OpenAI and led the...
- Falcon-H1v2Falcon-H1 is a family of open-weight large language models released in 2025 by the Technology Innovation Institute (TII), the applied-research arm of Abu...
- Brian Schimpfv2Brian Schimpf is the co-founder and chief executive officer of Anduril Industries, the defense AI and autonomous-systems company he started in 2017 with Palmer...
- AMI Labsv2AMI Labs (Advanced Machine Intelligence) is a Paris-based artificial intelligence research company founded by Yann LeCun, the Turing Award winner and former...
- Ali Ghodsiv2Ali Ghodsi (born 1978) is the co-founder and chief executive officer of Databricks, one of the most valuable private companies in the world and a leading...
- The Stack v2v2The Stack v2 is a large open dataset of source code released by BigCode in February 2024 as the training dataset behind the StarCoder2 family of code models....
- Project Astrav2Project Astra is a research prototype from Google DeepMind that explores what a universal AI assistant might look like: a single agent that can see and hear...
- Perception Encoderv2Perception Encoder (PE) is a family of vision and vision-language encoders from Meta AI's Fundamental AI Research (FAIR) group, released in April 2025, whose...
- Model welfarev2Model welfare is the research area that investigates whether advanced AI systems might have morally relevant experiences or interests, such as suffering or...
- Jian Sunv2Jian Sun (1976 to 2022) was a Chinese computer scientist and one of the most influential researchers in modern computer vision. He is best known as a...
- Meta AI Studiov2Meta AI Studio (often shortened to AI Studio) is a free platform from Meta that lets anyone build customizable AI characters and personas without writing code,...
- GenAI-Benchv2GenAI-Bench is an AI benchmark for evaluating compositional text-to-image and text-to-video generation, introduced in 2024 by researchers from Carnegie Mellon...
- Gemini Livev2Gemini Live is Google's natural-voice conversational mode in the Gemini app: a hands-free, interruptible spoken interface that lets a person talk to Google's...
- Ethan Mollickv2Ethan R. Mollick (born 1975) is an American academic and author who is an associate professor of management at the Wharton School of the University of...
- Emmett Shearv2Emmett Shear (born 1983) is an American entrepreneur, software engineer, and investor best known as a co-founder and longtime chief executive of the...
- Voice Engine (OpenAI)v2Voice Engine is a speech-generation and voice-cloning model developed by OpenAI that can produce natural-sounding speech resembling a specific person from a...
- Michael Truellv2Michael Truell (born September 4, 2000) is an American technology entrepreneur and the co-founder and chief executive officer of Anysphere, the San Francisco...
- MASKv2MASK (Model Alignment between Statements and Knowledge) is an AI safety benchmark that measures the honesty of large language models (LLMs) by testing whether...
- TPU v4v2TPU v4 is Google's fourth-generation Tensor Processing Unit, a custom application-specific integrated circuit (ASIC) that accelerates machine learning...
- Edwin Chenv2Edwin Chen is an American entrepreneur and machine learning engineer who is the founder and chief executive officer of Surge AI, a company that supplies...
- SWE-bench Multilingualv2SWE-bench Multilingual is an AI benchmark of 300 real-world software bug-fixing tasks drawn from 42 open-source repositories across nine programming languages,...
- Pathways (Google AI)v2Pathways is the name Google has used for two related but distinct things in its artificial-intelligence work: a research vision for a next-generation AI...
- OpenGVLabv2OpenGVLab is the open-source organization and project hub for general-vision and multimodal foundation models run by the General Vision group at Shanghai AI...
- Magenta (project)v2Magenta is an open-source research project from Google that explores the role of machine learning in creating art and music. Started by researchers and...
- InternVL3v2InternVL3 is an open-weights family of multimodal AI large language models released on April 11, 2025 by OpenGVLab, the general vision team associated with...
- OpenAI o1-prov2OpenAI o1-pro is the highest-compute variant of OpenAI's o1 reasoning model, designed to spend more inference-time compute so it "thinks harder" and returns...
- Lyria 2v2Lyria 2 is a high-fidelity, text-to-music generation model built by Google DeepMind that turns text prompts into professional-grade instrumental audio....
- Genie (DeepMind)v2Genie is an 11-billion-parameter generative interactive environment from Google DeepMind, described by its creators as the first foundation world model: it...
- Dreamer (reinforcement learning)v2Dreamer is a family of model-based reinforcement learning agents that learn a compact world model of their environment and then improve their behavior by...
- Cozev2Coze is a no-code and low-code platform from ByteDance for building, debugging, and deploying AI agents and chatbots without writing much code. Users assemble...
- Toy Models of Superpositionv2Toy Models of Superposition is a September 2022 mechanistic interpretability paper from Anthropic that shows how a neural network can represent more features...
- Pixtral Largev2Pixtral Large is a 124-billion-parameter multimodal (vision-language) large language model released by Mistral AI on November 18, 2024. It pairs a...
- Make-A-Videov2Make-A-Video is a text-to-video generation system from Meta AI, announced on September 29, 2022, whose defining contribution is generating moving video from a...
- Gemini 1.5 Flashv2Gemini 1.5 Flash is a lightweight, low-latency multimodal large language model from Google DeepMind, released at Google I/O on May 14, 2024, as the fast and...
- Denis Yaratsv2Denis Yarats (born 1987) is the co-founder and chief technology officer of Perplexity AI, the artificial intelligence company behind a conversational "answer...
- Gemma 3nv2Gemma 3n is an open, mobile-first multimodal model from Google built to run locally on phones, tablets, laptops, and other resource-constrained hardware,...
- Gemini 2.5 Deep Thinkv2Gemini 2.5 Deep Think is Google DeepMind's enhanced reasoning mode for the Gemini 2.5 Pro model that uses a technique called "parallel thinking" to explore...
- Collective Constitutional AIv2Collective Constitutional AI (CCAI) is a 2023 research project by Anthropic and the Collective Intelligence Project (CIP) that sourced the value principles, or...
- Claude Projectsv2Claude Projects is a feature of Claude, the chat assistant from Anthropic, that groups related conversations into a dedicated workspace with its own knowledge...
- Claude for Chromev2Claude for Chrome is an agentic browser extension from Anthropic that lets its Claude models perceive a Chrome browser window and take actions in it, including...
- Mistral Medium 3v5Mistral Medium 3 is a proprietary multimodal large language model developed by Mistral AI and released on May 7, 2025.[^1] Positioned as a mid-tier enterprise...
- Joelle Pineauv2Joelle Pineau (born 1974) is a Canadian computer scientist who is the first Chief AI Officer of Cohere, a professor and William Dawson Scholar at McGill...
- Claude Instantv2Claude Instant was Anthropic's fast, low-cost line of large language models, offered through the Anthropic API from March 2023 to November 2024 as the...
- AI Wikiv4AI Wiki (stylized AIWIKI.AI) is a community-driven online encyclopedia of artificial intelligence that is written and maintained by fans and enthusiasts of the...
- Programming with ChatGPTv4Programming with ChatGPT is the practice of using OpenAI's conversational chatbot to read, write, refactor, document, debug, test, and explain source code in...
- Paris AI Action Summitv6The Paris AI Action Summit (French: Sommet pour l'action sur l'IA) was the third major international summit on artificial intelligence in the so-called "AI...
- Mistral Large 3v6Mistral Large 3 is a sparse mixture-of-experts large language model released on December 2, 2025 by the French AI company Mistral AI, distributed as open...
- Healthv4AI in healthcare is the use of artificial intelligence, and especially machine learning, to diagnose disease, discover drugs, document care, predict patient...
- Ground Truthv6Ground truth is verified, correct information that serves as the authoritative reference for training and evaluating machine learning models. In supervised...
- Software Developmentv3AI in software development is the use of large language models and other machine learning systems to write, complete, review, test, document, and refactor...
- Minimum Viable Agentv3A Minimum Viable Agent (MVA) is the simplest valid implementation of an AI agent: a program in which a large language model calls tools in a loop, observes the...
- Liquid AIv5Liquid AI is an artificial intelligence company founded in 2023 by researchers from MIT CSAIL and headquartered at 314 Main Street in Cambridge,...
- Cohere Command Av5Cohere Command A is a 111 billion parameter dense large language model released by Cohere on March 13, 2025, built for enterprise agents, Retrieval-Augmented...
- Agent orchestrationv4Agent orchestration is the coordinated management of multiple AI agents working together as a unified system to accomplish tasks that exceed the capability of...
- SEOv3Search engine optimization (SEO) is the practice of preparing websites and other content so that search engines surface them in response to user queries. It...
- Neuromorphic computingv6Neuromorphic computing is brain-inspired computer hardware that processes information with spiking neural networks (SNNs) and event-driven, in-memory...
- ElevenLabs Musicv5ElevenLabs Music (also marketed as Eleven Music) is an AI music generation product developed by ElevenLabs, the voice AI company founded in 2022 by Piotr...
- EAGLE (speculative decoding)v3EAGLE (Extrapolation Algorithm for Greater Language-model Efficiency) is a lossless speculative decoding method that speeds up large language model (LLM)...
- ARIMAv3ARIMA (Autoregressive Integrated Moving Average) is a class of statistical models for analyzing and forecasting time series data, specified by three...
- Krea AIv9Krea AI (legally Krea, krea.ai) is a San Francisco generative AI company, founded in 2022, that builds a browser-based creative suite for image generation,...
- Diffusion Language Modelsv5Diffusion language models (DLMs, sometimes written dLLMs at frontier scale) are text generators that synthesize a sequence by iteratively denoising or...
- Businessv4Business uses of artificial intelligence are the application of machine learning, generative AI, and AI agents inside companies and other organizations to...
- Language Learningv3Language learning is one of the first consumer markets to be reshaped by generative AI: since 2023, apps such as Duolingo, Speak, Babbel, ELSA, and Pimsleur...
- DOBOT Atomv7The DOBOT Atom is a full-size humanoid robot developed by DOBOT Robotics (formally Shenzhen Yuejiang Technology Co., Ltd.), priced at approximately $27,500...
- On the Biology of a Large Language Modelv3On the Biology of a Large Language Model is a mechanistic interpretability paper published by Anthropic on March 27, 2025, in the Transformer Circuits...
- AI in Educationv4AI in education refers to the use of artificial intelligence technologies to support teaching, learning, assessment, and administrative processes across...
- AdvBenchv5AdvBench (Adversarial Behavior Benchmark) is a red-teaming benchmark dataset for measuring how easily an aligned large language model can be pushed into...
- Text Generation Modelsv5Text generation models are language models trained to produce coherent natural-language text by predicting tokens one at a time, each conditioned on the...
- PubMedQAv5PubMedQA is a biomedical question answering dataset and benchmark that evaluates whether machine learning models can answer yes/no/maybe research questions...
- Consistency Modelsv3Consistency models are a family of generative models, introduced by Yang Song, Prafulla Dhariwal, Mark Chen, and Ilya Sutskever at OpenAI in March 2023, that...
- Musicv3AI in music is the use of artificial intelligence, especially machine learning and generative AI, to compose, perform, mix, master, transcribe, voice-clone,...
- MGSM (Multilingual Grade School Math)v5Multilingual Grade School Math Abbreviation A multilingual benchmark evaluating mathematical reasoning across 10 typologically diverse languages using...
- Linear Probesv4A linear probe is a small linear classifier (or linear regressor) trained on the frozen internal activations of a neural network to test whether a particular...
- Landmarksv6Landmarks are reference points used as anchors in two largely separate areas of machine learning. In manifold learning and dimension reduction, landmarks are a...
- DeepSeek-R1-Distillv3DeepSeek-R1-Distill is a family of six open-weight reasoning language models released by DeepSeek on January 20, 2025, alongside the flagship DeepSeek-R1...
- Voice Activity Detection Modelsv5Voice activity detection (VAD), also called speech activity detection (SAD), is the task of deciding which segments of an audio signal contain human speech and...
- Translational invariancev5Translational invariance (also called translation invariance or shift invariance) is the property of a function, system, or machine learning model whose output...
- Out-Group Homogeneity Biasv5Out-group homogeneity bias, also called the out-group homogeneity effect, is the cognitive bias in which people perceive members of an out-group as more...
- Indirect prompt injectionv3Indirect prompt injection is a class of attack against large language model-integrated applications in which the malicious instructions that subvert the model...
- Financev4Artificial intelligence in finance is the use of machine learning, deep learning, and generative AI across banking, capital markets, insurance, fintech, and...