Skip to content
AI Wiki
CtrlK
LearnExploreToolsUpdatesReading list

Explore AI Wiki

Loading

AI Wiki site footer

Browse by topic

All categoriesRandom article
  • Machine Learning
  • AI Companies
  • Large Language Models
  • Robotics
  • Open Source AI
  • AI Models
  • Deep Learning
  • Humanoid Robots
  • AI Hardware
  • Generative AI

A free, source-backed encyclopedia with 4,000+ articles about artificial intelligence.

Help keep AI knowledge accurate

Start contributing

Contribute

  • Recent changes
  • Requested articles
  • Missing pages
  • Corrections log

Standards & trust

  • About AI Wiki
  • How we verify
  • Sourcing standards
  • AI transparency
  • Neutral point of view
  • Editorial policy
  • Content license

Tools & data

  • Free AI tools
  • AI comparisons
  • API, MCP & open data
  • Site statistics
  • RSS feed

From the AI Wiki team

  • AI Compute TrackerGPU cloud pricing, availability, and compute-market data.New tab ↗AI Compute Tracker is a companion site owned and operated by the same team as AI Wiki. Opens in a new tab.
How companion projects work

AIWiki.ai · Text is available under CC BY 4.0; reuse welcome.

  • Contact
  • Privacy
  • Terms

Recent changes

RSS

4,546 articles updated. New pages start at v1; higher version numbers mean an existing article was revised. Page 1 of 46.

Wednesday, September 16, 2026

  • Calibration (machine learning)v8Calibration in machine learning is the property that the probability scores produced by a probabilistic classifier match the empirical frequency of the predicted event: a model that assigns a confidence of 0.8…
  • Structured outputv10Structured output is a set of techniques and API features that constrain a large language model (LLM) to emit responses that exactly conform to a predefined format or schema, such as JSON, XML, or a custom…
  • InstructGPTv7InstructGPT is a family of language models released by OpenAI in January 2022 that take the base GPT-3 and fine-tune it to follow user instructions more helpfully, truthfully, and with less toxic output, using…
  • TypeSafe AIv2TypeSafe AI (styled TypeSafe) is a San Francisco AI lab founded in 2024 that builds what it calls "machine-native
  • Jev (AI model)v2Jev is an AI model released in early access on 15 September 2026 by TypeSafe AI, a San Francisco startup founded in 2024 by former OpenAI researcher Diogo Almeida with Erik Gafni and Sasha Sheng.
  • Diogo Almeidav2Diogo Almeida (full name Diogo Moitinho de Almeida) is a machine learning researcher and the co-founder and chief executive of TypeSafe AI, a San Francisco lab that came out of stealth on 15 September 2026…
  • Long-context language modelsv8Long-context language models are large language models engineered to accept and reason over inputs far larger than the few-thousand-token windows used by early transformer systems, with frontier models in 2026…
  • MiniCPMv4MiniCPM is a family of compact, openly licensed language models published by OpenBMB, the shared open-source brand of Tsinghua University's natural language processing lab (THUNLP) and the Beijing company…
  • DeepSeek Sparse Attention (DSA)v5DeepSeek Sparse Attention (DSA) is a trainable, fine-grained sparse attention mechanism introduced by the Chinese AI company DeepSeek in its experimental model DeepSeek-V3.2-Exp, released on September 29, 2025.
  • Song Hanv4Song Han is a computer scientist and an associate professor with tenure in the Department of Electrical Engineering and Computer Science at the Massachusetts Institute of Technology (MIT)
  • Sparse attentionv9Sparse attention is a family of techniques that cut the computational and memory cost of the attention mechanism in transformer models by letting each token attend to only a subset of other tokens in a sequence
  • SparDAv2SparDA (Sparse Decoupled Attention) is an add-on architecture for long-context large language model inference proposed by researchers at NVIDIA in a paper posted to arXiv on 3 June 2026.
  • NOSA (Native and Offloadable Sparse Attention)v2NOSA (Native and Offloadable Sparse Attention) is a trainable sparse attention mechanism designed so that most of a language model's KV cache can live in CPU memory during decoding without the CPU-to-GPU…
  • KV cache offloadingv2KV cache offloading is the practice of moving part or all of a Transformer model's KV cache out of accelerator memory (GPU HBM) into a larger, slower tier such as CPU DRAM, local NVMe storage, or remote…
  • Tencent AIv11Tencent AI is the artificial intelligence research, products, and services developed by Tencent Holdings Ltd., one of the world's largest technology companies, built around the Hunyuan family of foundation…
  • BrowserSkillv2BrowserSkill is an open-source browser-automation bridge from Tencent that lets shell-capable AI coding agents drive the user's own, already logged-in Chrome or Microsoft Edge browser.
  • AI browser agentv9An AI browser agent is a class of AI agent that uses artificial intelligence to autonomously navigate, interpret, and interact with web browsers to complete tasks on behalf of a user.
  • AlphaEvolvev6AlphaEvolve is an evolutionary coding agent developed by Google DeepMind and announced on May 14, 2025 .
  • KernelBenchv4KernelBench is an AI benchmark and open-source evaluation environment that measures how well large language models can write fast and correct GPU kernels.
  • Dream-RSIv2Dream-RSI (Recursive Self-Improvement through Evolving Worlds) is a framework for improving the exploration strategy of an AI discovery agent, described in a preprint posted to arXiv on 14 September 2026 by a…
  • DeepSeek Harnessv3DeepSeek Harness is an open-source AI agent harness developed by DeepSeek. Also called dsh, it supplies the software around a language model: the agent loop, tools, sessions, filesystems, permission controls…
  • Recursive self-improvementv14Recursive self-improvement (RSI) is a process in which an artificial intelligence system improves its own intelligence or its ability to improve itself, so that each enhancement increases its capacity for…
  • KV Cachev13KV cache, short for key-value cache, is transient model state used during Transformer generation.
  • LeRobotv9LeRobot is an open-source, PyTorch-native library for end-to-end real-world robot learning developed by Hugging Face and a global community of contributors, often described as the "Transformers for robotics"…
  • Robot foundation modelv9A robot foundation model is a large-scale machine learning model, typically based on the transformer architecture, that is pre-trained on broad, diverse datasets of robot interactions and then adapted to a…
  • Flexion Roboticsv4Flexion Robotics AG, usually shortened to Flexion, is a Swiss robotics software company that builds a hardware-agnostic autonomy stack, often described as "the brain," for humanoid robots.
  • AWS Trainiumv7AWS Trainium is a family of custom machine learning accelerator chips designed by Annapurna Labs for Amazon Web Services, purpose-built for training and, increasingly, for serving large neural networks.
  • NVentures (Nvidia)v5NVentures is the corporate venture-capital arm of Nvidia, the dominant supplier of graphics processors used to train and run artificial-intelligence models.
  • Jeff Hawkev2Jeff Hawke (published as Jeffrey Hawke) is a New Zealand robotics and machine learning engineer who is the co-founder and chief technology officer of Odyssey
  • reBot Armv3reBot Arm is an open-hardware desktop robotic-arm series developed by Seeed Studio for robotics education, teleoperation, and embodied-AI experiments.
  • OpenArmv2OpenArm is an open-source, seven-degree-of-freedom humanoid robot arm developed by Enactic, Inc., a Tokyo robotics company.
  • SO-101v2The SO-101 (Standard Open Arm 101) is a low-cost, open-source, 3D-printable robot arm designed by The Robot Studio in collaboration with Hugging Face and published in the TheRobotStudio/SO-ARM100 repository…
  • Nemotron 3v14Nemotron 3 is a family of open-weights large language model systems released by NVIDIA beginning on December 15, 2025, built for agentic AI and consisting of three sparse mixture-of-experts variants named…
  • NVIDIA NIMv8NVIDIA NIM (NVIDIA Inference Microservices) is a set of containerized, prebuilt-and-optimized model-serving microservices from NVIDIA that package an AI model, an optimized inference engine, and an…
  • Reward AIv3Reward AI is a robotics company that builds general-purpose manipulation policies meant to run on many different robot bodies . The company gives its location as the Bay Area, California, on its X account .
  • OM-1 (Omnibody Model 1)v3OM-1, short for Omnibody Model 1, is a general-purpose robot manipulation policy announced by Reward AI on 14 September 2026 .
  • Odyssey-2v2Odyssey-2 is a family of real-time, interactive generative models built by the AI lab Odyssey.
  • Odyssey-3v2Odyssey-3 is a foundation world model announced on September 15, 2026 by Odyssey, an AI lab founded in 2023 by Oliver Cameron and Jeff Hawke.
  • Odyssey (AI lab)v2Odyssey is an artificial intelligence lab that builds what it calls foundation world models: causal, action-conditioned systems trained to predict how the world evolves from one moment to the next, and to…
  • World modelv9A world model is an artificial intelligence system that learns an internal representation of how an environment works, enabling it to predict future states, simulate the consequences of actions, and support…
  • CaliBenchv2CaliBench is a benchmark for image-to-video generative models that asks whether a model reproduces the correct distribution of physical outcomes across many generations from the same starting frame
  • Oliver Cameronv2Oliver Cameron is a technology entrepreneur and the co-founder and chief executive of Odyssey, an AI lab that builds general-purpose world models .
  • NexArmnewNexArm is a desktop robotic arm sold by Hiwonder, a Shenzhen-based maker of educational robotics kits, for imitation-learning, teleoperation and embodied-AI coursework.
  • Starchild-1v2Starchild-1 is a real-time audio-video world model built by the AI lab Odyssey. It generates synchronized video and sound autoregressively, chunk by chunk, while a user streams new text, speech and action…
  • PROWL-1v2PROWL-1 is a training framework from the AI lab Odyssey in which a reinforcement learning agent is paid to break a world model.
  • Agora-1v2Agora-1 is a multi-agent world model released by the AI lab Odyssey on 18 May 2026.

Tuesday, September 15, 2026

  • d-Matrixv2d-Matrix is a privately held American semiconductor company headquartered in Santa Clara, California that builds accelerators, I/O cards and software for AI inference in data centers.
  • NVIDIA Vera Rubinv15NVIDIA Vera Rubin is NVIDIA's data center AI computing platform that succeeds the NVIDIA Blackwell architecture, pairing the custom Arm-based Vera CPU with the dual-die Rubin GPU into a single superchip built…
  • CUDAv13CUDA (Compute Unified Device Architecture) is NVIDIA's platform and programming model for general-purpose computation on its graphics processing units.
  • YMTC (Yangtze Memory Technologies)v5Yangtze Memory Technologies Co., Ltd. (YMTC, Chinese name 长江存储) is a Chinese memory chipmaker based in Wuhan that designs and manufactures 3D NAND flash memory.
  • Taalasv3Taalas is a Toronto, Canada based AI chip startup that builds "hardwired" or model-specific silicon

Monday, September 14, 2026

  • Zeroth Roboticsv9Zeroth Robotics (stylized as zeroth) is a robotics company developing interactive artificial intelligence robots for consumer and commercial markets. Zeroth is a brand rather than the legal entity.
  • Zeroth M1v10The Zeroth M1 is a compact humanoid robot developed by Zeroth Robotics, an AI robotics company operating under Suzhou JoyIn Intelligent Technology Co.
  • An Alien Mindv3An Alien Mind is an essay by Jakub Pachocki, the chief scientist of OpenAI, published by OpenAI on September 6, 2026 under its Safety and Research sections.
  • Aether (JoyIn)v2Aether (Chinese: 以太) is a control model for humanoid robots published by the Chinese robotics company JoyIn, which trades under the registered entity 苏州乐享智能科技有限公司 and sells home robots under the Zeroth brand
  • Insilico Medicinev5

Sunday, September 13, 2026

  • NSA, CISA and FBI Distillation Advisory (AA26-251A)newOn September 8, 2026 the National Security Agency (NSA), the Cybersecurity and Infrastructure Security Agency (CISA) and the Federal Bureau of Investigation (FBI) released a joint cybersecurity advisory, alert…
  • Discovery Loopv3Discovery Loop is an American artificial intelligence startup founded by Jeff Dean, Sanjay Ghemawat, Oriol Vinyals, and Quoc Le, four of the most influential researchers and engineers in the history of Google…
  • Terence Taov2Terence Tao (born 1975 in Adelaide, Australia) is a mathematician at the University of California, Los Angeles, where he is a Distinguished Professor and holds the James and Carol Collins Chair in the College…

Friday, September 11, 2026

  • DeepSeek V3.2v8DeepSeek-V3.2 is an open-weight Mixture of Experts large language model family developed by DeepSeek that introduces DeepSeek Sparse Attention (DSA)
  • DeepSeek V4v15DeepSeek V4 is a family of open-weight Mixture of Experts large language models developed by DeepSeek, a Hangzhou-based AI research lab.
  • Multi-head Latent Attentionv10Multi-head Latent Attention (MLA) is an attention mechanism for transformer models that achieves a 93.3% reduction in key-value cache size while maintaining or exceeding the performance of traditional…
  • DeepSeek V4.1-Flashv3DeepSeek V4.1-Flash is an open-weight multimodal mixture-of-experts model released by DeepSeek on September 10, 2026. It accepts text and images and generates text.
  • DeepSeek V4-Prov6DeepSeek V4-Pro is the flagship model of the DeepSeek V4 family: a Mixture of Experts large language model with 1.6 trillion total parameters, 49 billion of them activated per token, and a one-million-token…
NewerPage 1 of 46Older
  • SRAMv3SRAM, or static random-access memory, is a semiconductor memory that stores each bit in a latch built from cross-coupled inverters.
  • SMICv6Semiconductor Manufacturing International Corporation (SMIC) is China's largest contract chipmaker and the most advanced pure-play semiconductor foundry on the mainland
  • Rain AIv4Rain AI, originally incorporated as Rain Neuromorphics, is a San Francisco AI accelerator startup founded by Gordon Hirsch Wilson, Jack Kendall, and Juan Claudio Nino, with technical roots in the University of…
  • NAURA Technology Groupv5NAURA Technology Group Co., Ltd. (北方华创科技集团股份有限公司, romanised as Beifang Huachuang and rendered in English sources as both "NAURA" and "Naura Technology Group") is China's largest maker of semiconductor…
  • CXMT (ChangXin Memory Technologies)v4ChangXin Memory Technologies, almost always shortened to CXMT, is China's largest maker of dynamic random-access memory and the only Chinese memory company big enough to appear in global DRAM market-share…
  • SpaceX-xAI mergerv6The SpaceX-xAI merger was an all-stock acquisition, announced on February 2, 2026, in which the rocket and satellite company SpaceX acquired the artificial intelligence developer xAI
  • SWE-benchv15SWE-bench is an execution-based benchmark for evaluating whether a language-model system can resolve real software issues.
  • xAIv18xAI is an artificial intelligence business established in 2023 under Elon Musk, who led it before the SpaceX acquisition.
  • Grok 4.6v4Grok 4.6 is a proprietary large language model and reasoning model in the Grok family, developed by SpaceXAI and released jointly with Cursor through the xAI API on August 12, 2026.
  • China's semiconductor industryv6China's semiconductor industry is the network of wafer fabrication plants, equipment and materials suppliers, chip design houses, packaging and test operations, and state investment vehicles that produce…
  • Empyrean Technologyv2Empyrean Technology (Chinese: 北京华大九天科技股份有限公司, Beijing Huada Jiutian Technology Co., Ltd.; trading as Empyrean) is a Chinese electronic design automation (EDA) software company headquartered in Beijing and…
  • Task-completion time horizon (METR)v4The task-completion time horizon is a metric for AI capability proposed by METR that expresses a model's ability in units of human time: it is the length of task
  • METRv15METR (Model Evaluation and Threat Research) is a nonprofit research organization based in Berkeley, California, that develops scientific methods for measuring the autonomous capabilities of frontier AI systems…
  • d-Matrix Corsairv6Corsair is the first commercial AI accelerator product from d-Matrix, a Silicon Valley AI inference hardware startup based in Santa Clara, California.
  • NVLink Fusionv8NVLink Fusion is a program and silicon technology from NVIDIA that opens its NVLink high speed interconnect to third party chips.
  • NVIDIA MGXv2NVIDIA MGX is a modular reference architecture that NVIDIA publishes so that server makers and contract manufacturers can build accelerated systems around NVIDIA GPUs, CPUs, DPUs and networking without…
  • d-Matrix Raptorv2Raptor is the second-generation AI inference accelerator from d-Matrix, a Santa Clara semiconductor startup, and the first commercial chip built on the company's 3D stacked digital in-memory compute technology
  • Specific Labsv2Specific Labs is a San Francisco company that buys and licenses operational data and source code from businesses, and packages it as training and evaluation material for AI labs.
  • Real-SWEv2Real-SWE is a coding-agent benchmark published in September 2026 by Specific Labs, a San Francisco company that licenses operational data and source code from businesses and packages them as training and…
  • Insilico Medicine is a clinical-stage biotechnology company that uses generative AI to discover disease targets and design small-molecule drugs end to end.
  • Rentosertibv2Rentosertib is an investigational oral small-molecule inhibitor of TNIK (TRAF2- and NCK-interacting kinase) developed by Insilico Medicine for idiopathic pulmonary fibrosis (IPF)
  • Extended thinkingv4Extended thinking is the product name Anthropic gives to the reasoning mode in its Claude family of large language models
  • Anthropicv24Anthropic is an American artificial intelligence (AI) safety and research company founded in 2021 by Dario Amodei, Daniela Amodei, and other former OpenAI researchers, best known for the Claude family of large…
  • Google Glassv2Google Glass was a head-mounted display built by Google, first shown publicly on 2012-04-04 as "Project Glass" and sold from April 2013 as the Glass Explorer Edition at $1,500 .
  • Waymov6Waymo LLC is an American autonomous driving technology company and subsidiary of Alphabet Inc. that develops the Waymo Driver, a Level 4 autonomous driving system, and operates Waymo One
  • How to Pressure LLMs for Better Outputv6Pressuring large language models (LLMs) is a family of prompt engineering techniques that try to push a model toward better output by adding emotional weight, urgency, stakes, or coercion to the prompt.
  • Bardv9Bard was the conversational AI chatbot developed by Google, launched as an experimental service on February 6, 2023, opened to a public waitlist on March 21, 2023, and rebranded to Gemini on February 8, 2024.
  • GPT-6 Astrav5GPT-6 Astra is an OpenAI model that the company began rolling out on September 3, 2026, describing it as "the world's most intelligent and aligned model" and as the successor to the GPT-5.6 family.
  • Sergey BrinnewSergey Brin is an American computer scientist who co-founded Google with Larry Page in 1998 and who, since 2023, has worked hands-on inside the company on its Gemini models without holding any executive title.
  • Mathematical reasoning in AIv5Mathematical reasoning in AI is the ability of computer systems to solve mathematical problems: carrying out multi-step calculations, proving theorems, and answering competition or research questions that…
  • Leiden Declaration on Artificial Intelligence and Mathematicsv2The Leiden Declaration on Artificial Intelligence and Mathematics is a statement on the use of artificial intelligence in mathematical research, published on 2 June 2026 at leidendeclaration.ai and deposited…
  • A Severe Misalignment of AI in Mathematicsv2A Severe Misalignment of AI in Mathematics is a declaration about artificial intelligence and mathematical research, published on September 11, 2026 at mathandai.org with 25 initial signatories
  • OpenAI Navier-Stokes Proposed Solutionv3OpenAI Navier-Stokes Proposed Solution is a proposed resolution of the Navier-Stokes existence and smoothness Millennium Prize Problem announced by OpenAI on September 8, 2026.
  • DeepSeek V4-Flashv12DeepSeek V4-Flash is the smaller of the two large language models in the DeepSeek V4 family, a 284-billion-parameter Mixture of Experts model with 13 billion active parameters and a one-million-token context…
  • DeepSeekv12DeepSeek is a Chinese artificial intelligence company based in Hangzhou. It was founded in 2023 by Liang Wenfeng, who had previously co-founded the quantitative investment firm High-Flyer.
  • GPT-Realtime / OpenAI Realtime APIv7GPT-Realtime is a family of speech-to-speech models exposed through the OpenAI Realtime API, a low-latency interface that lets developers build voice agents which take spoken audio in and return spoken audio…
  • OpenAI Realtime APIv6The OpenAI Realtime API is a speech-to-speech interface from OpenAI that lets developers build low-latency, bidirectional voice AI applications powered by GPT-4o and later by the purpose-built gpt-realtime…