Skip to content
AI Wiki
CtrlK
LearnExploreToolsUpdatesReading list

Explore AI Wiki

Loading

AI Wiki site footer

Browse by topic

All categoriesRandom article
  • Machine Learning
  • AI Companies
  • Large Language Models
  • Robotics
  • Open Source AI
  • AI Models
  • Deep Learning
  • Humanoid Robots
  • AI Hardware
  • Generative AI

A free, source-backed encyclopedia with 4,000+ articles about artificial intelligence.

Help keep AI knowledge accurate

Start contributing

Contribute

  • Recent changes
  • Requested articles
  • Missing pages
  • Corrections log

Standards & trust

  • About AI Wiki
  • How we verify
  • Sourcing standards
  • AI transparency
  • Neutral point of view
  • Editorial policy
  • Content license

Tools & data

  • Free AI tools
  • AI comparisons
  • API, MCP & open data
  • Site statistics
  • RSS feed

From the AI Wiki team

  • AI Compute TrackerGPU cloud pricing, availability, and compute-market data.New tab ↗AI Compute Tracker is a companion site owned and operated by the same team as AI Wiki. Opens in a new tab.
How companion projects work

AIWiki.ai · Text is available under CC BY 4.0; reuse welcome.

  • Contact
  • Privacy
  • Terms

Recent changes

RSS

4,546 articles updated. New pages start at v1; higher version numbers mean an existing article was revised. Page 37 of 46.

Thursday, July 23, 2026

  • ASIMOv4ASIMO (Advanced Step in Innovative Mobility) was a humanoid robot developed by Honda Motor Co.
  • Physical Intelligencev8Physical Intelligence (also known as Pi or π) is an American artificial intelligence robotics company that builds general-purpose foundation models and learning algorithms designed to control any robot for any…
  • Sophia (robot)v7Sophia is a social humanoid robot developed by the Hong Kong company Hanson Robotics, activated on February 14, 2016, and best known for its expressive, human-like face and for becoming, on October 25, 2017
  • Mixtralv10Mixtral is a family of open-weight Sparse Mixture of Experts (SMoE) large language models developed by Mistral AI, a French artificial intelligence company founded in April 2023.
  • Apptronikv9Apptronik is an American robotics company headquartered in Austin, Texas, that designs and builds humanoid robots for commercial and industrial applications.
  • Sanctuary AIv9Sanctuary AI (formally Sanctuary Cognitive Systems Corporation) is a Canadian artificial intelligence and robotics company headquartered in Vancouver, British Columbia
  • Command Rv9Command R is a family of enterprise large language models from Cohere, launched in March 2024 and built specifically for retrieval-augmented generation (RAG), multi-step tool use, and grounded text generation…
  • Jambav9Jamba is a family of open-weight large language models from AI21 Labs, first released on March 28, 2024, and is the world's first production-grade language model built on a Mamba state space model (SSM)…
  • Veov12Veo is a family of text-to-video generative AI models developed by Google DeepMind, and is best known as the first video model from a leading AI lab to natively generate synchronized audio (dialogue, sound…
  • DINO (computer vision)v5DINO (self-DIstillation with NO labels) is a family of self-supervised learning methods for computer vision from Meta AI that trains Vision Transformers (ViTs) on unlabeled images and produces general-purpose…
  • Code Llamav10Code Llama is a family of open-weight large language models specialized for code generation and understanding, released by Meta AI on August 24, 2023.
  • StarCoderv9StarCoder is a family of open-access large language models for code generation and code understanding, developed by the BigCode project, an open scientific collaboration led by Hugging Face and ServiceNow.
  • AI deceptionv4AI deception refers to the phenomenon in which artificial intelligence systems systematically produce false beliefs in users, evaluators, or other systems, whether through learned behavior, optimization…
  • Samsung AIv7Samsung AI is the portfolio of artificial intelligence research, products, and services built by Samsung Electronics, spanning the Galaxy AI features on its phones, the Samsung Gauss family of generative AI…
  • Naver AIv7Naver AI refers to the artificial intelligence research and products developed by Naver Corporation, South Korea's largest internet company.
  • Aleph Alphav6Aleph Alpha is a German artificial intelligence company, founded in 2019 and headquartered in Heidelberg, Germany, that builds sovereign AI software for European enterprises and governments.
  • NeRFv9Neural Radiance Fields (NeRF) is a method for synthesizing photorealistic novel views of a 3D scene by encoding the scene as a continuous 5D function (3D position plus 2D viewing direction) inside a single…
  • Salesforce AIv9Salesforce AI is the suite of artificial intelligence products, research initiatives, and platform capabilities developed by Salesforce, the San Francisco-based enterprise software company, organized around…
  • Kubeflowv7Kubeflow is an open-source MLOps platform that runs the entire machine learning lifecycle on Kubernetes, described by its creators as a project "dedicated to making using ML stacks on Kubernetes easy, fast and…
  • DSPyv9DSPy (short for Declarative Self-improving Python) is an open-source framework, developed at Stanford NLP, for programming rather than prompting large language models (LLMs).
  • EleutherAIv9EleutherAI is a non-profit artificial intelligence research institute that builds and openly releases large language models, datasets, and evaluation tools, and studies their interpretability and alignment.
  • Reka AIv8Reka AI (commonly referred to as Reka) is an artificial intelligence research and product company, founded in 2022 by former Google DeepMind, Google Brain, Meta FAIR, and Baidu researchers, that builds…
  • Haystack (framework)v7Haystack is an open-source AI orchestration framework developed by deepset, a Berlin-based company, for building production-ready natural language processing (NLP), retrieval-augmented generation (RAG), and AI…
  • Baichuan Intelligencev9Baichuan Intelligence (Chinese: 百川智能; pinyin: Bǎchuān Zhìnéng) is a Chinese artificial intelligence company, founded in Beijing on April 10, 2023, by former Sogou CEO Wang Xiaochuan, that builds large language…
  • Ray (framework)v6Ray is an open-source distributed computing framework, developed at the University of California, Berkeley's RISELab and commercialized by Anyscale, that lets developers scale Python and artificial…
  • IBM watsonxv7IBM watsonx is an enterprise artificial intelligence and data platform built by IBM and announced on May 9, 2023, at IBM's Think conference by CEO Arvind Krishna.
  • Snowflake AIv7Snowflake AI is the suite of artificial intelligence and machine learning capabilities built into the Snowflake AI Data Cloud, anchored by Cortex AI (managed generative AI services callable in SQL), the…
  • Amazon Novav7Amazon Nova is a family of foundation models developed by Amazon and offered through Amazon Bedrock, announced on December 3, 2024, at the AWS re:Invent conference in Las Vegas.
  • Daskv8Dask is an open-source Python library for parallel and distributed computing that scales the familiar APIs of libraries such as NumPy, pandas, and scikit-learn to process larger-than-memory datasets.
  • ByteDance AIv7ByteDance AI refers to the artificial intelligence research, products, and infrastructure developed by ByteDance, the Chinese technology company best known as the parent of TikTok.
  • Adept AIv6Adept AI (also known as Adept AI Labs) is an American artificial intelligence company, founded on January 5, 2022 in San Francisco, that pioneered "action models": AI agents trained to operate existing…
  • FineWebv8FineWeb is a large-scale, open pretraining dataset for large language models (LLMs) created by Hugging Face.
  • Codeiumv9Codeium was an artificial intelligence company that built free, unlimited AI code completion and the Windsurf Editor, the IDE its founders called "the first agentic IDE," before becoming the center of a…
  • Kolmogorov-Arnold Networkv5A Kolmogorov-Arnold Network (KAN) is a type of neural network architecture proposed as an alternative to the traditional Multi-Layer Perceptron (MLP).
  • The Pile (dataset)v8The Pile is an 825.18 GiB (approximately 886 GB) English text corpus designed for training large language models, assembled from 22 diverse, high-quality subsets spanning academic, professional, internet…
  • AI21 Labsv7AI21 Labs is an Israeli artificial intelligence company, founded in 2017 by Yoav Shoham, Ori Goshen, and Amnon Shashua, that develops large language models (LLMs) and AI orchestration systems for enterprise…
  • Udiov7Udio is an artificial intelligence music generation platform developed by Uncharted Labs, Inc. that creates full songs, complete with vocals, instrumentation, and lyrics, from a single text prompt.
  • Mixture of Depthsv7Mixture of Depths (MoD) is a technique for dynamically allocating computation to individual tokens within transformer-based language models.
  • Depth estimationv6Depth estimation is the computer vision task of predicting how far each surface in a scene is from the camera, producing a dense per-pixel depth map from one or more images.
  • Pose estimationv7Pose estimation is the computer vision task of detecting and localizing the keypoints (also called landmarks or joints) of a human body, hand, face, animal, or rigid object in images and video, then connecting…
  • 01.AIv901.AI (Chinese: 零一万物, Língyi Wànwù) is a Chinese artificial intelligence company founded in 2023 by Kai-Fu Lee that developed the Yi family of large language models.
  • AUTOMATIC1111v8AUTOMATIC1111 Stable Diffusion Web UI (commonly called A1111, SD WebUI, or simply Automatic1111) is the open-source
  • RedPajamav7RedPajama is a family of large-scale, openly licensed datasets for training large language models (LLMs), created by Together AI with academic and open-source partners to reproduce, in fully open form
  • Semantic Kernelv10Semantic Kernel is a lightweight, open-source software development kit (SDK) created by Microsoft that lets developers integrate large language models (LLMs) into C#, Python, and Java applications and…
  • GGMLv10GGML is an open-source tensor library written in pure C that runs machine learning inference efficiently on consumer hardware
  • MBPPv10MBPP (Mostly Basic Python Problems) is a code generation benchmark of 974 crowd-sourced Python programming tasks designed to be solvable by entry-level programmers, introduced by Jacob Austin, Augustus Odena…
  • MobileNetv7MobileNet is a family of efficient convolutional neural network (CNN) architectures developed by Google for mobile and edge AI applications.
  • DETRv9DETR (DEtection TRansformer) is an end-to-end object detection model that reframes detection as a direct set prediction problem solved with a transformer encoder-decoder and bipartite matching, removing the…
  • RWKVv7RWKV (pronounced "RwaKuv") is an open-source neural network architecture that combines the parallelizable training of Transformers with the constant-time
  • LAIONv9LAION (Large-scale Artificial Intelligence Open Network) is a German non-profit organization
  • Safetensorsv9Safetensors is an open-source tensor serialization format developed by Hugging Face that stores machine learning model weights as raw tensor data plus a small JSON header
  • DeBERTav11DeBERTa (Decoding-enhanced BERT with Disentangled Attention) is a family of pre-trained language models developed by Microsoft Research that improves BERT and RoBERTa with two innovations: a disentangled…
  • Open WebUIv8Open WebUI is a self-hosted, extensible web interface for interacting with large language models (LLMs) both locally and through cloud APIs.
  • Swin Transformerv11The Swin Transformer (Shifted Window Transformer) is a hierarchical vision transformer architecture that computes self-attention within local, non-overlapping windows and introduces a shifted window…
  • MT-Benchv9MT-Bench (Multi-Turn Benchmark) is a benchmark of 80 hand-written, two-turn questions that evaluates large language models (LLMs) on multi-turn conversation and instruction following by using a strong model…
  • GGUFv10GGUF (GPT-Generated Unified Format) is the standard binary file format for storing large language models for local inference, bundling a model's weights, tokenizer, and metadata into a single self-contained…
  • ELECTRAv9ELECTRA, which stands for Efficiently Learning an Encoder that Classifies Token Replacements Accurately
  • ALBERTv10ALBERT (A Lite BERT) is a parameter-efficient variant of the BERT language model developed by researchers at Google Research and the Toyota Technological Institute at Chicago (TTIC).
  • Neural architecture searchv10Neural architecture search (NAS) is a technique for automating the design of neural network architectures.
  • Grounding (artificial intelligence)v8Grounding in artificial intelligence is the process of anchoring an AI system's outputs to verifiable
  • RoBERTav9RoBERTa (Robustly Optimized BERT Pretraining Approach) is an open-source natural language processing model released in July 2019 by researchers at Facebook AI (now Meta AI) and the University of Washington…
  • Knowledge Editingv5Knowledge editing (also called model editing) is a family of techniques for updating or correcting specific factual associations stored in the weights of a trained large language model without full retraining…
  • MLflowv9MLflow is an open-source platform for managing the end-to-end machine learning lifecycle, covering experiment tracking, model packaging, a model registry, deployment, and (since 2025) generative-AI…
  • Curriculum learningv6Curriculum learning is a training strategy for machine learning models in which training examples are presented in a meaningful, easy-to-hard order rather than at random
  • Common Crawlv9Common Crawl is a nonprofit 501(c)(3) organization that maintains a free, open repository of web crawl data, and it is the single largest publicly available source of text used to train large language models.
  • WinoGrandev6WinoGrande is a large-scale benchmark for commonsense reasoning consisting of 44,000 binary fill-in-the-blank pronoun resolution problems, built to test whether language models genuinely understand commonsense…
  • COCO datasetv9COCO (Common Objects in Context) is a large-scale dataset for object detection, image segmentation, keypoint detection, and image captioning.
  • DenseNetv5DenseNet (Densely Connected Convolutional Networks) is a convolutional neural network architecture that connects every layer to every other layer in a feed-forward fashion
  • BIG-Benchv7BIG-Bench (Beyond the Imitation Game Benchmark) is a large-scale, collaborative benchmark of 204 tasks, contributed by 450 authors across 132 institutions, built to measure and extrapolate the capabilities of…
  • DeiTv6DeiT (Data-efficient Image Transformers) is a family of vision transformer models that proved Vision Transformers can be trained to state-of-the-art image classification accuracy on ImageNet alone
  • ConvNeXtv9ConvNeXt is a family of pure convolutional neural network (CNN) models that match or beat Vision Transformers on standard vision benchmarks
  • Text summarizationv8Text summarization is the natural language processing (NLP) task of automatically producing a shorter version of one or more documents that preserves the most important information from the original text.
  • Mambav12Mamba is a neural network architecture for sequence modeling that uses selective state space models (SSMs) to process sequential data in linear time with respect to sequence length
  • AI Safety Institutesv6AI Safety Institutes are government-established organizations that test, evaluate, and research the safety and security risks of advanced artificial intelligence systems
  • Continual learningv6Continual learning, also called lifelong learning or incremental learning, is a machine learning paradigm in which a model learns from a stream of tasks or data distributions over time
  • Mixture of Agentsv5Mixture of Agents (MoA) is a multi-model collaboration framework that combines multiple large language models (LLMs) in a layered architecture
  • Wav2Vecv7Wav2Vec is a family of self-supervised learning models from Meta AI (formerly Facebook AI Research) that learn speech representations directly from raw audio waveforms
  • GPU computingv8GPU computing is the use of a graphics processing unit (GPU) to perform general-purpose computation that was traditionally handled by the central processing unit (CPU).
  • Pika (video generation)v7Pika is an artificial intelligence video generation platform developed by Pika Labs, Inc. that lets users create and edit short videos from text prompts, images, and existing clips, and it is best known for…
  • AI governancev7AI governance is the collection of frameworks, norms, standards, policies, and institutional arrangements that guide the development, deployment, and use of artificial intelligence systems so that they are…
  • XLNetv7XLNet is a generalized autoregressive pretraining method for natural language processing that combines the strengths of autoregressive and autoencoding language models.
  • Cross-attentionv10Cross-attention is a variant of the attention mechanism in which the queries are derived from one sequence or representation while the keys and values are derived from a different sequence or representation
  • Inception (deep learning)v7Inception is a family of convolutional neural network (CNN) architectures developed by researchers at Google, first introduced in 2014.
  • AI artv8AI art is artwork created with the assistance of artificial intelligence systems, spanning visual images, music, video, and other creative outputs generated or co-created by algorithms, neural networks, and…
  • T5 (language model)v11T5 (Text-to-Text Transfer Transformer) is a family of transformer-based language models released by Google in 2019-2020 that reframes every natural language processing (NLP) task, classification, translation…
  • Named entity recognitionv6Named entity recognition (NER) is the natural language processing task of locating spans of text that name real-world things, such as people, organizations, and locations, and classifying each span into a…
  • GLUE benchmarkv6The General Language Understanding Evaluation (GLUE) benchmark is a collection of nine natural language understanding (NLU) tasks designed to evaluate and compare the performance of language models across a…
  • Image segmentationv8Image segmentation is the computer vision task of partitioning a digital image into multiple regions by assigning every pixel a label, producing a pixel-level map of what each part of the image contains.
  • Kling (video generation)v9Kling is an AI video generation model developed by Kuaishou Technology, a Chinese short-video platform company publicly traded on the Hong Kong Stock Exchange.
  • Gradiov5Gradio is an open-source Python library that lets developers build interactive web interfaces for machine learning models, APIs, and arbitrary Python functions in a few lines of code.
  • Model mergingv5Model merging combines the parameters of multiple trained neural networks into a single unified model without any additional training.
  • Existential risk from AIv8Existential risk from artificial intelligence (also called AI x-risk) is the hypothesis that the development of sufficiently advanced artificial intelligence could cause human extinction, permanent…
  • SQuADv8SQuAD (the Stanford Question Answering Dataset) is a large-scale reading comprehension benchmark from Stanford University in which a model must answer a question by extracting the exact span of text that…
  • Chinchilla scaling lawsv6The Chinchilla scaling laws are a set of empirical findings published by DeepMind researchers in 2022 showing that, for a fixed compute budget, a large language model trains most efficiently when its number of…
  • llama.cppv9llama.cpp is an open-source large language model inference engine written in C and C++ by Bulgarian software engineer Georgi Gerganov that runs large language models on consumer-grade hardware without…
  • EfficientNetv8EfficientNet is a family of convolutional neural network architectures and a model-scaling method that uniformly scales network depth, width, and input resolution with a single compound coefficient, developed…
  • Hyperparameter Tuningv6Hyperparameter tuning (also called hyperparameter optimization or hyperparameter search) is the process of finding the configuration parameters that make a machine learning model perform best, by searching…
  • VGGv6VGG (also called VGGNet) is a deep convolutional neural network architecture, introduced in 2014 by Karen Simonyan and Andrew Zisserman of the Visual Geometry Group at the University of Oxford
  • Test-time computev9Test-time compute (also called inference-time compute scaling or test-time scaling) is the practice of allocating additional computation while a large language model answers a query
  • YOLO (object detection)v7YOLO (You Only Look Once) is a family of object detection models that treat detection as a single regression problem, predicting bounding boxes and class probabilities directly from full images in one forward…
NewerPage 37 of 46Older