Thursday, July 23, 2026
- Zero Moment Point (ZMP)v5The Zero Moment Point (ZMP) is the point on the ground where the horizontal component of the moment generated by the contact reaction force vanishes, leaving only a vertical reaction.
- VisualWebArenav3VisualWebArena (often abbreviated VWA) is a benchmark of 910 realistic, visually grounded web tasks for evaluating multimodal autonomous AI agents, released in January 2024 by researchers at Carnegie Mellon…
- Zooxv3Zoox, Inc. is an American autonomous-vehicle company, owned by Amazon since June 2020, that designs and manufactures a purpose-built electric robotaxi with no steering wheel, no pedals, and no defined front or…
- Simulation (in AI and robotics)v3Simulation in artificial intelligence and robotics is the use of computational physics, rendering, and procedural environments to recreate a synthetic version of a physical or virtual world inside which AI…
- RMSPropv5RMSProp (Root Mean Square Propagation) is an adaptive learning-rate optimizer that divides each parameter's gradient by a running root-mean-square of that parameter's recent gradients
- DDPG (Deep Deterministic Policy Gradient)v3DDPG (Deep Deterministic Policy Gradient) is an off-policy, model-free actor-critic algorithm in deep reinforcement learning that learns continuous-control policies by combining a deterministic actor with a…
- Latent Dirichlet allocationv5Latent Dirichlet allocation (LDA) is a generative probabilistic model that discovers the hidden thematic structure in a collection of documents by treating each document as a mixture of a small number of…
- mT5v3mT5 (multilingual T5) is a transformer-based encoder-decoder language model released by Google Research in October 2020 that covers 101 languages in a single model, pre-trained on a Common Crawl corpus called…
- OpenVLAv4OpenVLA is a 7-billion-parameter open-source vision-language-action model (VLA) for robotic manipulation, released in June 2024 by a collaboration of researchers from Stanford University, UC Berkeley, the…
- TurtleBotv3TurtleBot is an open source, low cost mobile robot platform built around the Robot Operating System (ROS) and intended primarily for robotics education and research.
- SimCLRv3SimCLR (Simple Framework for Contrastive Learning of Visual Representations) is a self-supervised learning method for computer vision in which a network is trained to recognise that two differently augmented…
- Arm Holdingsv5Arm Holdings plc is a British semiconductor IP and software design company headquartered in Cambridge, England, whose CPU, GPU, and neural-processor designs sit inside more than 99% of the world's smartphones…
- RMSNormv5RMSNorm (Root Mean Square Layer Normalization) is a feature normalization technique introduced by Biao Zhang and Rico Sennrich in 2019 that scales each activation vector by its root mean square only, dropping…
- Tree of Thoughtsv5Tree of Thoughts (ToT) is a prompting and inference-time search framework for large language models that lets the model explore multiple intermediate reasoning steps as nodes in a tree, evaluate each…
- DBSCANv4DBSCAN (Density-Based Spatial Clustering of Applications with Noise) is a density-based clustering algorithm that groups together points packed closely in feature space and labels points in low-density regions…
- Waseda Universityv4Waseda University (早稲田大学, Waseda Daigaku) is a private research university in Shinjuku, Tokyo, Japan, founded on 21 October 1882 by Ōkuma Shigenobu and best known in artificial intelligence and robotics as the…
- Cleanlabv4Cleanlab is an open source Python library for automatically finding and fixing label errors and other data quality problems in machine learning datasets, and the data-centric AI startup, incorporated in 2021
- ARIMAv4ARIMA (Autoregressive Integrated Moving Average) is a class of statistical models for analyzing and forecasting time series data, specified by three non-negative integer orders written as ARIMA(p, d, q): p is…
- CVPR (Conference on Computer Vision and Pattern Recognition)v3CVPR, the IEEE/CVF Conference on Computer Vision and Pattern Recognition, is the leading annual academic conference for computer vision research and, as of 2025, the single highest-ranked publication venue in…
- LeNetv4LeNet is the pioneering family of convolutional neural networks developed by Yann LeCun and collaborators at AT&T Bell Labs between roughly 1988 and 1998 to read handwritten characters
- Data preprocessingv3Data preprocessing is the set of operations applied to raw data to clean and transform it into a form a machine learning model can use, covering deduplication, type fixing, missing-value imputation, outlier…
- GELU (Gaussian Error Linear Unit)v5The Gaussian Error Linear Unit (GELU) is a smooth, non-monotonic activation function defined as GELU(x) = x · Φ(x), where Φ(x) is the cumulative distribution function of the standard normal distribution.
- Greedy decodingv3Greedy decoding (also called greedy search or argmax decoding) is the simplest text-generation strategy used by autoregressive language models: at every step it picks the single highest-probability next token…
- TikTokv6TikTok is a short-form video app owned by the Chinese technology company ByteDance, best known in artificial intelligence circles as the most widely studied production recommender system: its For You Page…
- Slack (software)v5Slack is a workplace messaging platform owned by Salesforce that organizes team communication into persistent channels, threads, and direct messages, and that Salesforce now markets as an "agentic OS" where…
- Xiaomi CyberDogv4The Xiaomi CyberDog (Chinese: 铁蛋, pinyin Tiědàn, literally "Iron Egg") is a family of low-cost, open-source, bio-inspired quadruped robots developed by Xiaomi's Robotics Lab in Beijing.
- 3D printingv33D printing, also called additive manufacturing (AM), is a family of fabrication processes that build three-dimensional objects from a digital model by adding material in successive layers
- Emacs Lispv3Emacs Lisp (often shortened to Elisp) is the dialect of Lisp used as the extension and scripting language of GNU Emacs.
- Common Lisp Object System (CLOS)v3The Common Lisp Object System (CLOS, often pronounced "see-loss" or "kloss") is the object-oriented programming facility built into ANSI Common Lisp, organized around generic functions and multiple dispatch…
- Cycloidal drivev4A cycloidal drive (also called a cycloidal speed reducer, cyclo drive, cycloidal gearbox, or, in industrial robotics, an RV reducer) is a compact mechanical speed reducer that converts high-speed, low-torque…
- Quanta Computerv4Quanta Computer Inc. (Chinese: 廣達電腦; pinyin: Guǎngdá Diànnǎo; TWSE: 2382) is a Taiwanese electronics manufacturer headquartered in Taoyuan, Taiwan, and one of the two largest contract manufacturers of AI…
- GiveWellv5GiveWell is an American nonprofit charity evaluator, founded in 2007 by Holden Karnofsky and Elie Hassenfeld, that uses cost-effectiveness analysis to recommend a short list of evidence-backed global health…
- VinFastv3VinFast Auto Ltd. is a Vietnamese electric-vehicle maker, headquartered in Hai Phong and listed on the Nasdaq Global Select Market under the ticker VFS, that builds battery-electric cars, e-scooters, and the…
- Smart homev3A smart home is a residence equipped with networked devices, sensors, and appliances that can be monitored or controlled remotely, often through a smartphone app, a wall-mounted hub, or a voice command, and…
- Sparse autoencoderv8A sparse autoencoder (SAE) is a neural network that adds a sparsity penalty to an autoencoder's training loss so that only a small number of hidden units activate for any given input, producing a wide
- GPTZerov4GPTZero is a commercial AI content detector that estimates the probability a block of text was written by a large language model rather than by a human.
- World Robot Conferencev3The World Robot Conference (WRC; Chinese: 世界机器人大会) is an annual international robotics conference and exhibition held in Beijing, China, every year since 2015.
- Liquid AIv6Liquid AI is an artificial intelligence company founded in 2023 by researchers from MIT CSAIL and headquartered at 314 Main Street in Cambridge, Massachusetts.
- Consensus (academic AI search)v5Consensus is an AI-powered academic search engine that uses large language models to find, summarise, and synthesise findings from peer-reviewed scientific literature across a corpus of more than 200 million…
- Model deploymentv3Model deployment is the MLOps process of taking a trained machine learning model and making it available in a production environment so it can serve predictions to applications, users, or downstream systems.
- Inductive biasv3Inductive bias (also called learning bias) is the set of assumptions that a learning algorithm uses to predict outputs for previously unseen inputs.
- Computational graphv4A computational graph is a directed acyclic graph (DAG) representation of a numerical computation, where nodes represent operations (or variables) and edges represent the data, typically tensors
- MLXv4MLX is an open-source array and machine-learning framework built by Apple Inc.'s machine-learning research team specifically for Apple silicon, the M1, M2, M3, M4, and M5 series of system-on-chip processors.
- Switch Transformerv6The Switch Transformer is a sparsely activated Mixture of Experts (MoE) Transformer architecture introduced by William Fedus, Barret Zoph, and Noam Shazeer at Google in January 2021.
- TIAGo (PAL Robotics)v4TIAGo (Take It And Go) is a modular mobile manipulator service robot developed by PAL Robotics, a Spanish robotics company headquartered in Barcelona.
- Toyota Human Support Robot (HSR)v4The Human Support Robot (often abbreviated HSR) is a compact, single-armed mobile manipulation robot built by Toyota Motor Corporation as a standardised research platform for domestic and assistive robotics.
- AI Test Kitchenv5AI Test Kitchen was a Google application for Android, iOS, and the web that let users interact with experimental generative AI models behind a controlled access program.
- Whole-body controlv4Whole-body control (WBC) is a class of robotics control techniques that coordinates all of a robot's degrees of freedom at once to achieve multiple prioritized tasks, such as balancing, reaching, and…
- LAION-5Bv3LAION-5B is an open dataset of approximately 5.85 billion CLIP-filtered image and text pairs scraped from the public internet, released by LAION (Large-scale Artificial Intelligence Open Network) on March 31
- Stable Audiov5Stable Audio is a family of generative AI models from Stability AI that turn a text prompt into music or sound effects as a stereo audio file.
- DeepLabv5DeepLab is a family of deep convolutional neural network architectures for semantic segmentation, developed by Liang-Chieh Chen and collaborators at UCLA and Google between 2014 and 2018 .
- SoftBank Vision Fundv3The SoftBank Vision Fund (SVF) is a series of large technology-focused investment vehicles managed by SB Investment Advisers (UK) Limited, a subsidiary of SoftBank Group Corp. of Japan.
- Hugging Face Transformersv7Hugging Face Transformers is an open-source Python library that provides general-purpose architectures, a unified API, and pretrained weights for state-of-the-art machine learning models across text, vision…
- NVIDIA Hopperv6NVIDIA Hopper is the codename for NVIDIA's ninth-generation datacenter GPU microarchitecture, announced March 22, 2022 by CEO Jensen Huang at the GTC keynote.
- Amazon Roboticsv5Amazon Robotics LLC is the wholly owned robotics and warehouse-automation subsidiary of Amazon that designs, manufactures, and operates the autonomous mobile robots, robotic arms, gantry systems, and…
- Mistral 7Bv7Mistral 7B is a 7.3-billion-parameter, decoder-only large language model released by mistral ai on September 27, 2023, under the apache 2 license.
- Service robotv4A service robot is a robot that performs useful tasks for humans or equipment, excluding industrial automation applications, according to the international vocabulary standard ISO 8373:2021.
- R-CNN (Regions with CNN features)v3R-CNN (short for Regions with CNN features) is a two-stage object detection method that generates about 2,000 candidate region proposals per image with Selective Search, warps each region and runs a…
- Wolfram GPTv4Wolfram GPT is the integration of Wolfram|Alpha and the Wolfram Language with OpenAI's ChatGPT, giving the language model on-demand access to authoritative computational knowledge, real-time data feeds…
- DCGAN (Deep Convolutional GAN)v3DCGAN (Deep Convolutional Generative Adversarial Network) is a family of generative adversarial network architectures, introduced in 2015 by Alec Radford, Luke Metz, and Soumith Chintala
- Apple Neural Enginev3The Apple Neural Engine (ANE) is the dedicated neural network accelerator, a type of neural processing unit (NPU), that Apple builds into the Apple silicon system-on-chip powering iPhone, iPad, Mac, Apple…
- Voice cloningv4Voice cloning is the use of machine learning to generate synthetic speech in the voice of a specific real person (the target speaker) from a sample of their recorded audio.
- Alibaba AIv3Alibaba AI refers to the artificial-intelligence work of Alibaba Group Holding Limited, a Chinese multinational technology conglomerate headquartered in Hangzhou.
- BioGPTv3BioGPT is a domain-specific generative pre-trained Transformer language model for biomedical text generation and mining, developed by Microsoft Research.
- Importance samplingv5Importance sampling (often abbreviated IS) is a Monte Carlo method for estimating the expectation of a function under a target probability distribution $$p$$ by drawing samples from a different proposal…
- Policy gradient methodsv4Policy gradient methods are a family of reinforcement learning algorithms that directly parameterise the agent's policy and optimise it by stochastic gradient ascent on the expected return.
- Feature Pyramid Network (FPN)v3A Feature Pyramid Network (FPN) is a generic feature-extraction architecture for object detection and other dense-prediction tasks that builds a multi-scale feature pyramid with strong semantics at every…
- Topic modelv3A topic model is a statistical model that discovers the abstract "topics" hidden in a collection of documents, where each document is represented as a mixture of a small number of latent topics and each topic…
- Temporal-difference learningv5Temporal-difference (TD) learning is a class of model-free reinforcement learning methods that learn value-function estimates by bootstrapping: updating each estimate of how good a state is toward a target…
- Dense Passage Retrieval (DPR)v3Dense Passage Retrieval (DPR) is a neural information retrieval method that uses a dual-encoder BERT architecture to map questions and passages into dense vectors
- Markov chainv5A Markov chain is a stochastic process in which the probability of the next state depends only on the current state and not on the sequence of states that came before it.
- Control theoryv5Control theory is the mathematical and engineering discipline concerned with designing and analysing systems that achieve desired behaviour through measurement and feedback.
- SSD (Single Shot MultiBox Detector)v3SSD (Single Shot MultiBox Detector) is a single stage object detection model that predicts bounding boxes and per class confidence scores in one forward pass through a convolutional network
- Bayesian statisticsv3Bayesian statistics is a statistical paradigm in which probability expresses a degree of belief that is updated as evidence arrives, using Bayes' theorem.
- TensorFlow Decision Forests (TF-DF)v3TensorFlow Decision Forests (often abbreviated TF-DF) is an open-source Google library for training, serving, and interpreting decision-forest models such as Random Forest and Gradient Boosted Decision Trees…
- Core MLv3Core ML is Apple's framework for running trained machine learning models on-device across iOS, iPadOS, macOS, watchOS, tvOS, and visionOS.
- BigGANv3BigGAN is a class-conditional generative adversarial network that, when introduced by DeepMind researchers Andrew Brock, Jeff Donahue, and Karen Simonyan in 2018, set a new state of the art for AI image…
- Weak supervisionv3Weak supervision is a machine learning paradigm in which models are trained from noisy, limited, imprecise, or programmatically generated labels rather than from large, expensively hand annotated datasets.
- BioBERTv4BioBERT (Bidirectional Encoder Representations from Transformers for Biomedical Text Mining) is a domain-specific language model that adapts BERT to biomedicine by continuing its pre-training on large…
- NVIDIA H200v7The NVIDIA H200 is a data center Tensor Core GPU for AI and high-performance computing that was the first GPU to ship with HBM3e memory, packing 141 GB at 4.8 TB/s on its SXM and NVL boards.
- Bayesian networkv3A Bayesian network (also called a belief network, Bayes net, directed graphical model, or probabilistic causal network) is a probabilistic graphical model that represents a set of random variables and their…
- Natural language inference (NLI)v3Natural language inference (NLI), also known as recognising textual entailment (RTE), is the natural language processing task of deciding whether a hypothesis sentence is entailed by, contradicts, or is…
- CycleGANv3CycleGAN (Cycle-Consistent Generative Adversarial Network) is a deep learning architecture for unpaired image-to-image translation.
- Neuralinkv6Neuralink Corp. is an American neurotechnology company that builds high-bandwidth, fully implantable brain-computer interfaces (BCIs).
- Domain adaptationv3Domain adaptation is the subfield of transfer learning that adapts a model trained on a labelled source domain so it performs well on a related but different target domain, where labels are scarce or absent
- Kai-Fu Leev4Kai-Fu Lee (Chinese: 李開復; born December 3, 1961) is a Taiwanese-American computer scientist, venture capitalist, and author who is the founder and CEO of 01.AI, the Beijing large language model startup behind…
- Differential privacyv5Differential privacy is a mathematical definition of privacy that guarantees the output of an analysis is essentially unchanged whether or not any single individual's record is included in the input
- StyleGANv3StyleGAN is a family of style-based generative adversarial network (GAN) architectures developed by NVIDIA Research for high-quality unconditional image synthesis
- Uncanny valleyv3The uncanny valley is the dip in human affinity that occurs when a robot, computer-generated character, or other artificial figure looks and moves almost like a real person but not quite, triggering a sudden…
- PyTorch Lightningv6PyTorch Lightning is an open-source deep-learning framework, created by William Falcon in 2019, that wraps PyTorch to abstract away the engineering boilerplate of training loops, distributed training, mixed…
- Sundar Pichaiv4Sundar Pichai (Pichai Sundararajan, born June 10, 1972) is the chief executive officer of Google and its parent holding company Alphabet Inc., and the executive who led Google's pivot from a "mobile-first" to…
- Gemini (app)v4The Gemini app is Google's consumer AI assistant, a chat application available on the web at gemini.google.com and as native Android and iOS apps
- Fairlearnv3Fairlearn is an open-source Python toolkit for assessing and improving the fairness of machine-learning models with respect to sensitive attributes such as race, gender, or age.
- Diffusion Transformer (DiT)v6A Diffusion Transformer (DiT) is a transformer-based neural network backbone for diffusion models that replaces the U-Net with a Vision Transformer operating on patches of an image latent.
- Voyage AIv4Voyage AI is an artificial intelligence company that builds state-of-the-art embedding and reranking models for retrieval-augmented generation and semantic search.
- GPT4Allv3GPT4All is an open-source ecosystem from Nomic AI for running large language models locally and privately on consumer laptops and desktops, with no internet connection, GPU, or API key required.
- Cognitive roboticsv3Cognitive robotics is the subfield of robotics and artificial intelligence concerned with endowing robots with higher-level cognitive capabilities: perception, attention, memory, reasoning, knowledge…
- OpenPosev3OpenPose is an open-source library for real-time multi-person 2D pose estimation that detects body, foot, hand, and facial keypoints in images and video.
- Actor modelv3The actor model is a mathematical model of concurrent computation whose universal primitive is the actor: an autonomous, isolated entity that owns private state and communicates with other actors only by…
- Dynamic inferencev3Dynamic inference, also called input-adaptive inference, conditional computation, or adaptive computation