Stargate Project
The Stargate Project is a private American AI infrastructure joint venture, announced on January 21, 2025 at the White House, under which SoftBank, OpenAI, Oracle, and MGX committed to deploy $100 billion…
Explore AI Infrastructure through related topics and the articles other pages reference most.
Articles that also belong to these categories. Counts cover all of AI Infrastructure.
Showing 241-281 of 281 articles
The Stargate Project is a private American AI infrastructure joint venture, announced on January 21, 2025 at the White House, under which SoftBank, OpenAI, Oracle, and MGX committed to deploy $100 billion…
Stargate UAE is a large artificial intelligence data center cluster being built in Abu Dhabi, United Arab Emirates, and is the first site of the Stargate Project located outside the United States.
Supabase is an open-source backend-as-a-service (BaaS) platform that bundles a hosted PostgreSQL database with authentication, file storage, real-time subscriptions, edge functions, and vector similarity…
TPU Ironwood (officially TPU v7 or TPU7x) is Google's seventh-generation Tensor Processing Unit and the first TPU designed specifically for inference, unveiled at Google Cloud Next 2025 in Las Vegas on April…
A TPU node is the legacy Google Cloud architecture for accessing Tensor Processing Unit (TPU) hardware, in which a user's virtual machine (VM) runs application code and communicates with a separate
A TPU Pod is a single Google supercomputer built from many Tensor Processing Unit (TPU) chips wired directly to each other by a high-speed Inter-Chip Interconnect (ICI) fabric arranged as a 2D or 3D torus, so…
A TPU worker is a virtual machine (VM) running Linux that has direct access to one or more Tensor Processing Unit (TPU) chips and executes the actual TPU computation on that attached hardware.
Tairos (Chinese: 钛螺丝, "titanium screw") is an embodied-AI platform for robotics built by Tencent.
Talen Energy Corporation (Nasdaq: TLN) is an American independent power producer headquartered in Houston, Texas, that owns and operates roughly 13 gigawatts of electricity generation
Tavily is a search engine and API platform built specifically for AI agents and retrieval-augmented generation (RAG) systems.
Tencent Holdings Limited (Chinese: 腾讯控股有限公司; pinyin: Téngxùn) is a Chinese multinational technology and entertainment conglomerate based in Shenzhen, Guangdong, and is one of the world's largest internet…
A Tensor Core is a specialized execution unit inside NVIDIA GPUs that computes a small matrix multiplication and accumulation, D = A x B + C
Tensor parallelism (TP) is a distributed training technique that splits the individual weight matrices of a neural network layer across multiple devices, so that each device computes a partial result that is…
A Tensor Processing Unit (TPU) is a family of custom application-specific integrated circuits developed by Google to accelerate machine-learning computation.
Tenstorrent is a North American artificial intelligence hardware and intellectual property company that designs processors for AI training and inference on the basis of the open-standard RISC-V instruction set…
Terafab (styled "Terafab" on the project's official site, "TERAFAB" in Elon Musk's announcement post, and frequently written "TeraFab" in press coverage) is a planned semiconductor fabrication venture between…
TerraPower is an American advanced nuclear energy company chaired by Bill Gates, the co-founder of Microsoft, that is building Natrium
ThunderKittens (often abbreviated TK) is an embedded C++ domain-specific language and header-only library for writing high-performance AI kernels on modern NVIDIA GPUs.
Trusted Execution Environments for machine learning (TEEs for ML, sometimes marketed as "Confidential AI" or "confidential inference") are deployments of hardware-isolated execution environments to run…
Turbopuffer is a serverless vector and full-text search database built from first principles on object storage such as Amazon S3 and Google Cloud Storage.
UALink (Ultra Accelerator Link) is an open industry standard for scale-up interconnect between AI accelerators that lets up to 1,024 accelerators inside a single pod read and write each other's memory directly…
Ultra Ethernet is an open networking specification that reworks Ethernet into a high performance fabric for large AI training clusters and high performance computing, giving operators an interoperable
Vast.ai is a cloud-based GPU rental marketplace that connects independent hardware operators (called hosts) with developers, researchers and companies looking to rent compute by the second.
A vector database is a database that stores data as high-dimensional vectors (numerical embeddings produced by a machine learning model) and retrieves records by similarity rather than exact match
Vercel (originally ZEIT) is an American cloud platform-as-a-service (PaaS) company headquartered in San Francisco, California, best known as the creator of Next.js and as a hosting platform for modern frontend…
Volcano Engine (Chinese: 火山引擎; pinyin: Huoshan Yinqing) is the enterprise cloud and artificial intelligence platform operated by ByteDance, the company that owns Douyin and TikTok.
Voltage Park is a United States cloud computing company that operates a fleet of NVIDIA H100 graphics processing units for artificial intelligence training and inference workloads.
Weaviate is an open-source vector database that stores both data objects and their vector embeddings, enabling a combination of vector similarity search with structured filtering, keyword retrieval, and…
Web scraping is the automated extraction of data from websites, performed by programs that request pages the way a browser does and then parse the returned HTML, or the APIs behind it, into structured records .
WebAssembly (often abbreviated Wasm) is a binary instruction format for a stack-based virtual machine, designed as a portable compilation target for high-level programming languages and intended to enable…
Wormhole is the second-generation AI accelerator application-specific integrated circuit (ASIC) designed by Tenstorrent, a Toronto-based hardware startup led by chip architect Jim Keller.
X-energy, Inc. is an American advanced nuclear reactor and fuel company, based in Rockville, Maryland, that develops the Xe-100, a high-temperature gas-cooled small modular reactor (SMR) rated at about 80…
XLA (Accelerated Linear Algebra) is Google's open-source machine learning compiler that takes computational graphs from frameworks such as TensorFlow, JAX, and PyTorch and transforms them into highly optimized…
bfloat16 (short for Brain Floating Point Format, sometimes written BF16) is a 16-bit floating-point number format that uses one sign bit, eight exponent bits, and seven mantissa bits
d-Matrix is a privately held American semiconductor company headquartered in Santa Clara, California that builds accelerators, I/O cards and software for AI inference in data centers.
Raptor is the second-generation AI inference accelerator from d-Matrix, a Santa Clara semiconductor startup, and the first commercial chip built on the company's 3D stacked digital in-memory compute technology
fal.ai (legally Features and Labels, Inc., often stylized fal) is a San Francisco-based artificial intelligence company that operates a specialized inference platform for generative-media models, including…
llm-d is an open-source, Kubernetes-native framework for serving large language models in a distributed way at production scale.
pgvector is an open-source PostgreSQL extension that adds vector similarity search to a standard PostgreSQL database, letting developers store, index, and query high-dimensional embeddings using ordinary SQL.
Colossus is an artificial-intelligence supercomputer and GPU training cluster operated by xAI in Memphis, Tennessee, and is widely described as one of the largest single-site AI GPU clusters in the world.
xAI Colossus 2 is the second-generation AI supercomputer and data-center complex built by xAI, the artificial intelligence company founded by Elon Musk, in the Memphis, Tennessee, metropolitan area.