Magic Patterns
Magic Patterns (at the domain magicpatterns.com) is a San Francisco-based AI design and prototyping tool that turns text prompts, screenshots, and Figma files into interactive user interfaces and…
Explore AI Code Generation through related topics and the articles other pages reference most.
Articles that also belong to these categories. Counts cover all of AI Code Generation.
Showing 61-114 of 114 articles
Magic Patterns (at the domain magicpatterns.com) is a San Francisco-based AI design and prototyping tool that turns text prompts, screenshots, and Figma files into interactive user interfaces and…
MetaGPT is an open-source multi-agent framework that organizes large language model agents into a simulated software-development company, with role-specialized agents (Product Manager, Architect, Project…
MiniMax Code is a desktop AI agent application from Shanghai-based AI company MiniMax that combines chat, local project context, file operations, terminal sessions, browser previews, skills, memory, and…
Mistral Vibe is a unified AI assistant and agent from Mistral AI for professional work, software development, and conversational tasks.
Multi-SWE-bench is a multilingual benchmark for evaluating the ability of large language model based coding systems to resolve real-world software issues across seven programming languages.
MultiChallenge is an AI benchmark for evaluating large language models on realistic multi-turn conversations.
Muse Code is a terminal coding agent developed by Meta Superintelligence Labs. Meta introduced it in beta on August 5, 2026 alongside the Muse Spark 1.2 model that powers it
North Mini Code is an open-weight large language model developed by Cohere for agentic software development and AI code generation.
OpenHands is an open-source, autonomous software development agent platform created by All Hands AI that lets AI agents write code, run shell commands, browse the web, and edit files inside an isolated Docker…
Ox Alpha was the anonymous preview alias for GLM-5.3-Flash, a natively multimodal large language model developed by Z.ai.
Pass@k is the standard metric for evaluating code generation models: it measures the probability that at least one of k generated candidate solutions passes all of a problem's unit tests.
Pi is an open-source coding agent and agent harness written in TypeScript, created in 2025 by the Austrian developer Mario Zechner and owned since April 2026 by Earendil Inc., a public benefit corporation…
Poolside AI is an American-founded artificial intelligence company that trains its own foundation models for software development and sells them to enterprises and regulated-industry customers that run the…
Prime Agent is an open-source coding agent with a terminal user interface, released by Prime Intellect on August 5, 2026.
Program synthesis is the task of automatically constructing a program that satisfies a specification expressed at a higher level than the code itself: a logical formula, a set of input-output examples, a…
Programming with ChatGPT is the practice of using OpenAI's conversational chatbot to read, write, refactor, document, debug, test, and explain source code in plain language instead of an editor or a…
Qodo (formerly CodiumAI or Codium) is an AI-powered code integrity platform that provides automated code review, test generation, and code quality tools for software developers.
Qwen2.5-Coder is the code-specialized series within the Qwen2.5 generation of large language models developed by the Qwen team at Alibaba.
Qwen3-Coder-Next is an open-weight code generation model released by Alibaba's Qwen team on February 3, 2026, under the Apache 2.0 license.
Qwen3.8-Max is a 2.4-trillion-parameter mixture-of-experts large language model developed by Alibaba's Qwen team and released on August 3, 2026 under the title "Qwen3.8-Max: A New Bar for Coding and Cowork".
Real-SWE is a coding-agent benchmark published in September 2026 by Specific Labs, a San Francisco company that licenses operational data and source code from businesses and packages them as training and…
Replit is an online integrated development environment (IDE) and AI-powered coding platform that lets users write, run, and deploy software directly from a web browser, including by describing an application…
RepoBench is an AI benchmark for repository-level code auto-completion, introduced in the 2023 paper "RepoBench: Benchmarking Repository-Level Code Auto-Completion Systems" by Tianyang Liu, Canwen Xu, and…
Roo Code is an open-source AI coding agent that runs as a Visual Studio Code extension.
SWE-Atlas is a benchmark for evaluating AI coding agents on professional software-engineering work that goes beyond fixing bugs and resolving issues.
SWE-Bench Pro (stylized SWE-BENCH PRO) is a contamination-resistant benchmark, released by Scale AI in September 2025, that measures whether an AI coding agent can resolve long-horizon
SWE-Lancer is a benchmark released by OpenAI in February 2025 that evaluates the ability of frontier large language models to perform real-world freelance software-engineering work.
SWE-agent is an open-source autonomous software engineering agent created by the Princeton NLP group (Princeton Language and Intelligence, with co-authors from Stanford) and first released on April 2, 2024.
SWE-bench Multilingual is an AI benchmark of 300 real-world software bug-fixing tasks drawn from 42 open-source repositories across nine programming languages
SWE-bench Multimodal (also written SWE-bench M) is a benchmark that measures whether autonomous software-engineering systems can resolve bugs in visual
SWE-bench Verified is a 500-problem, human-validated subset of the SWE-bench software engineering benchmark, released on August 13
SWE-rebench is a continuously refreshed, contamination-resistant AI benchmark and public leaderboard for evaluating AI agents on real-world software engineering tasks.
SantaCoder is a 1.1 billion parameter large language model for code generation, released in early 2023 by the BigCode project, an open scientific collaboration co-led by Hugging Face and ServiceNow Research.
Scott Wu is an American entrepreneur and competitive programmer who is the co-founder and chief executive officer of Cognition AI, the startup behind Devin, an autonomous coding agent marketed as an "AI…
Sourcegraph Cody is an enterprise AI coding assistant built by Sourcegraph that combines large language models with the company's code-search and code-graph technology to deliver context-aware chat…
Spec-driven development (SDD) is a software engineering methodology in which structured specification documents serve as the primary source of truth for a project, with code treated as a generated or verified…
Spider 2.0 is a benchmark for evaluating large language models on real-world enterprise text-to-SQL workflows.
StarCoder is a family of open-access large language models for code generation and code understanding, developed by the BigCode project, an open scientific collaboration led by Hugging Face and ServiceNow.
Supermaven was an AI code completion tool that used a proprietary neural network architecture to deliver fast, context-aware code suggestions directly inside a developer's editor.
Sweep (stylized with a broom emoji and reachable at sweep.dev) is an AI developer tool built by the San Francisco startup of the same name, founded in 2023 by Kevin Lu and William Zeng.
Tabby is an open-source, self-hosted AI code generation assistant developed by TabbyML, Inc. It is positioned as a privacy-preserving, on-premises alternative to GitHub Copilot: organizations run the Tabby…
Tabnine is a privacy-focused, enterprise artificial intelligence code assistant that provides inline code completions, chat-based assistance, test generation, and code review inside a developer's integrated…
Terminal-Bench is an open benchmark for evaluating AI agents on complex, real-world tasks performed through command-line terminal interfaces.
The Stack is a family of large, permissively-licensed source-code datasets built by the BigCode project, an open scientific collaboration jointly led by Hugging Face and ServiceNow Research, to train and…
TheAgentCompany is an AI benchmark that evaluates AI agents on long-horizon, economically valuable knowledge work inside a self-hosted simulation of a small software company.
Trae (styled TRAE, short for The Real AI Engineer) is an AI-native integrated development environment (IDE) developed by ByteDance, the Chinese technology company best known for TikTok and the Doubao large…
Vibe Code Bench (VCB) is a benchmark that measures whether AI models can build a complete, deployable web application from nothing but a written specification.
Vibe coding is the practice of building software by describing what you want in natural language and letting an AI write the code.
Vibe coding is a software development practice in which a programmer describes their intent in plain natural language and relies on a large language model (LLM) to generate the corresponding source code.
Vibe engineering is a term for the disciplined practice of building production software with the help of large language model coding agents
AI in web development is the use of large language models and related generative systems to design, build, deploy, and maintain websites and web applications.
Windsurf is an AI-powered code editor developed by the company formerly known as Codeium (originally Exafunction).
opencode is an open-source AI coding agent built for the terminal. Developed by Anomaly Innovations, formerly known as the SST (Serverless Stack) team, it provides a terminal user interface (TUI) through which…
v0 is an AI application builder from Vercel that turns natural-language prompts and images into working React code, generating user interfaces styled with Tailwind CSS and the shadcn/ui component library .