Revision History
Best AI Coding Assistants · 6 revisions
Sizes are character counts of the article source. The signed number is the change from the previous revision.
Recent edit summaries
Detailed summaries recorded by editors. Generic maintenance summaries are omitted here; every recorded revision remains below. Dates describe the edit, not necessarily the event it covers.
- Correction: Grok 4.7's 37.6% Terminal-Bench 4.0 score is over all 330 trials; only its cost figure is marked partial (324 of 330) on the leaderboard
Version 6 · Sep 23, 2026, 12:14 PM
- Correction: Augment 65.4 is a self-report, not a vals.ai result; citation fixes (18 July capture, Cursor Grok 4.7 post); Copilot Auto subset; Cursor pools; vals.ai archive wording; Grok 4.7 partial run (independent verification V8)
Version 5 · Sep 23, 2026, 12:08 PM
- Correction: Terminal-Bench 2.1 described as harder than 2.0 (it scored higher); Opus 4.8 self-report was 74.6% with Terminus-2, not 87.6% with an internal harness; Gemini CLI free personal tier ended 18 Jun 2026; Codex Pro 'up to $200' and Cursor '$20 usage included'/'4x speed' unsourced; Augment 51.8% SWE-bench Pro cited to wrong post; broken Opus 4.8 and Artificial Analysis links; update: Terminal-Bench 4.0 leaderboard, Opus 5.5, Fable 5.1, Opus 5, GPT-6 Astra/Sol, Grok 4.7/Grok Build, Cursor-SpaceX, Devin Desktop/SWE-2, Copilot/Cursor/Claude pricing, open-weight models
Version 4 · Sep 23, 2026, 11:35 AM
- Correction: the Sonnet 5 increase to $3/$15 was cancelled; $2/$10 became the standard price
Version 3 · Sep 23, 2026, 07:29 AM
- Added 6 contextual internal links
Version 2 · Jul 23, 2026, 04:06 PM