Anthropic
Claude Fable 5
claude-fable-5- Context
- 1M
- Max output
- 128K
- Distribution
- Hosted API
- Input / output price
- $10 / $50
Available until at least Jun 9, 2027
Find models that meet your actual constraints, then compare their context, capabilities, pricing, distribution, and lifecycle side by side. Every model card links back to a first-party source.
Finder
Results use reviewed first-party model facts, not benchmark rankings.
Side by side
3 of 4 selected
| Attribute | OpenAI GPT-5.6 Terra | Anthropic Claude Sonnet 5 | Gemini 3.6 Flash |
|---|---|---|---|
| Lifecycle | Active No retirement announced | Active Available until at least Jun 30, 2027 | Active No retirement announced |
| API / checkpoint ID | gpt-5.6-terra | claude-sonnet-5 | gemini-3.6-flash |
| Distribution | Hosted API | Hosted API | Hosted API |
| Context window | 1.05M tokens | 1M tokens | 1.05M tokens |
| Maximum output | 128K tokens | 128K tokens | 65.5K tokens |
| Modalities | text, image → text | text, image → text | text, image, video, audio, pdf → text |
| Capabilities | reasoningfunction callingstructured outputsstreamingtool use | adaptive thinkingvisiontool useprompt cachingbatch processing | thinkingfunction callingstructured outputscode executionsearch groundingcontext caching |
| Standard text price | $2.50 input / $15 output USD per 1M tokens | $2 input / $10 output USD per 1M tokens | $1.50 input / $7.50 output USD per 1M tokens |
| Open-weight details | Hosted API only | Hosted API only | Hosted API only |
| Official source | GPT-5.6 Terra model card ↗ | Claude models overview ↗ | Gemini 3.6 Flash model card ↗ |
17 of 27 reviewed records
Anthropic
claude-fable-5Available until at least Jun 9, 2027
Anthropic
claude-haiku-4-5-20251001Available until at least Oct 15, 2026
Anthropic
claude-opus-4-8Available until at least May 28, 2027
Anthropic
claude-sonnet-5Available until at least Jun 30, 2027
gemini-3.5-flashNo retirement announced
gemini-3.5-flash-liteNo retirement announced
gemini-3.6-flashNo retirement announced
gemma-4-31b-itNo retirement announced
OpenAI
gpt-5.6-lunaNo retirement announced
OpenAI
gpt-5.6-solNo retirement announced
OpenAI
gpt-5.6-terraNo retirement announced
Mistral AI
mistral-small-2603+1No retirement announced
OpenAI
gpt-oss-120bNo retirement announced
OpenAI
gpt-oss-20bNo retirement announced
Meta
llama-4-maverickNo retirement announced
Meta
llama-4-scoutNo retirement announced
gemini-3.1-pro-previewNo retirement date published
Specifications narrow a shortlist; they do not rank quality.
Benchmarks, latency, rate limits, regional availability, safety behavior, and real workload quality still require testing. For workload-specific token costs, use the API cost and context planner.
Method and limits
The catalog is a curated snapshot verified on July 23, 2026. It preserves model IDs, effective dates, caveats, and source links rather than silently treating a rolling alias as a permanent checkpoint.
Prices shown are standard base text-token rates where the provider publishes one. Long-context tiers, regions, caching, batches, media, tools, fine-tuning, and hosted open-model prices may differ. Confirm the linked source before committing a budget or migration.
The finder compares first-party published facts: lifecycle state, model or deployment ID, context window, maximum output, input and output modalities, selected capabilities, standard text-token pricing, distribution, parameter counts, and licenses where applicable.
No. Token counts, prompt caching, batch discounts, long-context tiers, tool calls, images, audio, reasoning tokens, retries, and provider-specific fees can change the bill. Use the linked cost planner with a realistic workload.
Not necessarily. Open-weight means downloadable model weights are available. Each model still has its own license, and some community licenses impose conditions that differ from standard open-source licenses.
There is no universal winner. This tool narrows models by documented constraints; it does not invent one composite quality score. Test the finalists on representative prompts and measure quality, latency, reliability, and total cost.