MOC: Frontier Model Releases
MOC: Frontier Model Releases
This map covers major frontier model releases: when each shipped, where it is logged here, and how the frontier has moved on benchmarks that stay comparable across releases. The pace in 2026 is the story. Anthropic shipped four flagship-tier releases between June and September, and OpenAI went from GPT-5.6 Sol to GPT-6 Astra in eight weeks.
Benchmark progress (Jun–Sep 2026)
Most benchmark rows change between announcements. GDPval-AA moved from v2 to v2.1, which changed the Elo scale. OSWorld 2.0 changed its task set, and AutomationBench's leaderboard re-scored older models. Only Terminal-Bench 4.0 and Terminal-Bench-Science 0.1 keep the same numbers across the Fable 5.1 and Opus 5.5 tables: Opus 5 reproduces at 52.3% and 29.0% in both. The chart uses only those two.
| Model | Released | Terminal-Bench 4.0 | Terminal-Bench-Science 0.1 |
|---|---|---|---|
| Claude Fable 5 | 2026-06-09 | 42.0% | 24.7% |
| GPT-5.6 Sol | 2026-07-09 | 37.3% | 22.4% |
| Claude Opus 5 | 2026-07-24 | 52.3% | 29.0% |
| Claude Fable 5.1 | 2026-09-01 | 55.8% | 52.6% |
| GPT-6 Astra | 2026-09-03 | 57.9% | 64.6% |
| Claude Opus 5.5 | 2026-09-22 | 66.4% | 58.7% |
Scores: Fable 5 from the Fable 5.1 announcement;1 all others from the Opus 5.5 announcement,2 where the shared rows match.
Claude Mythos 5.1, which has restricted access, scores 60.9% on Terminal-Bench 4.0. The science benchmark shows the steepest change: Claude went from 29% to 53% in one release (Opus 5 → Fable 5.1), and Astra reached 65%. Anthropic now says benchmark margins at this level are "a less reliable guide to real-world differences", so read the chart for direction rather than decimals. Scores come from vendors; OpenAI figures are as reported by OpenAI.
Timeline
2026
| Date | Model | Lab | Notes | Source |
|---|---|---|---|---|
| 2026-02-05 | Claude Opus 4.6 | Anthropic | 1M context in beta | 3 |
| 2026-02-17 | Claude Sonnet 4.6 | Anthropic | 3 | |
| 2026-03 | gpt-5-4 | OpenAI | Native computer use, 1M context | 4 |
| 2026-04 | gemma-4 | Google DeepMind | Open models built from Gemini 3 research | 5 |
| 2026-04-08 | meta-muse-spark-msl | Meta | Muse Spark, from Meta Superintelligence Labs | 6 |
| 2026-04 | Claude Mythos Preview | Anthropic | Restricted; announced with anthropic-glasswing | 7 |
| 2026-04-16 | Claude Opus 4.7 | Anthropic | 3 | |
| 2026-04-23 | gpt-5-5 | OpenAI | Agentic coding, computer use | 8 |
| 2026-04 | deepseek-v4-preview | DeepSeek | Open-weights MoE, 1M context, very low API prices | 9 |
| 2026-05-28 | Claude Opus 4.8 | Anthropic | Now the fallback model for Opus 5.5's cyber safeguards | 3 2 |
| 2026-06-09 | Claude Fable 5 + Mythos 5 | Anthropic | Access revoked Jun 12 after a US government letter; see fable-5-oneshot-sakura | 10 11 |
| 2026-06-30 | Claude Sonnet 5 | Anthropic | $2 / $10 per M tokens | 3 |
| 2026-07-09 | GPT-5.6 (Luna / Terra / Sol) | OpenAI | Limited preview from Jun 26 | 12 |
| 2026-07-24 | Claude Opus 5 | Anthropic | 1M context, adaptive thinking; see trq212-claude-5-context-engineering | 3 |
| 2026-09-01 | Claude Fable 5.1 + Mythos 5.1 | Anthropic | Most capable general release at the time | 1 |
| 2026-09-03 | gpt-6-astra | OpenAI | Saturates ARC-AGI-3; first "Critical" cyber rating | 13 14 |
| 2026-09-22 | claude-opus-5-5 | Anthropic | Fable 5.1-level for about 40% less than Opus 5; Sonnet/Haiku 5.5 "in the coming weeks" | 2 |
2024–2025 (Claude, for context)
All dates from the public Claude release timeline.3
| Date | Model |
|---|---|
| 2024-03-04 | Claude 3 (Haiku, Sonnet, Opus) |
| 2024-06-21 | Claude 3.5 Sonnet |
| 2024-10-22 | Claude 3.5 Haiku; upgraded 3.5 Sonnet (computer use) |
| 2025-02-24 | Claude 3.7 Sonnet (hybrid reasoning) |
| 2025-05-22 | Claude Opus 4 + Sonnet 4 |
| 2025-08-05 | Claude Opus 4.1 |
| 2025-09-29 | Claude Sonnet 4.5 |
| 2025-10-15 | Claude Haiku 4.5 |
| 2025-11-24 | Claude Opus 4.5 |
Related
- 2026-09-10-the-pace-is-the-story: OpenAI's internal model after Astra, and why pace matters
- pacing-the-frontier: the call to deliberately slow the frontier
- stephen-reid-coding-agents-benchmarks
- moc-ai-security-incidents
Sources
Where a row links to a reference entry here, that entry holds the fuller write-up and links.
Footnotes
-
Introducing Claude Fable 5.1 and Claude Mythos 5.1, Anthropic. ↩ ↩2
-
Introducing Claude Opus 5.5, Anthropic, 2026-09-22. ↩ ↩2 ↩3
-
anthropic-claude-timeline: a community-maintained timeline of Claude model releases. It is a secondary source; check against Anthropic's announcements where it matters. ↩ ↩2 ↩3 ↩4 ↩5 ↩6 ↩7
-
Introducing GPT-5.4, OpenAI. ↩
-
Introducing Muse Spark, Meta. ↩
-
Project Glasswing, Anthropic. ↩
-
Introducing GPT-5.5, OpenAI. ↩
-
DeepSeek V4 Preview release, DeepSeek API docs. ↩
-
Introducing Claude Fable 5 and Claude Mythos 5, Claude Platform docs. ↩
-
Claude Mythos, Wikipedia: covers the US government letter and the access revocation. ↩
-
GPT-6 Astra, Wikipedia: gives the Sep 3 release to approved users and general availability the next day. ↩