The Claude Lineage: A Consolidated Comparison, Claude 1 to Fable 5.1
This series wrote full case studies for Claude 1 and Claude 2. Rather than a dedicated article for every release since, this piece lines up the entire Claude lineage — Claude 2.1, the Claude 3 family, Claude 3.5, Claude 3.7, the Claude 4 family, the 4.x and 5.x point releases, and the new Mythos-class Fable/Mythos tier — across the same eight dimensions this series uses throughout, and closes with a direct comparison to the GPT lineage this site has also documented in full.
A Lineage Defined by What It Never Disclosed
This site's GPT lineage comparison found a trend: disclosure declined over time, from GPT-1's precise "one month, 8 GPUs" to GPT-6 Astra's near-total opacity. Claude's lineage has no such arc — because Anthropic never disclosed parameter counts, architecture, or training compute for any Claude model, including the very first one. What did change, repeatedly and dramatically, was capability, safety framework maturity, and the shape of the product itself.
Fifteen Releases in Three and a Half Years
Where the Detailed Story Already Lives
- Research Trigger
- Anthropic's own founding: ex-OpenAI researchers building a company around safety pace as the founding purpose.
- Learning Technique & Scaffolding
- RLHF plus Constitutional AI's self-critique-against-written-principles technique.
- Architecture
- Undisclosed from the very first release — not a later trend, the starting condition.
- Safety Framework
- None named yet — Constitutional AI itself functioned as the safety approach.
- Research Trigger
- Closing the direct-consumer-product gap ChatGPT had already opened eight months earlier.
- Learning Technique & Scaffolding
- Refined RLHF + Constitutional AI; first public claude.ai chat product, US/UK only at launch.
- Data
- Training cutoff January 1, 2023; sourcing undisclosed.
- Safety Framework
- Still none named — the second and final Claude release before the Responsible Scaling Policy existed.
The Multi-Tier Lineup and the First Reasoning Fork
The First "Critical"-Tier Safety Activation
- Research Trigger
- Push agentic coding and long-horizon task reliability further, while capability growth made a higher safety tier newly relevant.
- Learning Technique & Scaffolding
- Extended thinking with tool use, parallel tool execution, and substantially improved long-term memory; 65% less likely to take shortcuts on agentic tasks than Sonnet 3.7.
- Training Technique
- Extended-thinking token budgets up to 64K reasoning tokens, lifting AIME 2025 math accuracy from ~33% to 75% (Opus 4), and up to 90% with additional test-time computation.
- Safety Framework / ASL Classification
- The first Anthropic model release to activate ASL-3 Deployment and Security Standards — increased protection against model-weight theft and misuse for CBRN weapons, as a precautionary measure rather than a confirmed capability threshold crossing.
- Influence on Next
- Set the ASL-3 baseline every subsequent Claude release (4.1, 4.5, the 5.x line, and the Mythos-class tier) has maintained or built on.
A New Tier Appears Above Opus
- Research Trigger
- A genuinely new capability tier — internally developed under "Project Glasswing" — that Anthropic judged needed a distinct name and a distinct safety architecture, not just a new version number.
- Learning Technique & Scaffolding
- Fable and Mythos are reportedly identical models apart from their safeguards: Fable auto-redirects requests its classifiers flag as relating to cybersecurity, biology/chemistry, or model distillation to a less-capable fallback (Claude Opus), while Mythos is a separately access-controlled, government-adjacent variant without that redirect.
- Safety Framework / ASL Classification
- A twin-model, classifier-based safety-redirect architecture layered on top of the existing ASL framework — a structurally new approach to containment, distinct from GPT-6 Astra's isolation-and-monitoring-based approach to a comparable problem.
- Influence on Next
- Established "Mythos-class" as a tier sitting above the numbered Opus/Sonnet line, continued one update later with Fable 5.1/Mythos 5.1.
- Learning Technique & Scaffolding
- Continued refinement of the Fable/Mythos twin-model safety-redirect architecture introduced in June 2026.
- Training Technique
- Reported 52.6% on Terminal-Bench-Science, an agentic-science benchmark.
- Compute
- Undisclosed, as always — though a 75% reduction in Fable cache-read pricing was disclosed as a deployment-cost change.
- Influence on Next
- The most recent entry in this lineage at the time of writing — released roughly two weeks before this article.
Three Patterns Across the Whole Lineage
Two Labs, Two Different Disclosure Philosophies
Placing this lineage next to this site's GPT Lineage comparison makes the contrast concrete rather than abstract.
| Dimension | GPT Lineage (OpenAI) | Claude Lineage (Anthropic) |
|---|---|---|
| Architecture disclosure trend | Full transparency (GPT-1–3) → total opacity (GPT-4 onward) | Undisclosed from the first release onward — no decline, because there was no starting transparency |
| Named safety framework | Preparedness Framework, introduced alongside growing capability concerns | Responsible Scaling Policy / AI Safety Levels (ASL), introduced after two frameworkless releases |
| First "top-tier risk" classification | GPT-6 Astra — "Critical" cybersecurity capability (Sep 2026) | Claude Opus 4/Sonnet 4 — ASL-3 activation (May 2025), over a year earlier |
| Multi-model product structure | GPT-5's real-time router between a fast and a deep-reasoning model | Claude 3.7's single hybrid-reasoning model; Fable/Mythos's twin safety-redirect models |
| Team credit | Author counts (4 → 6 → 31) until GPT-4, then uncountable contributor lists | Never a numbered author list for any model — company-credited from day one |
The One Thread That Never Broke Here Either
Every Claude model across fourteen named releases remains, as far as any public source confirms, a Transformer-based language model trained via next-token prediction, refined with RLHF and Constitutional-AI-style techniques. Exactly like the GPT lineage, the foundational architecture family was never the interesting variable — what changed was scaffolding, safety architecture, and product shape, layered repeatedly on top of a base nobody outside Anthropic has ever been shown in detail.
Readiness Checklist
⚠️ What's Missing or Uncertain
🔗 Reference Links
- This site — Anthropic Model Case Study: Claude 1
- This site — Anthropic Model Case Study: Claude 2
- This site — The GPT Lineage: A Consolidated Comparison
- This site — The Gemini Lineage: A Consolidated Comparison
- This site — The Grok Lineage: A Consolidated Comparison
- This site — The DeepSeek Lineage: A Consolidated Comparison
- This site — The Meta AI (Llama) Lineage: A Consolidated Comparison
- This site — The Mistral AI Lineage: A Consolidated Comparison
- Anthropic — Claude 2 Model Card
- Anthropic — The Claude 3 Model Family Model Card
- Anthropic — "Claude 3.7 Sonnet and Claude Code"
- Anthropic — "Introducing Claude 4"
- Anthropic — "Activating AI Safety Level 3 Protections"
- Anthropic — "Claude Fable 5 and Claude Mythos 5"
- This site — The Qwen Lineage: A Consolidated Comparison
- This site — The AI Researcher Atlas: 50 People Who Built the Field
- This site — The Objective-Specification Problem (RSI series, Part 2)