run stats · from the Claude Code session transcripts
What each run cost.
Same prompt, same harness, one session per model. Bars share a scale within each panel; the longest bar is the largest value.
All numbers
| Fable 5.1 | Opus 5.5 | Sonnet 5 | Sonnet 5.5 | |
|---|---|---|---|---|
| Effort | medium | medium | high | high |
| Active time | 41.0 min | 72.2 min | 38.1 min | 37.9 min |
| Wall-clock time | 48.9 min | 77.2 min | 46.4 min | 45.8 min |
| Input tokens | 2,468 | 270 | 366 | 234 |
| Output tokens | 172,924 | 259,556 | 157,172 | 216,112 |
| Output tokens / active min | 4,218 | 3,595 | 4,125 | 5,702 |
| Cache reads | 13,452,183 | 39,664,379 | 40,309,719 | 29,213,598 |
| Cache writes | 190,273 | 355,735 | 256,127 | 305,503 |
| Estimated cost | $14.41 | $14.90 | $10.27 | $8.77 |
| Own code | 1,821 lines · 6 files | 3,958 lines · 16 files | 1,764 lines · 6 files | 2,516 lines · 14 files |
| Simulation grid | 160 × 160 · 2 m cells | 240 × 200 · 2 m cells | 120 × 100 · 1 m cells | 160 × 140 |
| Solver | Staggered-grid shallow water, upwind advection, wet/dry fronts, wave maker | MAC-grid shallow water, semi-Lagrangian momentum, Flather offshore boundary | Virtual-pipes flux scheme with outflow clamping and sponge boundaries | Inertial shallow water on a staggered grid, implicit Manning friction, flux limiting |
How these were measured
- Active time drops gaps over 5 minutes; wall-clock time is first to last message.
- Extracted from the Claude Code session transcripts of each one-shot run. Cost is an estimate at list API prices (cache writes billed at 1.25x input).
- Cache reads dominate the token totals in agentic runs: every tool call re-reads the conversation so far.
- Own code counts non-blank lines of the model's .js/.html/.css files, including its own tests and tools, excluding vendored three.js and generated bundles.
- The Fable 5.1 session hit the output-token limit once and was told to resume; the Sonnet 5.5 session received some background-monitor notifications. Neither run got any additional guidance.
- Effort differs: Fable 5.1 and Opus 5.5 ran at medium, both Sonnets at high.