Model Analytics

A native dashboard over the curated public benchmark snapshot: cost trends, per-model activity and the experiment record — real data rendered as themeable, accessible charts, not an embedded frame.

Data as of 2026-07-20.

At a glance

Suite cost over CLI versions

Benchmark-suite cost (USD) per model as the CLI advances. Each line is one model; gaps are versions where that model was not run.

Data table
Suite cost (USD) by CLI version and model
CLI versionclaude-haiku-4-5claude-opus-4-7claude-opus-4-8claude-sonnet-4-6claude-fable-5claude-sonnet-5
2.1.156$0.68$2.91$1.20$1.30
2.1.170$2.60
2.1.196$1.38$7.93$1.48
2.1.201$2.91$8.40
2.1.202$0.14$0.79$1.59$0.46
2.1.204$0.14$0.75$1.59$0.64
2.1.205$0.14$0.69$1.41$0.66
2.1.206$0.13$0.71$1.45$0.62
2.1.207$0.16$1.08$1.17$0.47
2.1.208$0.14$0.45$1.05$0.43
2.1.210$0.14$0.50$1.05$0.39
2.1.211$0.14$0.44$1.04$0.39
2.1.212$0.13$0.45$1.04$0.40
2.1.214$0.14$0.47$1.00$0.41
2.1.215$0.13$0.40$0.99$0.40

Output tokens by model

Total benchmark output tokens attributed to each model in the snapshot.

Data table
Benchmark activity by model
ModelFamilySessionsOutput tokens
claude-fable-5fable157341,301
claude-opus-4-8opus180164,977
claude-sonnet-4-6sonnet9081,464
claude-haiku-4-5haiku9058,767
claude-sonnet-5sonnet9053,471
claude-opus-4-7opus9052,384
global.anthropic.claude-haiku-4-5-20251001-v1:0haiku4528,669
global.anthropic.claude-opus-4-7opus4525,699
global.anthropic.claude-sonnet-4-6sonnet4523,653

Experiments

Capability probes

Hypotheses