AI benchmark results and model performance
Individual benchmark scores plotted by date.
| Organisation | Model | Reported | Top Score | Info | Self Reported | Source |
|---|---|---|---|---|---|---|
| Claude Sonnet 5 | 30 Jun 2026 | 5.80% | Harvey held-out set; all-pass rate; reported in Anthropic's system card | Yes | Source | |
| Claude Opus 5 | 24 Jul 2026 | 11.70% | Held-out all-pass rate reported by Harvey; 94.1% mean criterion-pass rate | No | Source | |
| Claude Fable 5 | 09 Jun 2026 | 13.30% | Higher of Mythos 5 / Fable 5 in Anthropic's launch table | Yes | Source | |
| Claude Mythos 5 | 09 Jun 2026 | 13.30% | Higher of Mythos 5 / Fable 5 in Anthropic's launch table | Yes | Source |