Compare
Put models head to head.
Pick up to four models and read them side by side. The best value in every row is called out. The selection lives in the URL, so any comparison is a link you can share.
2/4 selected
| Metric | Claude Fable 5.1#1 · Claude Fable 5.1 | GPT-5.5#7 · GPT-5.5 |
|---|---|---|
| Accuracyhigher is better | 79.4% | 76% |
| Hallucinationlower is better | 2.6% | 6.2% |
| Cost / 1klower is better | $674 | $316 |
| Valueacc. pts per $/1k | 0.12 | 0.24 |
| p50 latencylower is better | 61.2s | 90s |
| Run-to-run σlower is better | ±0 | ±0 |
| Reliabilityhigher is better | 100% | 100% |