Compare
Put models head to head.
Pick up to four models and read them side by side — the best value in every row is called out. The selection lives in the URL, so any comparison is a link you can share.
2/4 selected
| Metric | Gemini 3.6 Flash#1 · Gemini 3 | GPT-5.5#6 · GPT-5.5 |
|---|---|---|
| Accuracyhigher is better | 78.9% | 69.3% |
| Hallucinationlower is better | 14.5% | 14.8% |
| Cost / 1klower is better | $116 | $303 |
| Valueacc. pts per $/1k | 0.68 | 0.23 |
| p50 latencylower is better | 52.8s | 97.5s |
| Run-to-run σlower is better | ±5.6 | ±2.1 |
| Reliabilityhigher is better | 100% | 100% |