← leaderboard
Command R+
CommandRanks #39 of 43 on reading real contracts into structured billing data — -28 pts vs GPT-5.5.
41.3%
Accuracy
#39 of 43
32.8%
Hallucination
HIGH-confidence & wrong
$87.4
Cost / 1k contracts
$0.09 each · via OpenRouter
811.4s
Median latency
p90 1168.4s
±17.6
Run-to-run σ
3 runs
0.47
Value
acc. pts per $/1k
Tokens per contract
Input24,351
Output2,650
Reasoning0
Reliability 100% · valid structured output across 18 calls.
Want the full picture? How we score →