Leaderboard

Best intelligence per dollar

Quality index divided by reuse-adjusted blended cost at 50% reuse. Rewards models that are both capable and cheap to run once caching is working.

#ModelProviderQuality / $List blendedCache disc.
1OpenAI gpt-oss-120bFireworks325.7$0.2690%
2GPT-5 MiniOpenAI121.0$0.6990%
3Grok 4.3xAI71.9$1.5684%
4Kimi K2.6Fireworks61.4$1.7183%
5DeepSeek-V4-ProFireworks54.5$2.1792%
6Claude Haiku 4.5Anthropic45.7$2.0090%
7Gemini 3.5 FlashGoogle33.5$3.3890%
8Gemini 3.1 Pro PreviewGoogle24.3$4.5090%
9Claude Sonnet 4.6Anthropic17.6$6.0090%
10Claude Opus 4.7Anthropic11.8$10.0090%
11Claude Opus 4.8Anthropic11.8$10.0090%
12GPT-5.5OpenAI10.4$11.2590%
13Claude Fable 5Anthropic6.0$20.0090%
14GPT-5.5 ProOpenAI1.3$67.500%
Method

intelligence ÷ reuseAdjustedBlended(model, 0.5, 0.25).

Rank models for your workload.

A diagnostic measures your real reuse and re-ranks the catalog for the way you actually call models.