Leaderboard

Best intelligence per dollar

Quality index divided by reuse-adjusted blended cost at 50% reuse. Rewards models that are both capable and cheap to run once caching is working.

#ModelProviderQuality / $List blendedCache disc.
1OpenAI gpt-oss-120bFireworks325.7$0.2690%
2GPT-5 MiniOpenAI121.0$0.6990%
3Grok 4.3xAI72.7$1.5684%
4Kimi K2.6Fireworks61.4$1.7183%
5DeepSeek-V4-ProFireworks54.5$2.1792%
6GLM 5.1Fireworks51.7$2.1581%
7Claude Haiku 4.5Anthropic45.7$2.0090%
8Gemini 3.5 FlashGoogle32.8$3.3890%
9Gemini 3.1 Pro PreviewGoogle24.6$4.5090%
10Claude Sonnet 4.6Anthropic17.6$6.0090%
11Claude Opus 4.8Anthropic11.9$10.0090%
12Claude Opus 4.7Anthropic11.8$10.0090%
13GPT-5.5OpenAI10.4$11.2590%
14Claude Fable 5Anthropic6.0$20.0090%
15GPT-5.5 ProOpenAI1.3$67.500%
Method

intelligence ÷ reuseAdjustedBlended(model, 0.5, 0.25).

Rank models for your workload.

A diagnostic measures your real reuse and re-ranks the catalog for the way you actually call models.