Leaderboard

Best models for coding agents

A composite for repository-scale coding agents: 60% quality index, 40% measured reuse, limited to tool-capable models that teams actually wire into coding loops.

#ModelProviderCoding fitList blendedCache disc.
1Claude Fable 5Anthropic85.2$20.0090%
2Claude Opus 4.8Anthropic84.2$10.0090%
3GPT-5.5OpenAI81.0$11.2590%
4Claude Sonnet 4.6Anthropic76.0$6.0090%
5GLM 5.1Fireworks73.8$2.1581%
6Kimi K2.6Fireworks72.2$1.7183%
7DeepSeek-V4-ProFireworks70.4$2.1792%
Method

0.6 x intelligence + 0.4 x reuseMedianPct, filtered to tool-capable models tagged for coding or agentic use.

Rank models for your workload.

A diagnostic measures your real reuse and re-ranks the catalog for the way you actually call models.