The TC Score ranks 20 leading large language models across five weighted capability dimensions, giving you a single comparable number grounded in public benchmarks, community evaluations and our own assessments. Last recalculated July 2026.
The TC Score is a weighted composite of five capability dimensions. Each dimension is scored 0–100 and combined using the weights below.
| Rank | Model | Provider | Context | Category Scores | TC Score | Trend |
|---|---|---|---|---|---|---|
| 🥇 | Claude Fable 5Best Overall | Anthropic | 1M | 98.4 | ⇧ | |
| 🥈 | Claude Opus 5Best Agentic | Anthropic | 200K | 97.6 | ⇧ | |
| 🥉 | GPT-5.6 SolBest Reasoning | OpenAI | 1M | 96.1 | ⇧ | |
| #4 | GPT-5.6Best Coding | OpenAI | 1M | 95.2 | ⇧ | |
| #5 | Gemini 3.6 ProBest Context | 2M | 94.3 | ⇧ | ||
| #6 | Claude Sonnet 5Best Value | Anthropic | 1M | 93.5 | ⇧ | |
| #7 | Claude Opus 4.8 | Anthropic | 200K | 92.4 | — | |
| #8 | Gemini 3.6 FlashSpeed Pick | 1M | 89.6 | ⇧ | ||
| #9 | Grok 4Best Live Data | xAI | 1M | 88.7 | ⇧ | |
| #10 | Llama 5 BehemothBest Open Weight | Meta | 2M | 88.1 | ⇧ | |
| #11 | Kimi K3.7Biggest Jump | Moonshot AI | 512K | 87.2 | ⇧ | |
| #12 | GPT-5.6 Mini | OpenAI | 1M | 86.4 | ⇧ | |
| #13 | DeepSeek V4.5Best Price | DeepSeek | 256K | 85.9 | ⇧ | |
| #14 | Qwen 3.8Best Multilingual | Alibaba | 256K | 84.3 | ⇧ | |
| #15 | Llama 5 Maverick | Meta | 2M | 83.6 | ⇧ | |
| #16 | DeepSeek R2Open Reasoning | DeepSeek | 256K | 83.1 | — | |
| #17 | Claude Haiku 4.5 | Anthropic | 200K | 81.9 | — | |
| #18 | GLM-5 | Zhipu AI | 256K | 81.4 | ⇧ | |
| #19 | Mistral Large 3.1Best EU Residency | Mistral | 256K | 79.2 | — | |
| #20 | Llama 5 ScoutBest Local | Meta | 1M | 77.6 | ⇧ |