General AI Reasoning
A balanced profile weighting reasoning, knowledge and instruction-following most heavily.
| Rank | Model | Score | Confidence | Evidence | Reasoning | Knowledge | Instruction Following |
|---|---|---|---|---|---|---|---|
| 1 | Claude Sonnet · Anthropic | 82.6 | 80% | 2,882 | 100 | 76 | 100 |
| 2 | GPT Frontier · OpenAI | 79.0 | 80% | 2,882 | 87 | 88 | 81 |
| 3 | Gemini Pro · Google DeepMind | 75.4 | 80% | 2,882 | 73 | 100 | 69 |
| 4 | Llama Open · Meta AIPartial evidence | 9.7 | 66% | 567 | 0 | 0 | 0 |
Scores are based on sample data for demonstration purposes.
Formula version 1.0.0 · weights v1.0.0 · calculated 2026-09-21. See the methodology page for how scores, confidence and evidence are combined.