Cost & Latency Index
Sample dataOperational efficiency
Definition & methodology
- Methodology
- Sample methodology: median response latency and per-million-token cost under standard load.
- Dataset
- Sample workload of representative prompts of varying length.
- Metric
- composite index (lower is better) (lower is better)
- Version
- 1.0
- Last updated
- 2026-09-01
- Limitations
- Sample limitation note: pricing changes frequently and this figure may be stale.
Model results
| Model | Score | Sample size | Evidence | Evaluated |
|---|---|---|---|---|
| Claude Sonnet | 42 index | — | L1Source | 2026-09-01 |
| GPT Frontier | 48 index | — | L1Source | 2026-09-01 |