01 Market & Economics
Current roster
A deliberately compact provider roster: flagship, balanced or fast, and differentiated specialist models. Search, filter, or sort the columns; linked values open the primary provider evidence.
| Model | Provider | Scope | Status | Control / default | Available levels | Speed | Speed evidence | API input / output | Context | Region | Availability |
|---|
Benchmark register
Quality % is computed against the selected reference (top-right). SWE-bench Pro is the trustworthy benchmark; Verified is contaminated.
| Model | Provider | Quality vs ref | SWE-Pro | SWE-Ver | LCB | AIME | In $/Mtok | Out $/Mtok | Cache | Context | tok/s | Released |
|---|
Capability radar
Compare six editorial capability axes for two selectable models. Choose a focus to increase that axis's weight in the capability score; the quality column remains the separate benchmark-derived value relative to the selected reference.
Quota burn cross-matrix
Burn ratios shown as multiples of the selected reference model at medium effort. OpenAI (Codex /effort) and Anthropic (Claude Code /effort) share the same vocabulary — low / medium / high / xhigh — with Anthropic adding 'max' for Opus 4.7. Google uses thinking budgets. Multipliers stack: Fast mode ×2.0, cached input ×0.6, plan mode forces high. Switch the reference dropdown at the top of the page to recompute all ratios.
Subscription tiers
| Provider | Tier | $/mo | Limits | Models | Features |
|---|
Automated agent policy
| Provider | Sub on automation | Enforcement | First-party exception | API needed? |
|---|