01 Market & Economics

Current roster, benchmark register, capability radar, quota-burn cross-matrix, subscription tiers, and agent policy.

Current roster

A deliberately compact provider roster: flagship, balanced or fast, and differentiated specialist models. Search, filter, or sort the columns; linked values open the primary provider evidence.

ModelProviderScopeStatusControl / defaultAvailable levelsSpeedSpeed evidenceAPI input / outputContextRegionAvailability

Benchmark register

Quality % is computed against the selected reference (top-right). SWE-bench Pro is the trustworthy benchmark; Verified is contaminated.

ModelProviderQuality vs refSWE-ProSWE-VerLCBAIMEIn $/MtokOut $/MtokCacheContexttok/sReleased

Capability radar

Compare six editorial capability axes for two selectable models. Choose a focus to increase that axis's weight in the capability score; the quality column remains the separate benchmark-derived value relative to the selected reference.

Quota burn cross-matrix

Burn ratios shown as multiples of the selected reference model at medium effort. OpenAI (Codex /effort) and Anthropic (Claude Code /effort) share the same vocabulary — low / medium / high / xhigh — with Anthropic adding 'max' for Opus 4.7. Google uses thinking budgets. Multipliers stack: Fast mode ×2.0, cached input ×0.6, plan mode forces high. Switch the reference dropdown at the top of the page to recompute all ratios.

Subscription tiers

ProviderTier$/moLimitsModelsFeatures

Automated agent policy

ProviderSub on automationEnforcementFirst-party exceptionAPI needed?