Capability badges appear only where the provider catalog explicitly lists support. A blank cell means the source is silent, not that the feature is absent.
Listed rates
Price across providers
Input price per million tokens as published by each provider. Bars are drawn from listed rates only — no traffic weighting, since the catalog observes no requests.
Lowest input$5/M
Across 37 priced providers
Median input$5/M
Midpoint of listed rates
Highest input$6/M
1.2× the lowest listed rate
Output range$25 – $30
Per million output tokens
LLM Gateway$5/M
Cloudflare AI Gateway$5/M
GitHub Copilot$5/M
Impossibl$5/M
DigitalOcean$5/M
OpenCode Zen$5/M
Neon$5/M
Perplexity Agent$5/M
Kilo Gateway$5/M
Merge Gateway$5/M
Auriko$5/M
Databricks$5/M
Requesty$5/M
NEAR AI Cloud$5/M
Eden AI$5/M
Vertex (Anthropic)$5/M
OrcaRouter$5/M
OpenRouter$5/M
Anthropic$5/M
Cloudflare AI Gateway$5/M
Azure$5/M
AIHubMix$5/M
Vertex$5/M
GMI Cloud$5/M
FreeModel$5/M
Amazon Bedrock$5/M
Requesty$5/M
Vercel AI Gateway$5/M
Abacus$5/M
Ofox$5/M
FrogBot$5/M
Pioneer$5/M
Azure Cognitive Services$5/M
Opper$5/M
DevPass (LLM Gateway)$5/M
Cortecs$5.313/M
Venice AI$6/M
Cost calculator
Estimate a workload
$0.00
Excludes taxes, non-token charges, and tiered discounts.
35 providers list the identical $5 input rate, so price alone will not separate them — compare context limits, max output, and capabilities above.
Context limits also differ by provider, from 200K to 1M tokens. Compare the provider table above before choosing on price alone.
Specification
Capabilities
Recorded from the source catalog and provider listings.
Every result stays attached to its source, version, metric and harness. Scores from different versions are never merged, and the profile below plots each benchmark against its own population rather than on a shared scale.
Each axis is this model's percentile among the 7 benchmarks it has published results for, measured against every other model with a score on that same benchmark. Percentiles are used because benchmarks do not share a scale — a 60 on one is not a 60 on another. Hover any point for the raw score.
Terminal-Bench — 2.1
pass@1 · Claude Code
Source ↗
Terminal-Bench#2 of 6
70.2
63.1median 64.971.4
SWE-Atlas Refactoring
score · Claude Code
Source ↗
SWE-Atlas Refactoring#2 of 3
35.58
32.2median 35.648.6
SWE-Atlas Codebase QnA
score · Claude Code
Source ↗
33.32 models scored — too few for a distribution
SWE-Atlas Test Writing
score · Claude Code
Source ↗
36.672 models scored — too few for a distribution
Each strip shows every published score for that benchmark, with this model marked. The lighter shape behind the ticks is the density of results, and the dashed line is the median.
Cost against capability
Price and performance
Listed input price plotted against SWE-Bench Pro, the benchmark with the widest published coverage that this model appears in.
The stepped line is the efficient frontier: at each price, the best score available for that money or less. A model sitting on it is not being beaten by anything cheaper. Price is log-scaled because listed rates span four orders of magnitude. Only models with both a listed price and a score on this benchmark can appear.
Catalog activity
Change log
Field-level changes detected between successful source imports.