Newer StepFun flash model for faster agents, coding, and multimodal prompts
17 providers
Every model in the catalog with source-linked pricing, context limits, provider availability, and published benchmark results.
Newer StepFun flash model for faster agents, coding, and multimodal prompts
StepFun flash lane for quick multimodal reasoning and coding assistance
Late GLM-4 workhorse for coding agents, reasoning, and structured tasks
Flagship Qwen3 model for coding agents, complex reasoning, and tool use
Fast Mistral production model for chat, extraction, and cost-sensitive agents
Classic open reasoning model for transparent math, coding, and deliberate problem solving
Popular open Llama workhorse for multilingual chat, coding, and self-hosting
| Model | Creator | Score | Inputs | Context | Input | Output | Released | Compare |
|---|---|---|---|---|---|---|---|---|
| Step 3.7 Flashstepfun/step-3.7-flash | 35.6 | 256K | $0.185 | $1.11 | 2026-05-29 | |||
| Step 3.5 Flashstepfun/step-3.5-flash | 27.3 | 256K | $0.1 | $0.3 | 2026-01-29 | |||
| GLM-4.6zhipuai/glm-4.6 | 25.0 | 204.8K | $0.6 | $2.2 | 2025-09-30 | |||
| Qwen3 Maxalibaba/qwen3-max | 20.5 | 262.144K | $1.2 | $6 | 2025-09-23 | |||
| Mistral Small 4mistral/mistral-small-2603 | 17.4 | 256K | $0.15 | $0.6 | 2026-03-16 | |||
| DeepSeek-R1deepseek/deepseek-r1 | 6.1 | 128K | $0.7 | $2.5 | 2025-01-20 | |||
| Llama-3.3-70B-Instructmeta/llama-3.3-70b-instruct | 3.0 | 128K | $0.1 | $0.32 | 2024-12-06 |