StepFun flash lane for quick multimodal reasoning and coding assistance
14 providers
Every model in the catalog with source-linked pricing, context limits, provider availability, and published benchmark results.
StepFun flash lane for quick multimodal reasoning and coding assistance
Newer StepFun flash model for faster agents, coding, and multimodal prompts
Late GLM-4 workhorse for coding agents, reasoning, and structured tasks
Flagship Qwen3 model for coding agents, complex reasoning, and tool use
Fast Mistral production model for chat, extraction, and cost-sensitive agents
Classic open reasoning model for transparent math, coding, and deliberate problem solving
Popular open Llama workhorse for multilingual chat, coding, and self-hosting
| Model | Creator | Score | Inputs | Context | Input | Output | Released | Compare |
|---|---|---|---|---|---|---|---|---|
| Step 3.5 Flashstepfun/step-3.5-flash | 40.4 | 256K | $0.1 | $0.3 | 2026-01-29 | |||
| Step 3.7 Flashstepfun/step-3.7-flash | 40.0 | 256K | $0.185 | $1.11 | 2026-05-29 | |||
| GLM-4.6zhipuai/glm-4.6 | 38.4 | 204.8K | $0.6 | $2.2 | 2025-09-30 | |||
| Qwen3 Maxalibaba/qwen3-max | 38.3 | 262.144K | $1.2 | $6 | 2025-09-23 | |||
| Mistral Small 4mistral/mistral-small-2603 | 38.0 | 256K | $0.15 | $0.6 | 2026-03-16 | |||
| DeepSeek-R1deepseek/deepseek-r1 | 35.7 | 128K | $0.7 | $2.5 | 2025-01-20 | |||
| Llama-3.3-70B-Instructmeta/llama-3.3-70b-instruct | 26.0 | 128K | $0.1 | $0.32 | 2024-12-06 |