StepFun flash lane for quick multimodal reasoning and coding assistance
Models
Every model in the catalog with source-linked pricing, context limits, provider availability, and published benchmark results.
Newer StepFun flash model for faster agents, coding, and multimodal prompts
Late GLM-4 workhorse for coding agents, reasoning, and structured tasks
Flagship Qwen3 model for coding agents, complex reasoning, and tool use
Fast Mistral production model for chat, extraction, and cost-sensitive agents
Mistral's largest general model for enterprise agents, coding, and multilingual reasoning
Classic open reasoning model for transparent math, coding, and deliberate problem solving
GPT model for general reasoning, writing, coding, and tool-assisted tasks
GPT model for general reasoning, writing, coding, and tool-assisted tasks
Mistral model for multilingual chat, reasoning, and tool-assisted workflows
Mistral's coding-agent model for repository work, terminal tasks, and software fixes
Compact GPT model for low-latency assistance and high-volume workloads
Smaller Qwen coder for efficient local agents and repo-level fixes
Popular open Llama workhorse for multilingual chat, coding, and self-hosting
Small omni GPT for cheap multimodal assistance and production-scale traffic
Fast web-grounded Sonar for current answers, citations, and lightweight retrieval
Deeper Sonar search model with broader retrieval and stronger synthesis
| Model | Creator | Score | Inputs | Context | Input | Output | Released | Compare |
|---|---|---|---|---|---|---|---|---|
| Step 3.5 Flashstepfun/step-3.5-flash | 40.4 | 256K | $0.1 | $0.3 | 2026-01-29 | |||
| Step 3.7 Flashstepfun/step-3.7-flash | 40.0 | 256K | $0.185 | $1.11 | 2026-05-29 | |||
| GLM-4.6zhipuai/glm-4.6 | 38.4 | 204.8K | $0.6 | $2.2 | 2025-09-30 | |||
| Qwen3 Maxalibaba/qwen3-max | 38.3 | 262.144K | $1.2 | $6 | 2025-09-23 | |||
| Mistral Small 4mistral/mistral-small-2603 | 38.0 | 256K | $0.15 | $0.6 | 2026-03-16 | |||
| Mistral Large 3mistral/mistral-large-2512 | 36.2 | 262.144K | $0.5 | $1.5 | 2024-11-01 | |||
| DeepSeek-R1deepseek/deepseek-r1 | 35.7 | 128K | $0.7 | $2.5 | 2025-01-20 | |||
| GPT-4o (2024-11-20)openai/gpt-4o-2024-11-20 | 33.3 | 128K | $2.5 | $10 | 2024-11-20 | |||
| GPT-4o (2024-08-06)openai/gpt-4o-2024-08-06 | 33.1 | 128K | $2.5 | $10 | 2024-08-06 | |||
| Mistral Medium 3mistral/mistral-medium-2505 | 33.1 | 131.072K | $0.4 | $2 | 2025-05-07 | |||
| Devstral 2mistral/devstral-2512 | 33.1 | 262.144K | $0.4 | $2 | 2025-12-09 | |||
| GPT-4 Turboopenai/gpt-4-turbo | 31.9 | 128K | $10 | $30 | 2023-11-06 | |||
| Qwen3-Coder 30B-A3B Instructalibaba/qwen3-coder-30b-a3b-instruct | 27.8 | 262.144K | $0.45 | $2.25 | 2025-04 | |||
| Llama-3.3-70B-Instructmeta/llama-3.3-70b-instruct | 26.0 | 128K | $0.1 | $0.32 | 2024-12-06 | |||
| GPT-4o miniopenai/gpt-4o-mini | 22.9 | 128K | $0.15 | $0.6 | 2024-07-18 | |||
| Sonarperplexity/sonar | 22.9 | 128K | $1 | $1 | 2024-01-01 | |||
| Sonar Properplexity/sonar-pro | 22.6 | 200K | $3 | $15 | 2024-01-01 |