Step 3.7 Flash is StepFun's latest high-efficiency multimodal Mixture-of-Experts model. It pairs a 196B-parameter language backbone with a vision encoder for native image and video understanding, activating roughly 11B parameters...
Models
Every model in the catalog with source-linked pricing, context limits, provider availability, and published benchmark results.
Step 3.5 Flash is StepFun's most capable open-source foundation model. Built on a sparse Mixture of Experts (MoE) architecture, it selectively activates only 11B of its 196B parameters per token....
Mistral's coding-agent model for repository work, terminal tasks, and software fixes
Mistral's largest general model for enterprise agents, coding, and multilingual reasoning
DeepSeek R1 is here: Performance on par with [OpenAI o1](/openai/o1), but open-sourced and with fully open reasoning tokens. It's 671B parameters in size, with 37B active in an inference pass....
Popular open Llama workhorse for multilingual chat, coding, and self-hosting
| Model | Creator | Score | Inputs | Context | Input | Output | Released | Compare |
|---|---|---|---|---|---|---|---|---|
| StepFun: Step 3.7 Flashstepfun/step-3.7-flash | 35.6 | 256K | $0.2 | $1.15 | 2026-05-29 | |||
| StepFun: Step 3.5 Flashstepfun/step-3.5-flash | 27.3 | 262.144K | $0.1 | $0.3 | 2026-01-29 | |||
| Devstral 2mistral/devstral-2512 | 18.9 | 262.144K | $0.4 | $2 | 2025-12-09 | |||
| Mistral Large 3mistral/mistral-large-2512 | 15.9 | 262.144K | $0.5 | $1.5 | 2024-11-01 | |||
| DeepSeek: R1deepseek/deepseek-r1 | 6.1 | 64K | $0.7 | $2.5 | 2025-01-20 | |||
| Llama-3.3-70B-Instructmeta/llama-3.3-70b-instruct | 3.0 | 128K | $0.1 | $0.32 | 2024-12-06 |