Fast Mistral production model for chat, extraction, and cost-sensitive agents
12 providers
Every model in the catalog with source-linked pricing, context limits, provider availability, and published benchmark results.
Fast Mistral production model for chat, extraction, and cost-sensitive agents
Mistral's coding-agent model for repository work, terminal tasks, and software fixes
Mistral's largest general model for enterprise agents, coding, and multilingual reasoning
Smaller Qwen coder for efficient local agents and repo-level fixes
Popular open Llama workhorse for multilingual chat, coding, and self-hosting
| Model | Creator | Score | Inputs | Context | Input | Output | Released | Compare |
|---|---|---|---|---|---|---|---|---|
| Mistral Small 4mistral/mistral-small-2603 | 24.3 | 256K | $0.15 | $0.6 | 2026-03-16 | |||
| Devstral 2mistral/devstral-2512 | 23.7 | 262.144K | $0.4 | $2 | 2025-12-09 | |||
| Mistral Large 3mistral/mistral-large-2512 | 22.7 | 262.144K | $0.5 | $1.5 | 2024-11-01 | |||
| Qwen3-Coder 30B-A3B Instructalibaba/qwen3-coder-30b-a3b-instruct | 19.4 | 262.144K | $0.45 | $2.25 | 2025-04 | |||
| Llama-3.3-70B-Instructmeta/llama-3.3-70b-instruct | 10.7 | 128K | $0.1 | $0.32 | 2024-12-06 |