Hybrid-reasoning DeepSeek model with thinking and non-thinking modes, sparse attention, and tool-use
24 providers
Every model in the catalog with source-linked pricing, context limits, provider availability, and published benchmark results.
Hybrid-reasoning DeepSeek model with thinking and non-thinking modes, sparse attention, and tool-use
DeepSeek reasoning model for multi-step analysis, math, coding, and tools
Thinking Kimi model for slower research passes, planning, and hard technical questions
Efficient Mistral model for fast chat, extraction, and production assistants
Sparse MoE Qwen model with 3B active parameters for efficient chat and reasoning
Dense open Qwen model for self-hosted chat, reasoning, and coding
Efficient Mistral-NVIDIA open model for multilingual chat and local deployment
| Model | Creator | Inputs | Context | Input | Output | Released | Compare |
|---|---|---|---|---|---|---|---|
| DeepSeek V3.2deepseek/deepseek-v3.2 | 128K | $0.18 | $0.35 | 2025-12-01 | |||
| DeepSeek Reasonerdeepseek/deepseek-reasoner | 1M | $0.147 | $0.295 | 2025-12-01 | |||
| Kimi K2 Thinkingmoonshotai/kimi-k2-thinking | 262.144K | $0.4 | $2.5 | 2025-11-06 | |||
| Mistral Small 3.2mistral/mistral-small-2506 | 128K | $0.1 | $0.3 | 2025-06-20 | |||
| Qwen3 30B A3Balibaba/qwen3-30b-a3b | 131.072K | $0.08 | $0.29 | 2025-04-28 | |||
| Qwen3 32Balibaba/qwen3-32b | 131.072K | $0.7 | $2.8 | 2025-04 | |||
| Mistral Nemomistral/mistral-nemo | 128K | $0.15 | $0.15 | 2024-07-01 |