Hybrid-reasoning DeepSeek model with thinking and non-thinking modes, sparse attention, and tool-use
Models
Every model in the catalog with source-linked pricing, context limits, provider availability, and published benchmark results.
DeepSeek-V3.2-Exp is an experimental large language model released by DeepSeek as an intermediate step between V3.1 and future architectures. It introduces DeepSeek Sparse Attention (DSA), a fine-grained sparse attention mechanism...
Claude Haiku 4.5 is Anthropic’s fastest and most efficient model, delivering near-frontier intelligence at a fraction of the cost and latency of larger Claude models. Matching Claude Sonnet 4’s performance...
Claude Haiku 4.5 is Anthropic’s fastest and most efficient model, delivering near-frontier intelligence at a fraction of the cost and latency of larger Claude models. Matching Claude Sonnet 4’s performance...
Fast Gemini workhorse for multimodal apps where latency and price matter
Gemini 2.5 Flash is Google's state-of-the-art workhorse model, specifically designed for advanced reasoning, coding, mathematics, and scientific tasks. It includes built-in "thinking" capabilities, enabling it to provide responses with greater...
Reasoning-optimized 398B MoE agent model with extended thinking for long-horizon and multi-turn tool use
Qwen3-Max is an updated release built on the Qwen3 series, offering major improvements in reasoning, instruction following, multilingual support, and long-tail knowledge coverage compared to the January 2025 version. It...
Mistral Large 3 2512 is Mistral’s most capable model to date, featuring a sparse mixture-of-experts architecture with 41B active parameters (675B total), and released under the Apache 2.0 license.
Mistral Large 3 2512 is Mistral’s most capable model to date, featuring a sparse mixture-of-experts architecture with 41B active parameters (675B total), and released under the Apache 2.0 license.
Mistral Medium 3.1 is an updated version of Mistral Medium 3, which is a high-performance enterprise-grade language model designed to deliver frontier-level capabilities at significantly reduced operational cost. It balances...
Mistral Medium 3.1 is an updated version of Mistral Medium 3, which is a high-performance enterprise-grade language model designed to deliver frontier-level capabilities at significantly reduced operational cost. It balances...
Mercury 2 is an extremely fast reasoning LLM, and the first reasoning diffusion LLM (dLLM). Instead of generating tokens sequentially, Mercury 2 produces and refines multiple tokens in parallel, achieving...
Coding-optimized GPT model for repository edits, reviews, and agentic software work
DeepSeek chat model for instruction following, coding, and analysis
DeepSeek-V3.1 is a large hybrid reasoning model (671B parameters, 37B active) that supports both thinking and non-thinking modes via prompt templates. It extends the DeepSeek-V3 base with a two-phase long-context...
| Model | Creator | Score | Inputs | Context | Input | Output | Released | Compare |
|---|---|---|---|---|---|---|---|---|
| DeepSeek V3.2deepseek/deepseek-v3.2 | 1058.0 | 128K | $0.18 | $0.35 | 2025-12-01 | |||
| DeepSeek: DeepSeek V3.2 Expdeepseek/deepseek-v3.2-exp | 1057.0 | 163.84K | $0.27 | $0.41 | — | |||
| Anthropic: Claude Haiku 4.5 (batch)anthropic/claude-haiku-4.5:batch | 1051.0 | 200K | $0.5 | $2.5 | — | |||
| Anthropic: Claude Haiku 4.5anthropic/claude-haiku-4.5 | 1051.0 | 200K | $1 | $5 | — | |||
| Gemini 2.5 Flashgoogle/gemini-2.5-flash | 1045.0 | 1.04858M | $0.3 | $2.5 | 2025-06-17 | |||
| Google: Gemini 2.5 Flash (batch)google/gemini-2.5-flash:batch | 1045.0 | 1.04858M | $0.15 | $1.25 | — | |||
| Trinity Large Thinkingarcee-ai/trinity-large-thinking | 1041.0 | 524.288K | $0.25 | $0.9 | 2026-04-01 | |||
| Qwen: Qwen3 Maxqwen/qwen3-max | 1036.0 | 262.144K | $0.78 | $3.9 | — | |||
| Mistral: Mistral Large 3 2512 (batch)mistralai/mistral-large-2512:batch | 1018.0 | 262.144K | $0.25 | $0.75 | — | |||
| Mistral: Mistral Large 3 2512mistralai/mistral-large-2512 | 1018.0 | 262.144K | $0.5 | $1.5 | — | |||
| Mistral: Mistral Medium 3.1mistralai/mistral-medium-3.1 | 1018.0 | 131.072K | $0.4 | $2 | — | |||
| Mistral: Mistral Medium 3.1 (batch)mistralai/mistral-medium-3.1:batch | 1018.0 | 131.072K | $0.2 | $1 | — | |||
| Inception: Mercury 2inception/mercury-2 | 1012.0 | 128K | $0.25 | $0.75 | — | |||
| GPT-5.1 Codex miniopenai/gpt-5.1-codex-mini | 1003.0 | 400K | $0.22 | $1.8 | 2025-11-13 | |||
| DeepSeek Chatdeepseek/deepseek-chat | 1002.0 | 1M | $0.147 | $0.295 | 2025-12-01 | |||
| DeepSeek: DeepSeek V3.1deepseek/deepseek-chat-v3.1 | 992.0 | 163.84K | $0.25 | $0.95 | — |