Fast Gemini model balancing multimodal reasoning, tool use, and cost
Models
Every model in the catalog with source-linked pricing, context limits, provider availability, and published benchmark results.
Low-latency Gemini model for high-volume multimodal and agent workloads
Fast DeepSeek V4 lane for economical reasoning, coding, and long-context work
Open MoE flagship with million-token context for coding and long agent runs
Default frontier GPT for coding, computer use, research, and knowledge work
Highest-accuracy GPT-5.5 tier for slower, precision-heavy reasoning and coding
Qwen vision-language model for visual reasoning, documents, and agent tasks
Stronger MiMo Pro tier for multimodal reasoning and coding-agent execution
Open MiMo model for multimodal coding agents and long-context automation
Multimodal Kimi workhorse for agent loops, coding tasks, and visual context
xAI's default Grok for chat, coding, agentic tools, and lower hallucination risk
Stronger Opus tier for advanced software work and high-stakes reasoning
Fast Grok coding model tuned for agentic engineering and iterative edits
Muse Spark is a natively multimodal reasoning model with support for tool-use, visual chain of thought, and multi-agent orchestration.
Earlier Qwen multimodal workhorse for million-token agent and document tasks
Largest Gemma 4 instruction model for open, self-hosted chat and reasoning
Open Gemma instruction model for efficient chat and self-hosted deployments
Reasoning-optimized 398B MoE agent model with extended thinking for long-horizon and multi-turn tool use
Earlier MiMo Pro model for multimodal agents, reasoning, and code tasks
Cheapest GPT-5.4 lane for simple routing, extraction, and bulk automation
Strong small GPT for coding subagents, quick tool use, and high-volume work
Efficient Mistral model for fast chat, extraction, and production assistants
Nemotron middle tier for collaborative agents and high-volume reasoning workloads
Agent-ready GPT for coding and computer-use workflows at a lower cost
More exact GPT-5.4 tier for demanding professional reasoning and agent tasks
Low-latency Gemini model for high-volume multimodal and agent workloads
Image model for prompt-driven generation, editing, and visual design workflows
Qwen vision-language model for visual reasoning, documents, and agent tasks
Reasoning-first Gemini preview for agentic coding and complex problem solving
Claude workhorse for coding agents, careful analysis, and production cost control
Qwen vision-language model for visual reasoning, documents, and agent tasks
General GLM flagship for coding, analysis, and tool-heavy engineering workflows
High-end Claude for difficult coding, planning, and slower expert reasoning
Coding-optimized GPT model for repository edits, reviews, and agentic software work
Open-weight Qwen coding model for agents, repository edits, and multi-turn tool use
StepFun flash lane for quick multimodal reasoning and coding assistance
Earlier Kimi frontier model for long-context agents, coding, and multimodal work
ByteDance Seed model for multimodal reasoning, long-context analysis, and agent workflows
MiMo flash model for fast multimodal assistance and agent workflows
Small Nemotron 3 MoE for efficient coding, math, and long-context agents
Code-specialist GPT for repository edits, reviews, and long-running software agents
Reliable GPT generation for broad coding, writing, and tool-assisted product work
Higher-accuracy GPT-5.2 variant for tougher reasoning and review workflows
Mistral's coding-agent model for repository work, terminal tasks, and software fixes
Compact multimodal coding model for repository exploration, file editing, and software agents
Lightweight GLM vision model for visual reasoning, documents, and multimodal agents
GLM vision model for visual reasoning, documents, and multimodal agents
Multimodal reasoning model for visual analysis, planning, and tool use
Compact multimodal Mistral model for local assistants, edge agents, and efficient tool use
Reasoning-tuned 26B MoE model with 3B active parameters for agents, tools, and multi-step workloads
| Model | Creator | Inputs | Context | Input | Output | Released | Compare |
|---|---|---|---|---|---|---|---|
| Gemini 3.5 Flashgoogle/gemini-3.5-flash | 1.04858M | $1.5 | $9 | 2026-05-19 | |||
| Gemini 3.1 Flash Litegoogle/gemini-3.1-flash-lite | 1.04858M | $0.25 | $1.5 | 2026-05-07 | |||
| DeepSeek V4 Flashdeepseek/deepseek-v4-flash | 1M | $0.15 | $0.6 | 2026-04-24 | |||
| DeepSeek V4 Prodeepseek/deepseek-v4-pro | 1M | $0.435 | $0.87 | 2026-04-24 | |||
| GPT-5.5openai/gpt-5.5 | 1.05M | $5 | $30 | 2026-04-23 | |||
| GPT-5.5 Proopenai/gpt-5.5-pro | 1.05M | $30 | $180 | 2026-04-23 | |||
| Qwen3.6 27Balibaba/qwen3.6-27b | 262.144K | $0.6 | $3.6 | 2026-04-22 | |||
| MiMo-V2.5-Proxiaomi/mimo-v2.5-pro | 1.04858M | $0.435 | $0.87 | 2026-04-22 | |||
| MiMo-V2.5xiaomi/mimo-v2.5 | 1.04858M | $0.14 | $0.28 | 2026-04-22 | |||
| Kimi K2.6moonshotai/kimi-k2.6 | 262.144K | $0.95 | $4 | 2026-04-21 | |||
| Grok 4.3xai/grok-4.3 | 1M | $1.25 | $2.5 | 2026-04-17 | |||
| Claude Opus 4.7anthropic/claude-opus-4-7 | 1M | $5 | $25 | 2026-04-16 | |||
| Grok Build 0.1xai/grok-build-0.1 | 256K | $1 | $2 | 2026-04-16 | |||
| Muse Spark 1.1meta/muse-spark-1.1 | 1.04858M | $1.25 | $4.25 | 2026-04-08 | |||
| Qwen3.6 Plusalibaba/qwen3.6-plus | 1M | $0.5 | $3 | 2026-04-02 | |||
| Gemma 4 31B ITgoogle/gemma-4-31b-it | 262.144K | $0.09 | $0.34 | 2026-04-02 | |||
| Gemma 4 26B A4B ITgoogle/gemma-4-26b-a4b-it | 262.144K | $0.042 | $0.22 | 2026-04-02 | |||
| Trinity Large Thinkingarcee-ai/trinity-large-thinking | 524.288K | $0.25 | $0.9 | 2026-04-01 | |||
| MiMo-V2-Proxiaomi/mimo-v2-pro | 1.04858M | $0.435 | $0.87 | 2026-03-18 | |||
| GPT-5.4 nanoopenai/gpt-5.4-nano | 400K | $0.2 | $1.25 | 2026-03-17 | |||
| GPT-5.4 miniopenai/gpt-5.4-mini | 400K | $0.75 | $4.5 | 2026-03-17 | |||
| Mistral Small (latest)mistral/mistral-small-latest | 256K | $0.15 | $0.6 | 2026-03-16 | |||
| Nemotron 3 Super 120B A12Bnvidia/nemotron-3-super-120b-a12b | 262.144K | $0.2 | $0.8 | 2026-03-11 | |||
| GPT-5.4openai/gpt-5.4 | 1.05M | $2.5 | $15 | 2026-03-05 | |||
| GPT-5.4 Proopenai/gpt-5.4-pro | 1.05M | $30 | $180 | 2026-03-05 | |||
| Gemini 3.1 Flash Lite Previewgoogle/gemini-3.1-flash-lite-preview | 1.04858M | $0.25 | $1.5 | 2026-03-03 | |||
| Nano Banana 2 Previewgoogle/gemini-3.1-flash-image-preview | 65.536K | $0.5 | $3 | 2026-02-26 | |||
| Qwen3.5 Flashalibaba/qwen3.5-flash | 1M | $0.029 | $0.287 | 2026-02-23 | |||
| Gemini 3.1 Pro Previewgoogle/gemini-3.1-pro-preview | 1.04858M | $2 | $12 | 2026-02-19 | |||
| Claude Sonnet 4.6anthropic/claude-sonnet-4-6 | 1M | $3 | $15 | 2026-02-17 | |||
| Qwen3.5 Plusalibaba/qwen3.5-plus | 1M | $0.4 | $2.4 | 2026-02-16 | |||
| GLM-5zhipuai/glm-5 | 204.8K | $1 | $3.2 | 2026-02-12 | |||
| Claude Opus 4.6anthropic/claude-opus-4-6 | 1M | $5 | $25 | 2026-02-05 | |||
| GPT-5.3 Codexopenai/gpt-5.3-codex | 400K | $1.75 | $14 | 2026-02-05 | |||
| Qwen3 Coder Nextalibaba/qwen3-coder-next | 262.144K | $0.108 | $0.675 | 2026-02-03 | |||
| Step 3.5 Flashstepfun/step-3.5-flash | 256K | $0.1 | $0.3 | 2026-01-29 | |||
| Kimi K2.5moonshotai/kimi-k2.5 | 262.144K | $0.3 | $1.9 | 2026-01 | |||
| Seed 1.8bytedance-seed/seed-1-8 | 256K | $0.119 | $1.187 | 2025-12-28 | |||
| MiMo-V2-Flashxiaomi/mimo-v2-flash | 262.144K | $0.14 | $0.28 | 2025-12-16 | |||
| Nemotron 3 Nano 30B A3Bnvidia/nemotron-3-nano-30b-a3b | 262.144K | $0.05 | $0.2 | 2025-12-15 | |||
| GPT-5.2 Codexopenai/gpt-5.2-codex | 400K | $0.14 | $1.14 | 2025-12-11 | |||
| GPT-5.2openai/gpt-5.2 | 400K | $1.75 | $14 | 2025-12-11 | |||
| GPT-5.2 Proopenai/gpt-5.2-pro | 400K | $21 | $168 | 2025-12-11 | |||
| Devstral 2mistral/devstral-2512 | 262.144K | $0.4 | $2 | 2025-12-09 | |||
| Devstral Small 2mistral/devstral-small-2 | 262.144K | $0.1 | $0.3 | 2025-12-09 | |||
| GLM-4.6V-Flashzhipuai/glm-4.6v-flash | 128K | $0.3 | $0.9 | 2025-12-08 | |||
| GLM-4.6Vzhipuai/glm-4.6v | 128K | $0.3 | $0.9 | 2025-12-08 | |||
| Nova 2 Liteamazon/nova-2-lite | 1M | $0.3 | $2.5 | 2025-12-02 | |||
| Ministral 14Bmistral/ministral-14b | 262.144K | $0.2 | $0.2 | 2025-12-02 | |||
| Trinity Miniarcee-ai/trinity-mini | 131.072K | $0.045 | $0.15 | 2025-12-01 |