Reasoning Grok for document-heavy analysis and long-horizon tool use
Models
Every model in the catalog with source-linked pricing, context limits, provider availability, and published benchmark results.
More exact GPT-5.4 tier for demanding professional reasoning and agent tasks
Agent-ready GPT for coding and computer-use workflows at a lower cost
Low-latency Gemini model for high-volume multimodal and agent workloads
Image model for prompt-driven generation, editing, and visual design workflows
Reasoning-first Gemini preview for agentic coding and complex problem solving
Advanced Gemini model for complex reasoning, coding, and multimodal analysis
Claude workhorse for coding agents, careful analysis, and production cost control
ByteDance Seed coding model for multimodal software engineering and long-running agents
Lightweight ByteDance Seed 2.0 model for low-latency multimodal reasoning and high-volume tasks
High-end Claude for difficult coding, planning, and slower expert reasoning
Coding-optimized GPT model for repository edits, reviews, and agentic software work
Earlier Kimi frontier model for long-context agents, coding, and multimodal work
New Gemini flash lane bringing frontier-style multimodal reasoning to cheaper runs
Reliable GPT generation for broad coding, writing, and tool-assisted product work
Higher-accuracy GPT-5.2 variant for tougher reasoning and review workflows
Code-specialist GPT for repository edits, reviews, and long-running software agents
GLM vision model for visual reasoning, documents, and multimodal agents
Flagship Claude model for deep reasoning, coding, and long-horizon agents
Nano Banana Pro for higher-fidelity image generation and design-heavy edits
Codex GPT for repository edits, code review, and practical software agents
Coding-optimized GPT model for repository edits, reviews, and agentic software work
Sharper GPT-5 generation for coding, product work, and tool-assisted tasks
Coding-optimized GPT model for repository edits, reviews, and agentic software work
Flagship Claude model for deep reasoning, coding, and long-horizon agents
Higher-accuracy GPT-5 tier for tough analysis, coding reviews, and planning
Qwen vision-language thinking model for visual reasoning, documents, and agent tasks
Qwen vision-language instruct model for visual reasoning, documents, and agent tasks
Nano Banana image model for fast generation, edits, and character-consistent assets
Original GPT-5 workhorse for reasoning, coding, writing, and tool workflows
Small GPT-5 for responsive agents, coding help, and everyday automation
Tiny GPT-5 lane for routing, extraction, classification, and bulk jobs
High-effort o3 tier for difficult technical reasoning and careful answers
Mistral model for multilingual chat, reasoning, and tool-assisted workflows
Deliberate o-series reasoner for hard math, coding, and multi-step analysis
Fast o-series model for compact reasoning, coding, and tool use
Affordable GPT-4.1 lane for fast coding help and structured extraction
Long-lived GPT workhorse for coding, instruction following, and production apps
Tiny GPT-4.1 option for classification, routing, and very high-volume tasks
O-series reasoning model for hard analysis, math, coding, and planning
O-series reasoning model for hard analysis, math, coding, and planning
Efficient model for low-latency assistance, extraction, and routine automation
Flagship model for demanding analysis, coding, and production agent workflows
GPT model for general reasoning, writing, coding, and tool-assisted tasks
Mistral's largest general model for enterprise agents, coding, and multilingual reasoning
Flagship Mistral model for advanced reasoning, coding, and multilingual work
Open multimodal Llama model for image understanding, captioning, and visual QA
GPT model for general reasoning, writing, coding, and tool-assisted tasks
Small omni GPT for cheap multimodal assistance and production-scale traffic
Omni-era GPT for multimodal chat, practical coding, and general assistants
| Model | Creator | Inputs | Context | Input | Output | Released | Compare |
|---|---|---|---|---|---|---|---|
| Grok 4.20 (Reasoning)xai/grok-4.20-0309-reasoning | 1M | $1.25 | $2.5 | 2026-03-09 | |||
| GPT-5.4 Proopenai/gpt-5.4-pro | 1.05M | $30 | $180 | 2026-03-05 | |||
| GPT-5.4openai/gpt-5.4 | 1.05M | $2.5 | $15 | 2026-03-05 | |||
| Gemini 3.1 Flash Lite Previewgoogle/gemini-3.1-flash-lite-preview | 1.04858M | $0.25 | $1.5 | 2026-03-03 | |||
| Nano Banana 2 Previewgoogle/gemini-3.1-flash-image-preview | 65.536K | $0.5 | $3 | 2026-02-26 | |||
| Gemini 3.1 Pro Previewgoogle/gemini-3.1-pro-preview | 1.04858M | $2 | $12 | 2026-02-19 | |||
| Gemini 3.1 Pro Preview Custom Toolsgoogle/gemini-3.1-pro-preview-customtools | 1.04858M | $2 | $12 | 2026-02-19 | |||
| Claude Sonnet 4.6anthropic/claude-sonnet-4-6 | 1M | $3 | $15 | 2026-02-17 | |||
| Seed 2.0 Codebytedance-seed/seed-2.0-code | 262.144K | $0.4 | $2.4 | 2026-02-14 | |||
| Seed 2.0 Minibytedance-seed/seed-2.0-mini | 256K | $0.03 | $0.297 | 2026-02-14 | |||
| Claude Opus 4.6anthropic/claude-opus-4-6 | 1M | $5 | $25 | 2026-02-05 | |||
| GPT-5.3 Codexopenai/gpt-5.3-codex | 400K | $1.75 | $14 | 2026-02-05 | |||
| Kimi K2.5moonshotai/kimi-k2.5 | 262.144K | $0.3 | $1.9 | 2026-01 | |||
| Gemini 3 Flash Previewgoogle/gemini-3-flash-preview | 1.04858M | $0.5 | $3 | 2025-12-17 | |||
| GPT-5.2openai/gpt-5.2 | 400K | $1.75 | $14 | 2025-12-11 | |||
| GPT-5.2 Proopenai/gpt-5.2-pro | 400K | $21 | $168 | 2025-12-11 | |||
| GPT-5.2 Codexopenai/gpt-5.2-codex | 400K | $0.14 | $1.14 | 2025-12-11 | |||
| GLM-4.6Vzhipuai/glm-4.6v | 128K | $0.3 | $0.9 | 2025-12-08 | |||
| Claude Opus 4.5 (latest)anthropic/claude-opus-4-5 | 200K | $5 | $25 | 2025-11-24 | |||
| Nano Banana Pro Previewgoogle/gemini-3-pro-image-preview | 65.536K | $1 | $6 | 2025-11-20 | |||
| GPT-5.1 Codexopenai/gpt-5.1-codex | 400K | $1.07 | $8.5 | 2025-11-13 | |||
| GPT-5.1 Codex miniopenai/gpt-5.1-codex-mini | 400K | $0.22 | $1.8 | 2025-11-13 | |||
| GPT-5.1openai/gpt-5.1 | 400K | $1.25 | $10 | 2025-11-13 | |||
| GPT-5.1 Codex Maxopenai/gpt-5.1-codex-max | 400K | $1.1 | $9 | 2025-11-13 | |||
| Claude Opus 4.5anthropic/claude-opus-4-5-20251101 | 200K | $5 | $25 | 2025-11-01 | |||
| GPT-5 Proopenai/gpt-5-pro | 400K | $15 | $120 | 2025-10-06 | |||
| Qwen3 VL 235B A22B Thinkingalibaba/qwen3-vl-235b-a22b-thinking | 131.072K | $0.4 | $4 | 2025-09-23 | |||
| Qwen3 VL 235B A22B Instructalibaba/qwen3-vl-235b-a22b-instruct | 131.072K | $0.2 | $0.88 | 2025-09-23 | |||
| Nano Bananagoogle/gemini-2.5-flash-image | 32.768K | $0.3 | $30 | 2025-08-26 | |||
| GPT-5openai/gpt-5 | 400K | $1.25 | $10 | 2025-08-07 | |||
| GPT-5 Miniopenai/gpt-5-mini | 400K | $0.25 | $2 | 2025-08-07 | |||
| GPT-5 Nanoopenai/gpt-5-nano | 400K | $0.05 | $0.4 | 2025-08-07 | |||
| o3-proopenai/o3-pro | 200K | $20 | $80 | 2025-06-10 | |||
| Mistral Medium 3mistral/mistral-medium-2505 | 131.072K | $0.4 | $2 | 2025-05-07 | |||
| o3openai/o3 | 200K | $2 | $8 | 2025-04-16 | |||
| o4-miniopenai/o4-mini | 200K | $1.1 | $4.4 | 2025-04-16 | |||
| GPT-4.1 miniopenai/gpt-4.1-mini | 1.04758M | $0.4 | $1.6 | 2025-04-14 | |||
| GPT-4.1openai/gpt-4.1 | 1.04758M | $2 | $8 | 2025-04-14 | |||
| GPT-4.1 nanoopenai/gpt-4.1-nano | 1.04758M | $0.1 | $0.4 | 2025-04-14 | |||
| o1-proopenai/o1-pro | 200K | $150 | $600 | 2025-03-19 | |||
| o1openai/o1 | 200K | $15 | $60 | 2024-12-05 | |||
| Nova Liteamazon/nova-lite | 300K | $0.06 | $0.24 | 2024-12-03 | |||
| Nova Proamazon/nova-pro | 300K | $0.8 | $3.2 | 2024-12-03 | |||
| GPT-4o (2024-11-20)openai/gpt-4o-2024-11-20 | 128K | $2.5 | $10 | 2024-11-20 | |||
| Mistral Large 3mistral/mistral-large-2512 | 262.144K | $0.5 | $1.5 | 2024-11-01 | |||
| Mistral Large (latest)mistral/mistral-large-latest | 262.144K | $0.5 | $1.5 | 2024-11-01 | |||
| Llama-3.2-11B-Vision-Instructmeta/llama-3.2-11b-vision-instruct | 128K | $0.055 | $0.055 | 2024-09-25 | |||
| GPT-4o (2024-08-06)openai/gpt-4o-2024-08-06 | 128K | $2.5 | $10 | 2024-08-06 | |||
| GPT-4o miniopenai/gpt-4o-mini | 128K | $0.15 | $0.6 | 2024-07-18 | |||
| GPT-4oopenai/gpt-4o | 128K | $2.5 | $10 | 2024-05-13 |