Grok model for agentic tool use, reasoning, coding, and live assistance
Models
Every model in the catalog with source-linked pricing, context limits, provider availability, and published benchmark results.
Reasoning Grok for document-heavy analysis and long-horizon tool use
GPT-5.4 is OpenAI’s latest frontier model, unifying the Codex and GPT lines into a single system. It features a 1M+ token context window (922K input, 128K output) with support for...
GPT-5.4 Pro is OpenAI's most advanced model, building on GPT-5.4's unified architecture with enhanced reasoning capabilities for complex, high-stakes tasks. It features a 1M+ token context window (922K input, 128K...
Chat-tuned GPT model for conversational assistance, writing, and tool workflows
Gemini 3.1 Flash Lite Preview is Google's high-efficiency model optimized for high-volume use cases. It outperforms Gemini 2.5 Flash Lite on overall quality and approaches Gemini 2.5 Flash performance across...
Gemini 3.1 Flash Image Preview, a.k.a. "Nano Banana 2," is Google’s latest state of the art image generation and editing model, delivering Pro-level visual quality at Flash speed. It combines...
Qwen vision-language model for visual reasoning, documents, and agent tasks
Qwen instruction model for multilingual chat, reasoning, and tool use
Qwen vision-language model for visual reasoning, documents, and agent tasks
Qwen vision-language model for visual reasoning, documents, and agent tasks
Qwen vision-language model for visual reasoning, documents, and agent tasks
Gemini 3.1 Pro Preview is Google’s frontier reasoning model, delivering enhanced software engineering performance, improved agentic reliability, and more efficient token usage across complex workflows. Building on the multimodal foundation...
Gemini 3.1 Pro Preview Custom Tools is a variant of Gemini 3.1 Pro that improves tool selection behavior by preventing overuse of a general bash tool when more efficient third-party...
Turkish-language multimodal instruct model built on Gemma 3 12B for e-commerce text, chat, and image-text tasks
Claude workhorse for coding agents, careful analysis, and production cost control
Qwen vision-language model for visual reasoning, documents, and agent tasks
Large open Qwen multimodal MoE for visual agents and long technical tasks
Flagship ByteDance Seed 2.0 model for complex multimodal reasoning and long-horizon agent workflows
Seed-2.0-mini targets latency-sensitive, high-concurrency, and cost-sensitive scenarios, emphasizing fast response and flexible inference deployment. It delivers performance comparable to ByteDance-Seed-1.6, supports 256k context, four reasoning effort modes (minimal/low/medium/high), multimodal understanding,...
Seed-2.0-Lite is a versatile, cost‑efficient enterprise workhorse that delivers strong multimodal and agent capabilities while offering noticeably lower latency, making it a practical default choice for most production workloads across...
Seed 2.0 Code is a model from ByteDance Seed optimized for agentic coding. It is suited for frontend development, multilingual programming tasks, and coding-agent workflows in tools such as Claude...
Embedding model for semantic search, retrieval, clustering, and ranking pipelines
High-end Claude for difficult coding, planning, and slower expert reasoning
Coding-optimized GPT model for repository edits, reviews, and agentic software work
GPT-5.3-Codex is OpenAI’s most advanced agentic coding model, combining the frontier software engineering performance of GPT-5.2-Codex with the broader reasoning and professional knowledge capabilities of GPT-5.2. It achieves state-of-the-art results...
High-accuracy OCR model for extracting text from documents, screenshots, receipts, and natural scenes
Kimi K2.5 is Moonshot AI's native multimodal model, delivering state-of-the-art visual coding capability and a self-directed agent swarm paradigm. Built on Kimi K2 with continued pretraining over approximately 15T mixed...
ByteDance Seed model for multimodal reasoning, long-context analysis, and agent workflows
Gemini 3 Flash Preview is a high speed, high value thinking model designed for agentic workflows, multi turn chat, and coding assistance. It delivers near Pro level reasoning and tool...
GPT-5.2 Pro is OpenAI’s most advanced model, offering major improvements in agentic coding and long context performance over GPT-5 Pro. It is optimized for complex tasks that require step-by-step reasoning,...
GPT-5.2 is the latest frontier-grade model in the GPT-5 series, offering stronger agentic and long context perfomance compared to GPT-5.1. It uses adaptive reasoning to allocate computation dynamically, responding quickly...
Chat-tuned GPT model for conversational assistance, writing, and tool workflows
GPT-5.2-Codex is an upgraded version of GPT-5.1-Codex optimized for software engineering and coding workflows. It is designed for both interactive development sessions and long, independent execution of complex engineering tasks....
Compact multimodal coding model for repository exploration, file editing, and software agents
GLM vision model for visual reasoning, documents, and multimodal agents
Lightweight GLM vision model for visual reasoning, documents, and multimodal agents
Multimodal reasoning model for visual analysis, planning, and tool use
Compact multimodal Mistral model for local assistants, edge agents, and efficient tool use
Image model for prompt-driven generation, editing, and visual design workflows
Flagship Claude model for deep reasoning, coding, and long-horizon agents
Nano Banana Pro is Google’s most advanced image-generation and editing model, built on Gemini 3 Pro. It extends the original Nano Banana with significantly improved multimodal reasoning, real-world grounding, and...
xAI's fast agentic tool-calling model with a 2M context window; non-reasoning variant for low-latency responses
xAI's fast agentic tool-calling model with a 2M context window and built-in reasoning
Preview Gemini flagship for complex reasoning, coding, and rich multimodal prompts
GPT-5.1-Codex-Max is OpenAI’s latest agentic coding model, designed for long-running, high-context software development tasks. It is based on an updated version of the 5.1 reasoning stack and trained on agentic...
GPT-5.1 is the latest frontier-grade model in the GPT-5 series, offering stronger general-purpose reasoning, improved instruction adherence, and a more natural conversational style compared to GPT-5. It uses adaptive reasoning...
Chat-tuned GPT-5.1 for polished assistants, writing, and product conversations
GPT-5.1-Codex is a specialized version of GPT-5.1 optimized for software engineering and coding workflows. It is designed for both interactive development sessions and long, independent execution of complex engineering tasks....
GPT-5.1-Codex-Mini is a smaller and faster version of GPT-5.1-Codex
| Model | Creator | Inputs | Context | Input | Output | Released | Compare |
|---|---|---|---|---|---|---|---|
| Grok 4.20 (Non-Reasoning)xai/grok-4.20-0309-non-reasoning | 1M | $1.25 | $2.5 | 2026-03-09 | |||
| Grok 4.20 (Reasoning)xai/grok-4.20-0309-reasoning | 1M | $1.25 | $2.5 | 2026-03-09 | |||
| OpenAI: GPT-5.4openai/gpt-5.4 | 1.05M | $2.5 | $15 | 2026-03-05 | |||
| OpenAI: GPT-5.4 Proopenai/gpt-5.4-pro | 1.05M | $30 | $180 | 2026-03-05 | |||
| GPT-5.3 Chat (latest)openai/gpt-5.3-chat-latest | 128K | $1.75 | $14 | 2026-03-03 | |||
| Google: Gemini 3.1 Flash Lite Previewgoogle/gemini-3.1-flash-lite-preview | 1.04858M | $0.25 | $1.5 | 2026-03-03 | |||
| Google: Nano Banana 2 (Gemini 3.1 Flash Image Preview)google/gemini-3.1-flash-image-preview | 65.536K | $0.5 | $3 | 2026-02-26 | |||
| Qwen3.5 122B-A10Balibaba/qwen3.5-122b-a10b | 262.144K | $0.4 | $3.2 | 2026-02-23 | |||
| Qwen3.5 9Balibaba/qwen3.5-9b | 262.144K | $0.04 | $0.15 | 2026-02-23 | |||
| Qwen3.5 Flashalibaba/qwen3.5-flash | 1M | $0.029 | $0.287 | 2026-02-23 | |||
| Qwen3.5 27Balibaba/qwen3.5-27b | 262.144K | $0.3 | $2.4 | 2026-02-23 | |||
| Qwen3.5 35B-A3Balibaba/qwen3.5-35b-a3b | 262.144K | $0.25 | $2 | 2026-02-23 | |||
| Google: Gemini 3.1 Pro Previewgoogle/gemini-3.1-pro-preview | 1.04858M | $2 | $12 | 2026-02-19 | |||
| Google: Gemini 3.1 Pro Preview Custom Toolsgoogle/gemini-3.1-pro-preview-customtools | 1.04858M | $2 | $12 | 2026-02-19 | |||
| Trendyol Asure 12Btrendyol/asure-12b | 131.072K | $0.1 | $0.5 | 2026-02-19 | |||
| Claude Sonnet 4.6anthropic/claude-sonnet-4-6 | 1M | $3 | $15 | 2026-02-17 | |||
| Qwen3.5 Plusalibaba/qwen3.5-plus | 1M | $0.4 | $2.4 | 2026-02-16 | |||
| Qwen3.5 397B-A17Balibaba/qwen3.5-397b-a17b | 262.144K | $0.6 | $3.6 | 2026-02-15 | |||
| Seed 2.0 Probytedance-seed/seed-2.0-pro | 256K | $0.475 | $2.375 | 2026-02-14 | |||
| ByteDance Seed: Seed-2.0-Minibytedance-seed/seed-2.0-mini | 262.144K | $0.1 | $0.4 | 2026-02-14 | |||
| ByteDance Seed: Seed-2.0-Litebytedance-seed/seed-2.0-lite | 262.144K | $0.25 | $2 | 2026-02-14 | |||
| ByteDance Seed: Seed-2.0-Codebytedance-seed/seed-2.0-code | 262.144K | $0.5 | $3 | 2026-02-14 | |||
| Llama Nemotron Embed VL 1B v2nvidia/llama-nemotron-embed-vl-1b-v2 | 32.768K | — | — | 2026-02-10 | |||
| Claude Opus 4.6anthropic/claude-opus-4-6 | 1M | $5 | $25 | 2026-02-05 | |||
| GPT-5.3 Codex Sparkopenai/gpt-5.3-codex-spark | 128K | $1.75 | $14 | 2026-02-05 | |||
| OpenAI: GPT-5.3-Codexopenai/gpt-5.3-codex | 400K | $1.75 | $14 | 2026-02-05 | |||
| DeepSeek OCR 2deepseek/deepseek-ocr-2 | 8.192K | $0.03 | $0.03 | 2026-01-27 | |||
| MoonshotAI: Kimi K2.5moonshotai/kimi-k2.5 | 262.144K | $0.45 | $2.25 | 2026-01 | |||
| Seed 1.8bytedance-seed/seed-1-8 | 256K | $0.119 | $1.187 | 2025-12-28 | |||
| Google: Gemini 3 Flash Previewgoogle/gemini-3-flash-preview | 1.04858M | $0.5 | $3 | 2025-12-17 | |||
| OpenAI: GPT-5.2 Proopenai/gpt-5.2-pro | 400K | $21 | $168 | 2025-12-11 | |||
| OpenAI: GPT-5.2openai/gpt-5.2 | 400K | $1.75 | $14 | 2025-12-11 | |||
| GPT-5.2 Chatopenai/gpt-5.2-chat-latest | 128K | $1.75 | $14 | 2025-12-11 | |||
| OpenAI: GPT-5.2-Codexopenai/gpt-5.2-codex | 400K | $1.75 | $14 | 2025-12-11 | |||
| Devstral Small 2mistral/devstral-small-2 | 262.144K | $0.1 | $0.3 | 2025-12-09 | |||
| GLM-4.6Vzhipuai/glm-4.6v | 128K | $0.3 | $0.9 | 2025-12-08 | |||
| GLM-4.6V-Flashzhipuai/glm-4.6v-flash | 128K | $0.3 | $0.9 | 2025-12-08 | |||
| Nova 2 Liteamazon/nova-2-lite | 1M | $0.3 | $2.5 | 2025-12-02 | |||
| Ministral 14Bmistral/ministral-14b | 262.144K | $0.2 | $0.2 | 2025-12-02 | |||
| GPT-Image-1.5openai/gpt-image-1.5 | Not documented | $5 | $32 | 2025-11-25 | |||
| Claude Opus 4.5 (latest)anthropic/claude-opus-4-5 | 200K | $5 | $25 | 2025-11-24 | |||
| Google: Nano Banana Pro (Gemini 3 Pro Image Preview)google/gemini-3-pro-image-preview | 65.536K | $2 | $12 | 2025-11-20 | |||
| Grok 4.1 Fastxai/grok-4.1-fast | 2M | $0.2 | $0.5 | 2025-11-19 | |||
| Grok 4.1 Fast (Reasoning)xai/grok-4.1-fast-reasoning | 2M | $0.2 | $0.5 | 2025-11-19 | |||
| Gemini 3 Pro Previewgoogle/gemini-3-pro-preview | 1.04858M | $0.57 | $3.43 | 2025-11-18 | |||
| OpenAI: GPT-5.1-Codex-Maxopenai/gpt-5.1-codex-max | 400K | $1.25 | $10 | 2025-11-13 | |||
| OpenAI: GPT-5.1openai/gpt-5.1 | 400K | $1.25 | $10 | 2025-11-13 | |||
| GPT-5.1 Chatopenai/gpt-5.1-chat-latest | 128K | $1.25 | $10 | 2025-11-13 | |||
| OpenAI: GPT-5.1-Codexopenai/gpt-5.1-codex | 400K | $1.25 | $10 | 2025-11-13 | |||
| OpenAI: GPT-5.1-Codex-Miniopenai/gpt-5.1-codex-mini | 400K | $0.25 | $2 | 2025-11-13 |