Gemini 3.5 Flash is Google's high-efficiency multimodal model, bringing near-Pro level coding and reasoning at Flash-tier cost and speed. It is highly optimized for coding proficiency and parallel agentic execution...
Models
Every model in the catalog with source-linked pricing, context limits, provider availability, and published benchmark results.
MiniMax-M3 is a multimodal foundation model from MiniMax. It supports text, image, and video inputs with text output, a 1M-token context window, and is suited for long-horizon agentic work, coding,...
Agent-ready GPT for coding and computer-use workflows at a lower cost
Fast Gemini model balancing multimodal reasoning, tool use, and cost
MiniMax-M2.7 is a next-generation large language model designed for autonomous, real-world productivity and continuous improvement. Built to actively participate in its own evolution, M2.7 integrates advanced agentic capabilities through multi-agent...
Qwen 3.6 Plus builds on a hybrid architecture that combines efficient linear attention with sparse mixture-of-experts routing, enabling strong scalability and high-performance inference. Compared to the 3.5 series, it delivers...
GLM-5 is Z.ai’s flagship open-source foundation model engineered for complex systems design and long-horizon agent workflows. Built for expert developers, it delivers production-grade performance on large-scale programming tasks, rivaling leading...
MoonshotAI: Kimi K2.7 Code is a coding-focused model in Moonshot AI's Kimi K2 family, built to complete end-to-end programming tasks reliably over long contexts. It uses a native multimodal mixture-of-experts...
Nex-N2-Pro is an agentic mixture-of-experts model from Nex AGI, with 17B active parameters out of 397B total. Built on the Qwen3.5 architecture, it accepts text and image input and produces...
Gemini 2.5 Pro is Google’s state-of-the-art AI model designed for advanced reasoning, coding, mathematics, and scientific tasks. It employs “thinking” capabilities, enabling it to reason through responses with enhanced accuracy...
Google's proven reasoning model for coding, math, and multimodal analysis
GPT-5 is OpenAI’s most advanced model, offering major improvements in reasoning, code quality, and user experience. It is optimized for complex tasks that require step-by-step reasoning, instruction following, and accuracy...
Original GPT-5 workhorse for reasoning, coding, writing, and tool workflows
Grok 4.20 is a reasoning model from SpaceXAI with industry-leading speed and agentic tool calling capabilities. It combines the lowest hallucination rate on the market with strict prompt adherance, delivering...
Codex GPT for repository edits, code review, and practical software agents
MiniMax-M2.1 is a lightweight, state-of-the-art large language model optimized for coding, agentic workflows, and modern application development. With only 10 billion activated parameters, it delivers a major jump in real-world...
Sharper GPT-5 generation for coding, product work, and tool-assisted tasks
GPT-5.1 is the latest frontier-grade model in the GPT-5 series, offering stronger general-purpose reasoning, improved instruction adherence, and a more natural conversational style compared to GPT-5. It uses adaptive reasoning...
Reliable GPT generation for broad coding, writing, and tool-assisted product work
GPT-5.2 is the latest frontier-grade model in the GPT-5 series, offering stronger agentic and long context perfomance compared to GPT-5.1. It uses adaptive reasoning to allocate computation dynamically, responding quickly...
GLM-5V-Turbo is Z.ai’s first native multimodal agent foundation model, built for vision-based coding and agent-driven tasks. It natively handles image, video, and text inputs, excels at long-horizon planning, complex coding,...
GLM-4.7 is Z.ai’s latest flagship model, featuring upgrades in two key areas: enhanced programming capabilities and more stable multi-step reasoning/execution. It demonstrates significant improvements in executing complex agent tasks while...
GLM-4.5-Air is the lightweight variant of our latest flagship model family, also purpose-built for agent-centric applications. Like GLM-4.5, it adopts the Mixture-of-Experts (MoE) architecture but with a more compact parameter...
o3 is a well-rounded and powerful model across domains. It sets a new standard for math, science, coding, and visual reasoning tasks. It also excels at technical writing and instruction-following....
Deliberate o-series reasoner for hard math, coding, and multi-step analysis
Grok 4.3 is a reasoning model from SpaceXAI. It accepts text and image inputs with text output, and is suited for agentic workflows, instruction-following tasks, and applications requiring high factual...
Grok 4.3 is a reasoning model from SpaceXAI. It accepts text and image inputs with text output, and is suited for agentic workflows, instruction-following tasks, and applications requiring high factual...
May 28th update to the [original DeepSeek R1](/deepseek/deepseek-r1) Performance on par with [OpenAI o1](/openai/o1), but open-sourced and with fully open reasoning tokens. It's 671B parameters in size, with 37B active...
DeepSeek V4 Flash 0731 is a sparse mixture-of-experts model from DeepSeek, with 13B active parameters out of 284B total. This re-post-trained revision is suited for coding, reasoning, and agent workflows....
Upstage's flagship model, specialized for agentic use
Inkling is an open-weight multimodal mixture-of-experts model from Thinking Machines Lab, with 41B active parameters out of 975B total. It is designed for general-purpose reasoning, coding, agentic and tool-use systems,...
Inkling is an open-weight multimodal mixture-of-experts model from Thinking Machines Lab, with 41B active parameters out of 975B total. It is designed for general-purpose reasoning, coding, agentic and tool-use systems,...
Hy3 is a 295B-parameter Mixture-of-Experts model from Tencent (21B active, 192 experts with top-8 routing) built for reasoning, agentic workflows, and real-world production use. It supports a configurable reasoning effort:...
The Qwen3.5 series 397B-A17B native vision-language model is built on a hybrid architecture that integrates a linear attention mechanism with a sparse mixture-of-experts model, achieving higher inference efficiency. It delivers...
MiniMax-M2.5 is a SOTA large language model designed for real-world productivity. Trained in a diverse range of complex real-world digital working environments, M2.5 builds upon the coding expertise of M2.1...
DeepSeek-V3.1 Terminus is an update to [DeepSeek V3.1](/deepseek/deepseek-chat-v3.1) that maintains the model's original capabilities while addressing issues reported by users, including language consistency and agent capabilities, further optimizing the model's...
Claude Sonnet 4.5 is Anthropic’s most advanced Sonnet model to date, optimized for real-world agents and coding workflows. It delivers state-of-the-art performance on coding benchmarks such as SWE-bench Verified, with...
Claude Sonnet 4.5 is Anthropic’s most advanced Sonnet model to date, optimized for real-world agents and coding workflows. It delivers state-of-the-art performance on coding benchmarks such as SWE-bench Verified, with...
Claude Opus 4.1 is an updated version of Anthropic’s flagship model, offering improved performance in coding, reasoning, and agentic tasks. It achieves 74.5% on SWE-bench Verified and shows notable gains...
Claude Opus 4.1 is an updated version of Anthropic’s flagship model, offering improved performance in coding, reasoning, and agentic tasks. It achieves 74.5% on SWE-bench Verified and shows notable gains...
Coding-optimized GPT model for repository edits, reviews, and agentic software work
Compared with GLM-4.5, this generation brings several key improvements: Longer context window: The context window has been expanded from 128K to 200K tokens, enabling the model to handle more complex...
GLM-4.5 is our latest flagship foundation model, purpose-built for agent-based applications. It leverages a Mixture-of-Experts (MoE) architecture and supports a context length of up to 128k tokens. GLM-4.5 delivers significantly...
Claude Sonnet 4 significantly enhances the capabilities of its predecessor, Sonnet 3.7, excelling in both coding and reasoning tasks with improved precision and controllability. Achieving state-of-the-art performance on SWE-bench (72.7%),...
DeepSeek-V3.2-Exp is an experimental large language model released by DeepSeek as an intermediate step between V3.1 and future architectures. It introduces DeepSeek Sparse Attention (DSA), a fine-grained sparse attention mechanism...
Claude Opus 4 is benchmarked as the world’s best coding model, at time of release, bringing sustained performance on complex, long-running tasks and agent workflows. It sets new benchmarks in...
Mistral Medium 3.1 is an updated version of Mistral Medium 3, which is a high-performance enterprise-grade language model designed to deliver frontier-level capabilities at significantly reduced operational cost. It balances...
Mistral Medium 3.1 is an updated version of Mistral Medium 3, which is a high-performance enterprise-grade language model designed to deliver frontier-level capabilities at significantly reduced operational cost. It balances...
NVIDIA Nemotron 3 Ultra is an open frontier-reasoning and orchestration model from NVIDIA, with 55B active parameters out of 550B total (MoE). Built on a hybrid Transformer-Mamba mixture-of-experts architecture, it...
MiniMax-M2 is a compact, high-efficiency large language model optimized for end-to-end coding and agentic workflows. With 10 billion activated parameters (230 billion total), it delivers near-frontier intelligence across general reasoning,...
| Model | Creator | Score | Inputs | Context | Input | Output | Released | Compare |
|---|---|---|---|---|---|---|---|---|
| Google: Gemini 3.5 Flash (batch)google/gemini-3.5-flash:batch | 1250.0 | 1.04858M | $0.75 | $4.5 | — | |||
| MiniMax: MiniMax M3 (batch)minimax/minimax-m3:batch | 1250.0 | 524.288K | $0.3 | $1.2 | — | |||
| GPT-5.4openai/gpt-5.4 | 1250.0 | 1.05M | $2.5 | $15 | 2026-03-05 | |||
| Gemini 3.5 Flashgoogle/gemini-3.5-flash | 1250.0 | 1.04858M | $1.5 | $9 | 2026-05-19 | |||
| MiniMax: MiniMax M2.7minimax/minimax-m2.7 | 1249.0 | 204.8K | $0.3 | $1.2 | — | |||
| Qwen: Qwen3.6 Plusqwen/qwen3.6-plus | 1245.0 | 1M | $0.325 | $1.95 | — | |||
| Z.ai: GLM 5z-ai/glm-5 | 1244.0 | 198K | $0.6 | $1.92 | — | |||
| MoonshotAI: Kimi K2.7 Code (batch)moonshotai/kimi-k2.7-code:batch | 1243.0 | 262.144K | $0.95 | $4 | — | |||
| Nex AGI: Nex-N2-Pronex-agi/nex-n2-pro | 1242.0 | 262.144K | $0.25 | $1 | — | |||
| Google: Gemini 2.5 Pro (batch)google/gemini-2.5-pro:batch | 1237.0 | 1.04858M | $0.625 | $5 | — | |||
| Gemini 2.5 Progoogle/gemini-2.5-pro | 1237.0 | 1.04858M | $1.25 | $10 | 2025-06-17 | |||
| OpenAI: GPT-5 (batch)openai/gpt-5:batch | 1237.0 | 400K | $0.625 | $5 | — | |||
| GPT-5openai/gpt-5 | 1237.0 | 400K | $1.25 | $10 | 2025-08-07 | |||
| SpaceXAI: Grok 4.20x-ai/grok-4.20 | 1233.0 | 2M | $1.25 | $2.5 | — | |||
| GPT-5.1 Codexopenai/gpt-5.1-codex | 1229.0 | 400K | $1.07 | $8.5 | 2025-11-13 | |||
| MiniMax: MiniMax M2.1minimax/minimax-m2.1 | 1222.0 | 204.8K | $0.3 | $1.2 | — | |||
| GPT-5.1openai/gpt-5.1 | 1220.0 | 400K | $1.25 | $10 | 2025-11-13 | |||
| OpenAI: GPT-5.1 (batch)openai/gpt-5.1:batch | 1220.0 | 400K | $0.625 | $5 | — | |||
| GPT-5.2openai/gpt-5.2 | 1217.0 | 400K | $1.75 | $14 | 2025-12-11 | |||
| OpenAI: GPT-5.2 (batch)openai/gpt-5.2:batch | 1217.0 | 400K | $0.875 | $7 | — | |||
| Z.ai: GLM 5V Turboz-ai/glm-5v-turbo | 1216.0 | 202.752K | $1.2 | $4 | — | |||
| Z.ai: GLM 4.7z-ai/glm-4.7 | 1211.0 | 202.752K | $0.4 | $1.75 | — | |||
| Z.ai: GLM 4.5 Airz-ai/glm-4.5-air | 1202.0 | 131.072K | $0.13 | $0.85 | — | |||
| OpenAI: o3 (batch)openai/o3:batch | 1200.0 | 200K | $1 | $4 | — | |||
| o3openai/o3 | 1200.0 | 200K | $2 | $8 | 2025-04-16 | |||
| SpaceXAI: Grok 4.3x-ai/grok-4.3 | 1198.0 | 1M | $1.25 | $2.5 | — | |||
| SpaceXAI: Grok 4.3 (batch)x-ai/grok-4.3:batch | 1198.0 | 1M | $1 | $2 | — | |||
| DeepSeek: R1 0528deepseek/deepseek-r1-0528 | 1197.0 | 163.84K | $0.5 | $2.15 | — | |||
| DeepSeek: DeepSeek V4 Flash 0731 (batch)deepseek/deepseek-v4-flash-0731:batch | 1197.0 | 1.04858M | $0.11 | $0.33 | — | |||
| Solar Pro 4upstage/solar-pro4 | 1196.0 | 524.288K | $0.3 | $1.2 | 2026-08-06 | |||
| Thinking Machines: Inkling (batch)thinkingmachines/inkling:batch | 1193.0 | 524.288K | $1 | $4.05 | — | |||
| Thinking Machines: Inkling (free)thinkingmachines/inkling:free | 1193.0 | 1.04858M | Free | Free | — | |||
| Tencent: Hy3 (free)tencent/hy3:free | 1192.0 | 262.144K | Free | Free | — | |||
| Qwen: Qwen3.5 397B A17Bqwen/qwen3.5-397b-a17b | 1191.0 | 262.144K | $0.55 | $3.5 | — | |||
| MiniMax: MiniMax M2.5minimax/minimax-m2.5 | 1188.0 | 204.8K | $0.3 | $1.2 | — | |||
| DeepSeek: DeepSeek V3.1 Terminusdeepseek/deepseek-v3.1-terminus | 1183.0 | 131.072K | $0.27 | $1 | — | |||
| Anthropic: Claude Sonnet 4.5 (batch)anthropic/claude-sonnet-4.5:batch | 1183.0 | 1M | $1.5 | $7.5 | — | |||
| Anthropic: Claude Sonnet 4.5anthropic/claude-sonnet-4.5 | 1183.0 | 1M | $3 | $15 | — | |||
| Anthropic: Claude Opus 4.1anthropic/claude-opus-4.1 | 1182.0 | 200K | $15 | $75 | — | |||
| Anthropic: Claude Opus 4.1 (batch)anthropic/claude-opus-4.1:batch | 1182.0 | 200K | $7.5 | $37.5 | — | |||
| GPT-5.3 Codexopenai/gpt-5.3-codex | 1180.0 | 400K | $1.75 | $14 | 2026-02-05 | |||
| Z.ai: GLM 4.6z-ai/glm-4.6 | 1177.0 | 198K | $0.43 | $1.75 | — | |||
| Z.ai: GLM 4.5z-ai/glm-4.5 | 1175.0 | 131.072K | $0.6 | $2.2 | — | |||
| Anthropic: Claude Sonnet 4anthropic/claude-sonnet-4 | 1173.0 | 200K | $3 | $15 | — | |||
| DeepSeek: DeepSeek V3.2 Expdeepseek/deepseek-v3.2-exp | 1168.0 | 163.84K | $0.27 | $0.41 | — | |||
| Anthropic: Claude Opus 4anthropic/claude-opus-4 | 1161.0 | 200K | $15 | $75 | — | |||
| Mistral: Mistral Medium 3.1mistralai/mistral-medium-3.1 | 1160.0 | 131.072K | $0.4 | $2 | — | |||
| Mistral: Mistral Medium 3.1 (batch)mistralai/mistral-medium-3.1:batch | 1160.0 | 131.072K | $0.2 | $1 | — | |||
| NVIDIA: Nemotron 3 Ultra (batch)nvidia/nemotron-3-ultra-550b-a55b:batch | 1158.0 | 512.288K | $0.6 | $3.6 | — | |||
| MiniMax: MiniMax M2minimax/minimax-m2 | 1156.0 | 204.8K | $0.255 | $1.02 | — |