GPT-5.5 is OpenAI’s frontier model designed for complex professional workloads, building on GPT-5.4 with stronger reasoning, higher reliability, and improved token efficiency on hard tasks. It features a 1M+ token...
Models
Every model in the catalog with source-linked pricing, context limits, provider availability, and published benchmark results.
Claude Sonnet 4 significantly enhances the capabilities of its predecessor, Sonnet 3.7, excelling in both coding and reasoning tasks with improved precision and controllability. Achieving state-of-the-art performance on SWE-bench (72.7%),...
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
Gemini 3.5 Flash is Google's high-efficiency multimodal model, bringing near-Pro level coding and reasoning at Flash-tier cost and speed. It is highly optimized for coding proficiency and parallel agentic execution...
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
Qwen3.8 Flash is a multimodal reasoning model from Alibaba. It is suited for coding assistance, agentic workflows, visual understanding, document and codebase analysis, desktop interaction, chart analysis, and long-video analysis.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
Qwen3-14B is a dense 14.8B parameter causal language model from the Qwen3 series, designed for both complex reasoning and efficient dialogue. It supports seamless switching between a "thinking" mode for...
No provider description is available for this model yet.
OpenAI o4-mini is a compact reasoning model in the o-series, optimized for fast, cost-efficient performance while retaining strong multimodal and agentic capabilities. It supports tool use and demonstrates competitive reasoning...
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
Fugu Ultra v2 is the higher-performance model in Sakana AI's Fugu family. Rather than a single monolithic model, Fugu is a learned multi-agent orchestration system: a language model trained to...
For tasks that demand low latency, GPT‑4.1 nano is the fastest and cheapest model in the GPT-4.1 series. It delivers exceptional performance at a small size with its 1 million...
No provider description is available for this model yet.
No provider description is available for this model yet.
Relace Apply 3 is a specialized code-patching LLM that merges AI-suggested edits straight into your source files. It can apply updates from GPT-4o, Claude, and others into your files at...
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
MiniMax-M2.7 is a next-generation large language model designed for autonomous, real-world productivity and continuous improvement. Built to actively participate in its own evolution, M2.7 integrates advanced agentic capabilities through multi-agent...
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
| Model | Creator | Inputs | Context | Input | Output | Released | Compare |
|---|---|---|---|---|---|---|---|
| OpenAI: GPT-5.5 (batch)openai/gpt-5.5:batch | 1.05M | $2.5 | $15 | — | |||
| Anthropic: Claude Sonnet 4anthropic/claude-sonnet-4 | 200K | $3 | $15 | — | |||
| openai.gpt-5oci/openai.gpt-5 | 272K | $1.25 | $10 | — | |||
| qwen3-vl-235b-a22b-thinkingqwencloud/qwen3-vl-235b-a22b-thinking | 131.072K | $0.4 | $4 | — | |||
| qwen3-vl-32b-instructqwencloud/qwen3-vl-32b-instruct | 131.072K | $0.16 | $0.64 | — | |||
| qwen3-vl-32b-thinkingqwencloud/qwen3-vl-32b-thinking | 131.072K | $0.16 | $2.87 | — | |||
| qwen3-vl-plusqwencloud/qwen3-vl-plus | 260.096K | — | — | — | |||
| Google: Gemini 3.5 Flash (batch)google/gemini-3.5-flash:batch | 1.04858M | $0.75 | $4.5 | — | |||
| qwen3.5-plusqwencloud/qwen3.5-plus | 991.808K | — | — | — | |||
| qwen3.7-maxqwencloud/qwen3.7-max | 991.808K | $2.5 | $7.5 | — | |||
| qwen3.7-plusqwencloud/qwen3.7-plus | 991.808K | — | — | — | |||
| xai.grok-code-fast-1oci/xai.grok-code-fast-1 | 131.072K | $5 | $25 | — | |||
| qwen-turbo-latestdashscope/qwen-turbo-latest | 1M | $0.05 | $0.2 | — | |||
| Qwen: Qwen3.8 Flashqwen/qwen3.8-flash | 1M | $0.15 | $0.47 | — | |||
| qwen3.8-maxqwencloud/qwen3.8-max | 991.808K | $2 | $6 | — | |||
| qwq-plusqwencloud/qwq-plus | 98.304K | $0.8 | $2.4 | — | |||
| gpt-4o-2024-08-06azure/gpt-4o-2024-08-06 | 128K | $2.5 | $10 | — | |||
| deepseek-v4-flashqwen_ai_platform/deepseek-v4-flash | 1M | $0.2 | $0.4 | — | |||
| Qwen: Qwen3 14Bqwen/qwen3-14b | 40.96K | $0.12 | $0.24 | — | |||
| deepseek-v4-flash-0731qwen_ai_platform/deepseek-v4-flash-0731 | 1M | $0.2 | $0.4 | — | |||
| OpenAI: o4 Mini (batch)openai/o4-mini:batch | 200K | $0.55 | $2.2 | — | |||
| glm-5.1qwen_ai_platform/glm-5.1 | 202.745K | $1.4 | $4.4 | — | |||
| glm-5.2qwen_ai_platform/glm-5.2 | 1.04858M | $1.4 | $4.4 | — | |||
| kimi-k2.7-codeqwen_ai_platform/kimi-k2.7-code | 229.376K | $0.95 | $4 | — | |||
| qwen-coderqwen_ai_platform/qwen-coder | 1M | $0.3 | $1.5 | — | |||
| qwen-flashqwen_ai_platform/qwen-flash | 997.952K | — | — | — | |||
| qwen-flash-2025-07-28qwen_ai_platform/qwen-flash-2025-07-28 | 997.952K | — | — | — | |||
| google/gemma-4-31B-itdeepinfra/google/gemma-4-31b-it | 262.144K | $0.13 | $0.38 | — | |||
| xai.grok-4.20-multi-agentoci/xai.grok-4.20-multi-agent | 131.072K | $3 | $15 | — | |||
| qwen-plusqwen_ai_platform/qwen-plus | 129.024K | $0.4 | $1.2 | — | |||
| qwen-plus-2025-01-25qwen_ai_platform/qwen-plus-2025-01-25 | 129.024K | $0.4 | $1.2 | — | |||
| Sakana: Fugu Ultra v2sakana/fugu-ultra-v2 | 1M | $5 | $30 | — | |||
| OpenAI: GPT-4.1 Nano (batch)openai/gpt-4.1-nano:batch | 1.04758M | $0.05 | $0.2 | — | |||
| qwen-plus-2025-04-28qwen_ai_platform/qwen-plus-2025-04-28 | 129.024K | $0.4 | $1.2 | — | |||
| qwen-plus-2025-07-14qwen_ai_platform/qwen-plus-2025-07-14 | 129.024K | $0.4 | $1.2 | — | |||
| Relace: Relace Apply 3relace/relace-apply-3 | 256K | $0.85 | $1.25 | — | |||
| qwen-plus-2025-07-28qwen_ai_platform/qwen-plus-2025-07-28 | 997.952K | — | — | — | |||
| qwen-plus-2025-09-11qwen_ai_platform/qwen-plus-2025-09-11 | 997.952K | — | — | — | |||
| qwen-plus-latestqwen_ai_platform/qwen-plus-latest | 997.952K | — | — | — | |||
| qwen-turboqwen_ai_platform/qwen-turbo | 129.024K | $0.05 | $0.2 | — | |||
| qwen-turbo-2024-11-01qwen_ai_platform/qwen-turbo-2024-11-01 | 1M | $0.05 | $0.2 | — | |||
| qwen-turbo-2025-04-28qwen_ai_platform/qwen-turbo-2025-04-28 | 1M | $0.05 | $0.2 | — | |||
| zai-org/GLM-4.7deepinfra/zai-org/glm-4.7 | 202.752K | $0.4 | $1.75 | — | |||
| qwen-turbo-latestqwen_ai_platform/qwen-turbo-latest | 1M | $0.05 | $0.2 | — | |||
| qwen3-30b-a3bqwen_ai_platform/qwen3-30b-a3b | 129.024K | — | — | — | |||
| MiniMax: MiniMax M2.7minimax/minimax-m2.7 | 204.8K | $0.3 | $1.2 | — | |||
| qwen3-coder-flashqwen_ai_platform/qwen3-coder-flash | 997.952K | — | — | — | |||
| qwen3-coder-flash-2025-07-28qwen_ai_platform/qwen3-coder-flash-2025-07-28 | 997.952K | — | — | — | |||
| qwen3-coder-plusqwen_ai_platform/qwen3-coder-plus | 997.952K | — | — | — | |||
| MiniMaxAI/MiniMax-M2.7-Turbodeepinfra/minimaxai/minimax-m2.7-turbo | 196.608K | $0.38 | $1.7 | — |