GPT-5.6 Sol is the flagship model in OpenAI's GPT-5.6 series. It is suited for complex reasoning, coding, and agentic workflows, and is particularly strong at command-line and multi-step coding tasks...
Models
Every model in the catalog with source-linked pricing, context limits, provider availability, and published benchmark results.
Kimi K3 is a 2.8T parameter open-weight multimodal reasoning model from Moonshot AI. It is suited for complex coding, knowledge work, and long-horizon agentic workflows, and is particularly strong at...
Claude Fable 5 is a Mythos-class model from Anthropic, built for autonomous knowledge work and coding. It supports text, image, and file inputs with text output, with reasoning support and...
GPT-5.6 Terra is a balanced model in OpenAI's GPT-5.6 series, positioned between the flagship Sol tier and the cost-efficient Luna tier. It is suited for everyday coding, reasoning, and agentic...
Gemini 3.7 Flash is a multimodal model from Google for fast agentic workflows, coding, and complex multi-step reasoning. It is designed for tasks that require responsive performance and reliable multi-step...
GPT-5.6 Luna is a fast, cost-efficient model in OpenAI's GPT-5.6 series. It is suited for high-volume, latency-sensitive tasks such as chat, classification, and lightweight agentic workflows, providing capable reasoning for...
xAI's Grok model for chat, coding, agentic tools, and lower hallucination risk
DeepSeek V4 Flash 0731 is a sparse mixture-of-experts model from DeepSeek, with 13B active parameters out of 284B total. This re-post-trained revision is suited for coding, reasoning, and agent workflows....
Sonnet 5 is Anthropic's most capable Sonnet-class model, with frontier performance across coding, agents, and professional work. It supports adaptive thinking with selectable reasoning effort levels (low, medium, high, max,...
Muse Spark 1.1 is a multimodal reasoning model from Meta, built for agentic tasks. It accepts text, images, video, audio, and PDF documents and returns text, with a 1M-token context...
GPT-5.5 is OpenAI’s frontier model designed for complex professional workloads, building on GPT-5.4 with stronger reasoning, higher reliability, and improved token efficiency on hard tasks. It features a 1M+ token...
Gemini 3.6 Flash is a high-efficiency model from Google for coding, agentic workflows, and web and app development. It is designed to produce polished outputs with fewer unnecessary edits and...
Gemini 3.5 Flash is Google's high-efficiency multimodal model, bringing near-Pro level coding and reasoning at Flash-tier cost and speed. It is highly optimized for coding proficiency and parallel agentic execution...
Top Claude Opus tier for the hardest reasoning, coding, and long-horizon agents
Stronger Opus tier for advanced software work and high-stakes reasoning
Gemini 3.1 Pro Preview is Google’s frontier reasoning model, delivering enhanced software engineering performance, improved agentic reliability, and more efficient token usage across complex workflows. Building on the multimodal foundation...
High-end Claude for difficult coding, planning, and slower expert reasoning
GPT-5.4 is OpenAI’s latest frontier model, unifying the Codex and GPT lines into a single system. It features a 1M+ token context window (922K input, 128K output) with support for...
Claude workhorse for coding agents, careful analysis, and production cost control
DeepSeek V4 Pro is a large-scale Mixture-of-Experts model from DeepSeek with 1.6T total parameters and 49B activated parameters, supporting a 1M-token context window. It is designed for advanced reasoning, coding,...
Kimi K2.6 is Moonshot AI's next-generation multimodal model, designed for long-horizon coding, coding-driven UI/UX generation, and multi-agent orchestration. It handles complex end-to-end coding tasks across Python, Rust, and Go, and...
Step 3.7 Flash is StepFun's latest high-efficiency multimodal Mixture-of-Experts model. It pairs a 196B-parameter language backbone with a vision encoder for native image and video understanding, activating roughly 11B parameters...
NVIDIA Nemotron 3 Ultra is an open frontier-reasoning and orchestration model from NVIDIA, with 55B active parameters out of 550B total (MoE). Built on a hybrid Transformer-Mamba mixture-of-experts architecture, it...
Gemini 3.5 Flash Lite is a high-efficiency model from Google with upgraded agentic capabilities. It is suited for subagents that execute focused tasks within complex, multi-agent workflows.
Muse Glimmer 30B is a dense, open-weight multimodal model from Meta Superintelligence Labs, distilled from Muse Spark and optimized for autonomous agents on consumer hardware. It is suited for long-horizon...
| Model | Creator | Score | Inputs | Context | Input | Output | Released | Compare |
|---|---|---|---|---|---|---|---|---|
| OpenAI: GPT-5.6 Solopenai/gpt-5.6-sol | 88.8 | 1.05M | $2 | $10 | 2026-07-09 | |||
| MoonshotAI: Kimi K3moonshotai/kimi-k3 | 88.3 | 1.04858M | $1.89 | $9.48 | 2026-07-16 | |||
| Anthropic: Claude Fable 5anthropic/claude-fable-5 | 88.0 | 1M | $10 | $50 | 2026-06-09 | |||
| OpenAI: GPT-5.6 Terraopenai/gpt-5.6-terra | 87.4 | 1.05M | $2 | $12 | 2026-07-09 | |||
| Google: Gemini 3.7 Flashgoogle/gemini-3.7-flash | 85.8 | 1.04858M | $0.75 | $3.75 | 2026-08-13 | |||
| OpenAI: GPT-5.6 Lunaopenai/gpt-5.6-luna | 84.7 | 1.05M | $0.2 | $1.2 | 2026-07-09 | |||
| Grok 4.5xai/grok-4.5 | 83.3 | 500K | $2 | $6 | 2026-07-08 | |||
| DeepSeek: DeepSeek V4 Flash 0731deepseek/deepseek-v4-flash-0731 | 82.7 | 1.04858M | $0.065 | $0.18 | 2026-07-31 | |||
| Anthropic: Claude Sonnet 5anthropic/claude-sonnet-5 | 80.4 | 1M | $2 | $10 | 2026-06-30 | |||
| Meta: Muse Spark 1.1meta/muse-spark-1.1 | 80.0 | 1.04858M | $1.25 | $4.25 | 2026-04-08 | |||
| OpenAI: GPT-5.5openai/gpt-5.5 | 78.2 | 1.05M | $5 | $30 | 2026-04-23 | |||
| Google: Gemini 3.6 Flashgoogle/gemini-3.6-flash | 78.0 | 1.04858M | $0.75 | $3.75 | 2026-07-21 | |||
| Google: Gemini 3.5 Flashgoogle/gemini-3.5-flash | 76.2 | 1.04858M | $1.5 | $9 | 2026-05-19 | |||
| Claude Opus 4.8anthropic/claude-opus-4-8 | 74.6 | 1M | $5 | $25 | 2026-05-28 | |||
| Claude Opus 4.7anthropic/claude-opus-4-7 | 71.4 | 1M | $5 | $25 | 2026-04-16 | |||
| Google: Gemini 3.1 Pro Previewgoogle/gemini-3.1-pro-preview | 70.3 | 1.04858M | $2 | $12 | 2026-02-19 | |||
| Claude Opus 4.6anthropic/claude-opus-4-6 | 70.2 | 1M | $5 | $25 | 2026-02-05 | |||
| OpenAI: GPT-5.4openai/gpt-5.4 | 69.8 | 1.05M | $2.5 | $15 | 2026-03-05 | |||
| Claude Sonnet 4.6anthropic/claude-sonnet-4-6 | 67.0 | 1M | $3 | $15 | 2026-02-17 | |||
| DeepSeek: DeepSeek V4 Pro 0423deepseek/deepseek-v4-pro | 64.7 | 1.024M | $0.948 | $1.896 | 2026-04-24 | |||
| MoonshotAI: Kimi K2.6moonshotai/kimi-k2.6 | 64.3 | 262.144K | $0.95 | $4 | 2026-04-21 | |||
| StepFun: Step 3.7 Flashstepfun/step-3.7-flash | 59.6 | 256K | $0.2 | $1.15 | 2026-05-29 | |||
| NVIDIA: Nemotron 3 Ultranvidia/nemotron-3-ultra-550b-a55b | 56.4 | 256K | $0.625 | $3.125 | 2026-06-04 | |||
| Google: Gemini 3.5 Flash Litegoogle/gemini-3.5-flash-lite | 54.0 | 1.04858M | $0.3 | $2.5 | 2026-07-21 | |||
| Meta: Muse Glimmer 30Bmeta/muse-glimmer-30b | 51.7 | 131.072K | $0.3 | $1.1 | 2026-08-10 |