Claude Opus 5 is Anthropic’s flagship model for demanding reasoning, coding, and long-horizon agentic work. It is particularly strong at end-to-end software tasks, code review and bug finding, visual analysis...
Models
Every model in the catalog with source-linked pricing, context limits, provider availability, and published benchmark results.
Claude Fable 5 is a Mythos-class model from Anthropic, built for autonomous knowledge work and coding. It supports text, image, and file inputs with text output, with reasoning support and...
Muse Spark 1.2 is a reasoning model from Meta, designed for complex agentic tasks. It accepts text, images, video, audio, and PDF documents, returns text, and offers a 1M-token context...
Muse Spark 1.1 is a multimodal reasoning model from Meta, built for agentic tasks. It accepts text, images, video, audio, and PDF documents and returns text, with a 1M-token context...
Gemini 3.6 Flash is a high-efficiency model from Google for coding, agentic workflows, and web and app development. It is designed to produce polished outputs with fewer unnecessary edits and...
Gemini 3.1 Pro Preview is Google’s frontier reasoning model, delivering enhanced software engineering performance, improved agentic reliability, and more efficient token usage across complex workflows. Building on the multimodal foundation...
GPT-5.5 is OpenAI’s frontier model designed for complex professional workloads, building on GPT-5.4 with stronger reasoning, higher reliability, and improved token efficiency on hard tasks. It features a 1M+ token...
Gemini 3.5 Flash is Google's high-efficiency multimodal model, bringing near-Pro level coding and reasoning at Flash-tier cost and speed. It is highly optimized for coding proficiency and parallel agentic execution...
Gemini 3.7 Flash is a multimodal model from Google for fast agentic workflows, coding, and complex multi-step reasoning. It is designed for tasks that require responsive performance and reliable multi-step...
Sonnet 5 is Anthropic's most capable Sonnet-class model, with frontier performance across coding, agents, and professional work. It supports adaptive thinking with selectable reasoning effort levels (low, medium, high, max,...
MoonshotAI: Kimi K2.7 Code is a coding-focused model in Moonshot AI's Kimi K2 family, built to complete end-to-end programming tasks reliably over long contexts. It uses a native multimodal mixture-of-experts...
GPT-5.4 is OpenAI’s latest frontier model, unifying the Codex and GPT lines into a single system. It features a 1M+ token context window (922K input, 128K output) with support for...
Kimi K2.5 is Moonshot AI's native multimodal model, delivering state-of-the-art visual coding capability and a self-directed agent swarm paradigm. Built on Kimi K2 with continued pretraining over approximately 15T mixed...
Gemini 3.1 Flash Lite Preview is Google's high-efficiency model optimized for high-volume use cases. It outperforms Gemini 2.5 Flash Lite on overall quality and approaches Gemini 2.5 Flash performance across...
Kimi K2.6 is Moonshot AI's next-generation multimodal model, designed for long-horizon coding, coding-driven UI/UX generation, and multi-agent orchestration. It handles complex end-to-end coding tasks across Python, Rust, and Go, and...
GPT-5.2 is the latest frontier-grade model in the GPT-5 series, offering stronger agentic and long context perfomance compared to GPT-5.1. It uses adaptive reasoning to allocate computation dynamically, responding quickly...
GPT-5.3-Codex is OpenAI’s most advanced agentic coding model, combining the frontier software engineering performance of GPT-5.2-Codex with the broader reasoning and professional knowledge capabilities of GPT-5.2. It achieves state-of-the-art results...
MiMo-V2.5-Pro is Xiaomi’s flagship model, delivering strong performance in general agentic capabilities, complex software engineering, and long-horizon tasks, with top rankings on benchmarks such as ClawEval, GDPVal, and SWE-bench Pro....
Step 3.7 Flash is StepFun's latest high-efficiency multimodal Mixture-of-Experts model. It pairs a 196B-parameter language backbone with a vision encoder for native image and video understanding, activating roughly 11B parameters...
DeepSeek V4 Pro is a large-scale Mixture-of-Experts model from DeepSeek with 1.6T total parameters and 49B activated parameters, supporting a 1M-token context window. It is designed for advanced reasoning, coding,...
MiMo-V2.5 is a native omnimodal model by Xiaomi. It delivers Pro-level agentic performance at roughly half the inference cost, while surpassing MiMo-V2-Omni in multimodal perception across image and video understanding...
GPT-5 is OpenAI’s most advanced model, offering major improvements in reasoning, code quality, and user experience. It is optimized for complex tasks that require step-by-step reasoning, instruction following, and accuracy...
GPT-5 Mini is a compact version of GPT-5, designed to handle lighter-weight reasoning tasks. It provides the same instruction-following and safety-tuning benefits as GPT-5, but with reduced latency and cost....
DeepSeek V4 Flash is an efficiency-optimized Mixture-of-Experts model from DeepSeek with 284B total parameters and 13B activated parameters, supporting a 1M-token context window. It is designed for fast inference and...
GPT-5.1-Codex-Mini is a smaller and faster version of GPT-5.1-Codex
Inkling is an open-weight multimodal mixture-of-experts model from Thinking Machines Lab, with 41B active parameters out of 975B total. It is designed for general-purpose reasoning, coding, agentic and tool-use systems,...
DeepSeek V4 Flash 0731 is a sparse mixture-of-experts model from DeepSeek, with 13B active parameters out of 284B total. This re-post-trained revision is suited for coding, reasoning, and agent workflows....
DeepSeek-V3.2 is a large language model designed to harmonize high computational efficiency with strong reasoning and agentic tool-use performance. It introduces DeepSeek Sparse Attention (DSA), a fine-grained sparse attention mechanism...
NVIDIA Nemotron 3 Ultra is an open frontier-reasoning and orchestration model from NVIDIA, with 55B active parameters out of 550B total (MoE). Built on a hybrid Transformer-Mamba mixture-of-experts architecture, it...
Trinity Large Thinking is a powerful open source reasoning model from the team at Arcee AI. It shows strong performance in PinchBench, agentic workloads, and reasoning tasks. Launch video: https://youtu.be/Gc82AXLa0Rg?si=4RLn6WBz33qT--B7...
| Model | Creator | Score | Inputs | Context | Input | Output | Released | Compare |
|---|---|---|---|---|---|---|---|---|
| Claude Opus 5anthropic/claude-opus-5 | 1386.0 | 1M | $5 | $25 | 2026-07-24 | |||
| Anthropic: Claude Fable 5anthropic/claude-fable-5 | 1354.0 | 1M | $10 | $50 | 2026-06-09 | |||
| Meta: Muse Spark 1.2meta/muse-spark-1.2 | 1333.0 | 1.04858M | $1.25 | $4.25 | 2026-08-05 | |||
| Meta: Muse Spark 1.1meta/muse-spark-1.1 | 1305.0 | 1.04858M | $1.25 | $4.25 | 2026-04-08 | |||
| Google: Gemini 3.6 Flashgoogle/gemini-3.6-flash | 1291.0 | 1.04858M | $0.75 | $3.75 | 2026-07-21 | |||
| Google: Gemini 3.1 Pro Previewgoogle/gemini-3.1-pro-preview | 1291.0 | 1.04858M | $2 | $12 | 2026-02-19 | |||
| OpenAI: GPT-5.5openai/gpt-5.5 | 1277.0 | 1.05M | $5 | $30 | 2026-04-23 | |||
| Google: Gemini 3.5 Flashgoogle/gemini-3.5-flash | 1276.0 | 1.04858M | $1.5 | $9 | 2026-05-19 | |||
| Google: Gemini 3.7 Flashgoogle/gemini-3.7-flash | 1252.0 | 1.04858M | $0.75 | $3.75 | 2026-08-13 | |||
| Anthropic: Claude Sonnet 5anthropic/claude-sonnet-5 | 1229.0 | 1M | $2 | $10 | 2026-06-30 | |||
| MoonshotAI: Kimi K2.7 Codemoonshotai/kimi-k2.7-code | 1226.0 | 262.144K | $0.71 | $3.5 | 2026-06-12 | |||
| OpenAI: GPT-5.4openai/gpt-5.4 | 1219.0 | 1.05M | $2.5 | $15 | 2026-03-05 | |||
| MoonshotAI: Kimi K2.5moonshotai/kimi-k2.5 | 1187.0 | 262.144K | $0.45 | $2.25 | 2026-01 | |||
| Google: Gemini 3.1 Flash Lite Previewgoogle/gemini-3.1-flash-lite-preview | 1187.0 | 1.04858M | $0.25 | $1.5 | 2026-03-03 | |||
| MoonshotAI: Kimi K2.6moonshotai/kimi-k2.6 | 1186.0 | 262.144K | $0.95 | $4 | 2026-04-21 | |||
| OpenAI: GPT-5.2openai/gpt-5.2 | 1172.0 | 400K | $1.75 | $14 | 2025-12-11 | |||
| OpenAI: GPT-5.3-Codexopenai/gpt-5.3-codex | 1171.0 | 400K | $1.75 | $14 | 2026-02-05 | |||
| Xiaomi: MiMo-V2.5-Proxiaomi/mimo-v2.5-pro | 1170.0 | 1.04858M | $0.435 | $0.87 | 2026-04-22 | |||
| StepFun: Step 3.7 Flashstepfun/step-3.7-flash | 1170.0 | 256K | $0.2 | $1.15 | 2026-05-29 | |||
| DeepSeek: DeepSeek V4 Pro 0423deepseek/deepseek-v4-pro | 1170.0 | 1.024M | $0.86 | $1.72 | 2026-04-24 | |||
| Xiaomi: MiMo-V2.5xiaomi/mimo-v2.5 | 1164.0 | 1.04858M | $0.14 | $0.28 | 2026-04-22 | |||
| OpenAI: GPT-5openai/gpt-5 | 1161.0 | 400K | $1.25 | $10 | 2025-08-07 | |||
| OpenAI: GPT-5 Miniopenai/gpt-5-mini | 1142.0 | 400K | $0.25 | $2 | 2025-08-07 | |||
| DeepSeek: DeepSeek V4 Flash 0423deepseek/deepseek-v4-flash | 1131.0 | 1.024M | $0.086 | $0.172 | 2026-04-24 | |||
| OpenAI: GPT-5.1-Codex-Miniopenai/gpt-5.1-codex-mini | 1125.0 | 400K | $0.25 | $2 | 2025-11-13 | |||
| Thinking Machines: Inklingthinkingmachines/inkling | 1119.0 | 1.04858M | $1 | $4.05 | 2026-07-15 | |||
| DeepSeek: DeepSeek V4 Flash 0731deepseek/deepseek-v4-flash-0731 | 1116.0 | 1.04858M | $0.065 | $0.18 | 2026-07-31 | |||
| DeepSeek: DeepSeek V3.2deepseek/deepseek-v3.2 | 1102.0 | 163.84K | $0.269 | $0.4 | 2025-12-01 | |||
| NVIDIA: Nemotron 3 Ultranvidia/nemotron-3-ultra-550b-a55b | 1101.0 | 256K | $0.625 | $3.125 | 2026-06-04 | |||
| Arcee AI: Trinity Large Thinkingarcee-ai/trinity-large-thinking | 1061.0 | 262.144K | $0.25 | $0.8 | 2026-04-01 |