Claude Opus 5 is Anthropic’s flagship model for demanding reasoning, coding, and long-horizon agentic work. It is particularly strong at end-to-end software tasks, code review and bug finding, visual analysis...
Models
Every model in the catalog with source-linked pricing, context limits, provider availability, and published benchmark results.
GPT-6 Astra is OpenAI's flagship model for demanding end-to-end work. It is suited for advanced analysis, software engineering, deep research, scientific work, and document creation, with particular strengths in long-horizon...
Claude Fable 5 is a Mythos-class model from Anthropic, built for autonomous knowledge work and coding. It supports text, image, and file inputs with text output, with reasoning support and...
GPT-5.6 Sol is the flagship model in OpenAI's GPT-5.6 series. It is suited for complex reasoning, coding, and agentic workflows, and is particularly strong at command-line and multi-step coding tasks...
Sonnet 5 is Anthropic's most capable Sonnet-class model, with frontier performance across coding, agents, and professional work. It supports adaptive thinking with selectable reasoning effort levels (low, medium, high, max,...
GPT-5.4 is OpenAI’s latest frontier model, unifying the Codex and GPT lines into a single system. It features a 1M+ token context window (922K input, 128K output) with support for...
GPT-5.6 Terra is a balanced model in OpenAI's GPT-5.6 series, positioned between the flagship Sol tier and the cost-efficient Luna tier. It is suited for everyday coding, reasoning, and agentic...
GPT-5.6 Luna is a fast, cost-efficient model in OpenAI's GPT-5.6 series. It is suited for high-volume, latency-sensitive tasks such as chat, classification, and lightweight agentic workflows, providing capable reasoning for...
GPT-5.5 is OpenAI’s frontier model designed for complex professional workloads, building on GPT-5.4 with stronger reasoning, higher reliability, and improved token efficiency on hard tasks. It features a 1M+ token...
Kimi K2.5 is Moonshot AI's native multimodal model, delivering state-of-the-art visual coding capability and a self-directed agent swarm paradigm. Built on Kimi K2 with continued pretraining over approximately 15T mixed...
DeepSeek-V3.2 is a large language model designed to harmonize high computational efficiency with strong reasoning and agentic tool-use performance. It introduces DeepSeek Sparse Attention (DSA), a fine-grained sparse attention mechanism...
Gemma 4 26B A4B IT is an instruction-tuned Mixture-of-Experts (MoE) model from Google DeepMind. Despite 25.2B total parameters, only 3.8B activate per token during inference — delivering near-31B quality at...
Gemma 4 31B Instruct is Google DeepMind's 30.7B dense multimodal model supporting text and image input with text output. Features a 256K token context window, configurable thinking/reasoning mode, native function...
Kimi K2 Thinking is Moonshot AI’s most advanced open reasoning model to date, extending the K2 series into agentic, long-horizon reasoning. Built on the trillion-parameter Mixture-of-Experts (MoE) architecture introduced in...
DeepSeek R1 is here: Performance on par with [OpenAI o1](/openai/o1), but open-sourced and with fully open reasoning tokens. It's 671B parameters in size, with 37B active in an inference pass....
Gemma 3 introduces multimodality, supporting vision-language input and text outputs. It handles context windows up to 128k tokens, understands over 140 languages, and offers improved math, reasoning, and chat capabilities,...
Gemma 3 introduces multimodality, supporting vision-language input and text outputs. It handles context windows up to 128k tokens, understands over 140 languages, and offers improved math, reasoning, and chat capabilities,...
| Model | Creator | Score | Inputs | Context | Input | Output | Released | Compare |
|---|---|---|---|---|---|---|---|---|
| Claude Opus 5anthropic/claude-opus-5 | 56.2 | 1M | $5 | $25 | 2026-07-24 | |||
| OpenAI: GPT-6 Astraopenai/gpt-6-astra | 51.5 | 1.05M | $10 | $50 | 2026-09-04 | |||
| Anthropic: Claude Fable 5anthropic/claude-fable-5 | 51.0 | 1M | $10 | $50 | 2026-06-09 | |||
| OpenAI: GPT-5.6 Solopenai/gpt-5.6-sol | 50.5 | 1.05M | $2 | $10 | 2026-07-09 | |||
| Anthropic: Claude Sonnet 5anthropic/claude-sonnet-5 | 44.3 | 1M | $2 | $10 | 2026-06-30 | |||
| OpenAI: GPT-5.4openai/gpt-5.4 | 44.2 | 1.05M | $2.5 | $15 | 2026-03-05 | |||
| OpenAI: GPT-5.6 Terraopenai/gpt-5.6-terra | 43.7 | 1.05M | $2 | $12 | 2026-07-09 | |||
| OpenAI: GPT-5.6 Lunaopenai/gpt-5.6-luna | 42.7 | 1.05M | $0.2 | $1.2 | 2026-07-09 | |||
| OpenAI: GPT-5.5openai/gpt-5.5 | 37.3 | 1.05M | $5 | $30 | 2026-04-23 | |||
| MoonshotAI: Kimi K2.5moonshotai/kimi-k2.5 | 21.7 | 262.144K | $0.45 | $2.25 | 2026-01 | |||
| DeepSeek: DeepSeek V3.2deepseek/deepseek-v3.2 | 18.3 | 163.84K | $0.269 | $0.4 | 2025-12-01 | |||
| Google: Gemma 4 26B A4B google/gemma-4-26b-a4b-it | 11.0 | 131.072K | $0.042 | $0.22 | 2026-04-02 | |||
| Google: Gemma 4 31Bgoogle/gemma-4-31b-it | 6.7 | 262.144K | $0.09 | $0.34 | 2026-04-02 | |||
| MoonshotAI: Kimi K2 Thinkingmoonshotai/kimi-k2-thinking | 1.8 | 262.144K | $0.6 | $2.5 | 2025-11-06 | |||
| DeepSeek: R1deepseek/deepseek-r1 | 1.1 | 64K | $0.7 | $2.5 | 2025-01-20 | |||
| Google: Gemma 3 12Bgoogle/gemma-3-12b-it | 0.1 | 131.072K | $0.05 | $0.15 | 2025-03-12 | |||
| Google: Gemma 3 27Bgoogle/gemma-3-27b-it | 0.1 | 131.072K | $0.08 | $0.45 | 2025-03-12 |