Thinking Kimi model for slower research passes, planning, and hard technical questions
Models
Every model in the catalog with source-linked pricing, context limits, provider availability, and published benchmark results.
Flagship Claude model for deep reasoning, coding, and long-horizon agents
Efficient open MiniMax model built for coding agents and tool-heavy workflows
Higher-accuracy GPT-5 tier for tough analysis, coding reviews, and planning
Late GLM-4 workhorse for coding agents, reasoning, and structured tasks
Qwen vision-language instruct model for visual reasoning, documents, and agent tasks
Flagship Qwen3 model for coding agents, complex reasoning, and tool use
Qwen vision-language thinking model for visual reasoning, documents, and agent tasks
Gemma 3 27B tuned by AI Singapore for Southeast Asian languages and instruction following
Efficient Qwen thinking model for local reasoning, math, and coding agents
Qwen instruction model for multilingual chat, reasoning, and tool use
Nano Banana image model for fast generation, edits, and character-consistent assets
Compact Nemotron model for efficient reasoning and deployable AI agents
Original GPT-5 workhorse for reasoning, coding, writing, and tool workflows
Small GPT-5 for responsive agents, coding help, and everyday automation
Tiny GPT-5 lane for routing, extraction, classification, and bulk jobs
Open GPT reasoning model for self-hosted agents and controllable deployments
Open GPT reasoning model for self-hosted agents and controllable deployments
Qwen coding model for software agents, repository edits, and code reasoning
Hosted Qwen coder for software agents, repo edits, and long-context code
Updated large open Qwen3 MoE instruct model for multilingual chat, coding, and tool use
Instruct model with native audio input for speech understanding and tool use
High-effort o3 tier for difficult technical reasoning and careful answers
Mistral model for multilingual chat, reasoning, and tool-assisted workflows
Deliberate o-series reasoner for hard math, coding, and multi-step analysis
Fast o-series model for compact reasoning, coding, and tool use
Nemotron model for efficient reasoning, coding, and specialized AI agents
Tiny GPT-4.1 option for classification, routing, and very high-volume tasks
Affordable GPT-4.1 lane for fast coding help and structured extraction
Long-lived GPT workhorse for coding, instruction following, and production apps
Smaller Qwen coder for efficient local agents and repo-level fixes
Open Qwen coding heavyweight for repository reasoning and agentic engineering
March 2025 checkpoint of DeepSeek-V3 with improved reasoning and coding
O-series reasoning model for hard analysis, math, coding, and planning
Mistral reasoning model for transparent analysis, math, and complex decisions
Cohere command model for multilingual enterprise agents, tools, and chat
Qwen reasoning model for deliberate problem solving, math, and coding
Sonar search model for autonomous research and citation-backed long-form reports
Classic open reasoning model for transparent math, coding, and deliberate problem solving
Open DeepSeek MoE chat model for coding, math, and general reasoning
Smaller o-series reasoner for economical coding, math, and planning tasks
Popular open Llama workhorse for multilingual chat, coding, and self-hosting
O-series reasoning model for hard analysis, math, coding, and planning
Flagship model for demanding analysis, coding, and production agent workflows
Efficient model for low-latency assistance, extraction, and routine automation
Efficient model for low-latency assistance, extraction, and routine automation
Cohere retrieval model for long-context chat and enterprise RAG workflows
GPT model for general reasoning, writing, coding, and tool-assisted tasks
Open coding-focused Qwen model for code generation, repair, and repository reasoning
Flagship Mistral model for advanced reasoning, coding, and multilingual work
| Model | Creator | Inputs | Context | Input | Output | Released | Compare |
|---|---|---|---|---|---|---|---|
| Kimi K2 Thinkingmoonshotai/kimi-k2-thinking | 262.144K | $0.4 | $2.5 | 2025-11-06 | |||
| Claude Opus 4.5anthropic/claude-opus-4-5-20251101 | 200K | $5 | $25 | 2025-11-01 | |||
| MiniMax-M2minimax/MiniMax-M2 | 204.8K | $0.3 | $1.2 | 2025-10-27 | |||
| GPT-5 Proopenai/gpt-5-pro | 400K | $15 | $120 | 2025-10-06 | |||
| GLM-4.6zhipuai/glm-4.6 | 204.8K | $0.6 | $2.2 | 2025-09-30 | |||
| Qwen3 VL 235B A22B Instructalibaba/qwen3-vl-235b-a22b-instruct | 131.072K | $0.2 | $0.88 | 2025-09-23 | |||
| Qwen3 Maxalibaba/qwen3-max | 262.144K | $1.2 | $6 | 2025-09-23 | |||
| Qwen3 VL 235B A22B Thinkingalibaba/qwen3-vl-235b-a22b-thinking | 131.072K | $0.4 | $4 | 2025-09-23 | |||
| Gemma-SEA-LION-v4-27B-ITaisingapore/gemma-sea-lion-v4-27b-it | 128K | — | — | 2025-09-23 | |||
| Qwen3-Next 80B-A3B (Thinking)alibaba/qwen3-next-80b-a3b-thinking | 131.072K | $0.5 | $6 | 2025-09 | |||
| Qwen3-Next 80B-A3B Instructalibaba/qwen3-next-80b-a3b-instruct | 131.072K | $0.5 | $2 | 2025-09 | |||
| Nano Bananagoogle/gemini-2.5-flash-image | 32.768K | $0.3 | $30 | 2025-08-26 | |||
| Nemotron Nano 9B v2nvidia/nemotron-nano-9b-v2 | 131.072K | $0.06 | $0.23 | 2025-08-18 | |||
| GPT-5openai/gpt-5 | 400K | $1.25 | $10 | 2025-08-07 | |||
| GPT-5 Miniopenai/gpt-5-mini | 400K | $0.25 | $2 | 2025-08-07 | |||
| GPT-5 Nanoopenai/gpt-5-nano | 400K | $0.05 | $0.4 | 2025-08-07 | |||
| GPT OSS 120Bopenai/gpt-oss-120b | 131.072K | $0.03 | $0.17 | 2025-08-05 | |||
| GPT OSS 20Bopenai/gpt-oss-20b | 131.072K | $0.02 | $0.1 | 2025-08-05 | |||
| Qwen3 Coder Flashalibaba/qwen3-coder-flash | 1M | $0.3 | $1.5 | 2025-07-28 | |||
| Qwen3 Coder Plusalibaba/qwen3-coder-plus | 1.04858M | $1 | $5 | 2025-07-23 | |||
| Qwen3 235B-A22B Instruct 2507alibaba/qwen3-235b-a22b-instruct-2507 | 262.144K | $0.069 | $0.455 | 2025-07-21 | |||
| Voxtral Small (latest)mistral/voxtral-small-latest | 32K | $0.1 | $0.3 | 2025-07-15 | |||
| o3-proopenai/o3-pro | 200K | $20 | $80 | 2025-06-10 | |||
| Mistral Medium 3mistral/mistral-medium-2505 | 131.072K | $0.4 | $2 | 2025-05-07 | |||
| o3openai/o3 | 200K | $2 | $8 | 2025-04-16 | |||
| o4-miniopenai/o4-mini | 200K | $1.1 | $4.4 | 2025-04-16 | |||
| Llama 3.1 Nemotron 70B Instructnvidia/llama-3.1-nemotron-70b-instruct | 128K | — | — | 2025-04-15 | |||
| GPT-4.1 nanoopenai/gpt-4.1-nano | 1.04758M | $0.1 | $0.4 | 2025-04-14 | |||
| GPT-4.1 miniopenai/gpt-4.1-mini | 1.04758M | $0.4 | $1.6 | 2025-04-14 | |||
| GPT-4.1openai/gpt-4.1 | 1.04758M | $2 | $8 | 2025-04-14 | |||
| Qwen3-Coder 30B-A3B Instructalibaba/qwen3-coder-30b-a3b-instruct | 262.144K | $0.45 | $2.25 | 2025-04 | |||
| Qwen3-Coder 480B-A35B Instructalibaba/qwen3-coder-480b-a35b-instruct | 262.144K | $1.5 | $7.5 | 2025-04 | |||
| DeepSeek V3 0324deepseek/deepseek-v3-0324 | 163.84K | $0.2 | $0.8 | 2025-03-24 | |||
| o1-proopenai/o1-pro | 200K | $150 | $600 | 2025-03-19 | |||
| Magistral Medium (latest)mistral/magistral-medium-latest | 128K | $2 | $5 | 2025-03-17 | |||
| Command Acohere/command-a-03-2025 | 256K | $2.5 | $10 | 2025-03-13 | |||
| QwQ Plusalibaba/qwq-plus | 131.072K | $0.8 | $2.4 | 2025-03-05 | |||
| Sonar Deep Researchperplexity/sonar-deep-research | 128K | $2 | $8 | 2025-02-01 | |||
| DeepSeek-R1deepseek/deepseek-r1 | 128K | $0.7 | $2.5 | 2025-01-20 | |||
| DeepSeek-V3deepseek/deepseek-v3 | 131.072K | $0.27 | $1.12 | 2024-12-26 | |||
| o3-miniopenai/o3-mini | 200K | $1.1 | $4.4 | 2024-12-20 | |||
| Llama-3.3-70B-Instructmeta/llama-3.3-70b-instruct | 128K | $0.1 | $0.32 | 2024-12-06 | |||
| o1openai/o1 | 200K | $15 | $60 | 2024-12-05 | |||
| Nova Proamazon/nova-pro | 300K | $0.8 | $3.2 | 2024-12-03 | |||
| Nova Microamazon/nova-micro | 128K | $0.035 | $0.14 | 2024-12-03 | |||
| Nova Liteamazon/nova-lite | 300K | $0.06 | $0.24 | 2024-12-03 | |||
| Command R7Bcohere/command-r7b-12-2024 | 128K | $0.037 | $0.15 | 2024-12-02 | |||
| GPT-4o (2024-11-20)openai/gpt-4o-2024-11-20 | 128K | $2.5 | $10 | 2024-11-20 | |||
| Qwen2.5-Coder-32B-Instructalibaba/qwen2.5-coder-32b-instruct | 131.072K | $0.06 | $0.2 | 2024-11-12 | |||
| Mistral Large (latest)mistral/mistral-large-latest | 262.144K | $0.5 | $1.5 | 2024-11-01 |