DeepSeek V4.1 Flash model for reasoning and agentic coding
Models
Every model in the catalog with source-linked pricing, context limits, provider availability, and published benchmark results.
GPT-6 Astra is OpenAI's most capable model for complex reasoning, coding, computer use, research, and document creation.
Muse Spark 1.3 is a multimodal reasoning model from Meta for long-running agentic, multi-agent, and coding workflows. It improves long-horizon agent collaboration, instruction following, and coding efficiency relative to Muse Spark 1.2.
Google's most intelligent Flash model, engineered for long-horizon software engineering, autonomous agents, and complex enterprise workflows
2026-09-02 upgraded snapshot of Qwen3.8 Max with stronger coding, collaborative agents, and multimodal document understanding
Claude model for demanding reasoning and long-horizon agentic work
Qwen vision-language model for visual reasoning, documents, and agent tasks
Native multimodal GLM model for efficient coding and long-horizon agent tasks
Experimental multimodal DeepSeek V4 Flash model for image understanding, coding, and agentic work
Dense 27B vision-language model for coding, agent tasks, and image and video understanding
High-efficiency Gemini model for agentic workflows, coding, and multimodal reasoning
xAI's frontier model for long-running agents, coding, knowledge work, and visual projects
Muse Glimmer is a 30-billion-parameter open-weight multimodal model from Meta Superintelligence Labs, distilled from Muse Spark for always-on local agents, tool use, coding, and image understanding.
Muse Spark 1.2 is a coding-focused update to Muse Spark 1.1 with improvements in code generation, complex debugging, codebase understanding, and end-to-end developer workflows.
Japanese-specialized reasoning model based on Kimi K2.6 and tuned for Japanese language, culture, and business workflows
2.4-trillion-parameter MoE flagship for coding, professional work, multimodal understanding, and long-horizon agentic workflows
Multimodal MoE reasoning model (276B total, 12B active) for text, image, and audio
Strongest Claude Opus model for coding, agents, and professional work
Fast Gemini model balancing multimodal reasoning, tool use, and cost
Fast Gemini model balancing multimodal reasoning, tool use, and cost
Multimodal Kimi model with 1M context and toggleable max-effort thinking for long-horizon agent work
Multimodal MoE reasoning model (975B total, 41B active) for text, image, and audio
Lightweight multimodal Qwen model for high-throughput text, image, and video tasks
Frontier GPT-5.6 model for complex professional work, coding, and agentic workflows
Balanced GPT-5.6 model for capable, cost-efficient everyday work
Cost-efficient GPT-5.6 model for fast, high-volume workloads
xAI's Grok model for chat, coding, agentic tools, and lower hallucination risk
Everyday Claude agent model for coding, planning, browsing, and general work
Quality-first multi-agent model for hard research, analysis, and competitions
Coding-focused Kimi model, stronger on long-horizon repo work with less overthinking
Claude model for creative writing, analysis, and controlled agent workflows
Safety model for policy screening, moderation, and risk-aware routing workflows
Multimodal Qwen workhorse for long-context agents, visual inputs, and coding
MiniMax multimodal model for long-context coding, perception, and agent planning
Newer StepFun flash model for faster agents, coding, and multimodal prompts
Top Claude Opus tier for the hardest reasoning, coding, and long-horizon agents
Image model for prompt-driven generation, editing, and visual design workflows
Fast Gemini model balancing multimodal reasoning, tool use, and cost
Low-latency Gemini model for high-volume multimodal and agent workloads
Qwen vision-language model for visual reasoning, documents, and agent tasks
Highest-accuracy GPT-5.5 tier for slower, precision-heavy reasoning and coding
Default frontier GPT for coding, computer use, research, and knowledge work
Qwen vision-language model for visual reasoning, documents, and agent tasks
Open MiMo model for multimodal coding agents and long-context automation
Multimodal Kimi workhorse for agent loops, coding tasks, and visual context
xAI's default Grok for chat, coding, agentic tools, and lower hallucination risk
Open multimodal Qwen MoE for local agents that need vision, audio, and code
Fast Grok coding model tuned for agentic engineering and iterative edits
Stronger Opus tier for advanced software work and high-stakes reasoning
Muse Spark is a natively multimodal reasoning model with support for tool-use, visual chain of thought, and multi-agent orchestration.
| Model | Creator | Inputs | Context | Input | Output | Released | Compare |
|---|---|---|---|---|---|---|---|
| DeepSeek V4.1 Flashdeepseek/deepseek-v4.1-flash | 1M | $0.15 | $0.6 | 2026-09-10 | |||
| GPT-6 Astraopenai/gpt-6-astra | 1.05M | $10 | $50 | 2026-09-04 | |||
| Muse Spark 1.3meta/muse-spark-1.3 | 1.04858M | $1.25 | $4.25 | 2026-09-02 | |||
| Gemini 3.8 Flashgoogle/gemini-3.8-flash | 1.04858M | $0.75 | $3.75 | 2026-09-02 | |||
| Qwen3.8 Max 0902alibaba/qwen3.8-max-0902 | 1M | $1.71 | $5.14 | 2026-09-02 | |||
| Claude Fable 5.1anthropic/claude-fable-5-1 | 1M | $10 | $50 | 2026-09-01 | |||
| Qwen3.8 Flashalibaba/qwen3.8-flash | 1M | $0.15 | $0.47 | 2026-08-26 | |||
| GLM-5.3-Flashzhipuai/glm-5.3-flash | 1M | $0.075 | $0.25 | 2026-08-26 | |||
| DeepSeek V4 Flash Vision Expdeepseek/deepseek-v4-flash-vision-exp | 1M | $0.15 | $0.6 | 2026-08-21 | |||
| Qwen3.8 27Balibaba/qwen3.8-27b | 262.144K | $0.1 | $0.4 | 2026-08-14 | |||
| Gemini 3.7 Flashgoogle/gemini-3.7-flash | 1.04858M | $0.75 | $3.75 | 2026-08-13 | |||
| Grok 4.6xai/grok-4.6 | 500K | $2 | $6 | 2026-08-12 | |||
| Muse Glimmer 30Bmeta/muse-glimmer-30b | 131.072K | $0.2 | $0.8 | 2026-08-10 | |||
| Muse Spark 1.2meta/muse-spark-1.2 | 1.04858M | $1.25 | $4.25 | 2026-08-05 | |||
| Sakana Namazusakana/sakana-namazu | 262.144K | $0.95 | $4 | 2026-08-03 | |||
| Qwen3.8 Maxalibaba/qwen3.8-max | 1M | $2 | $6 | 2026-08-03 | |||
| Inkling Smallthinkingmachines/inkling-small | 1.04858M | $0.45 | $1.2 | 2026-07-30 | |||
| Claude Opus 5anthropic/claude-opus-5 | 1M | $5 | $25 | 2026-07-24 | |||
| Gemini 3.5 Flash Litegoogle/gemini-3.5-flash-lite | 1.04858M | $0.3 | $2.5 | 2026-07-21 | |||
| Gemini 3.6 Flashgoogle/gemini-3.6-flash | 1.04858M | $0.75 | $3.75 | 2026-07-21 | |||
| Kimi K3moonshotai/kimi-k3 | 1.04858M | $3 | $15 | 2026-07-16 | |||
| Inklingthinkingmachines/inkling | 1.04858M | $1.87 | $4.68 | 2026-07-15 | |||
| Qwen3.7 Flashalibaba/qwen3.7-flash | 1M | $0.028 | $0.113 | 2026-07-15 | |||
| GPT-5.6 Solopenai/gpt-5.6-sol | 1.05M | $4 | $20 | 2026-07-09 | |||
| GPT-5.6 Terraopenai/gpt-5.6-terra | 1.05M | $2 | $12 | 2026-07-09 | |||
| GPT-5.6 Lunaopenai/gpt-5.6-luna | 1.05M | $0.2 | $1.2 | 2026-07-09 | |||
| Grok 4.5xai/grok-4.5 | 500K | $2 | $6 | 2026-07-08 | |||
| Claude Sonnet 5anthropic/claude-sonnet-5 | 1M | $2 | $10 | 2026-06-30 | |||
| Fugu Ultrasakana/fugu-ultra | 1M | $5 | $30 | 2026-06-15 | |||
| Kimi K2.7 Codemoonshotai/kimi-k2.7-code | 262.144K | $0.95 | $4 | 2026-06-12 | |||
| Claude Fable 5anthropic/claude-fable-5 | 1M | $10 | $50 | 2026-06-09 | |||
| Nemotron 3.5 Content Safetynvidia/nemotron-3.5-content-safety | 128K | $0.2 | $0.2 | 2026-06-04 | |||
| Qwen3.7 Plusalibaba/qwen3.7-plus | 1M | $0.5 | $3 | 2026-06-02 | |||
| MiniMax-M3minimax/MiniMax-M3 | 1.04858M | $0.3 | $1.2 | 2026-06-01 | |||
| Step 3.7 Flashstepfun/step-3.7-flash | 256K | $0.185 | $1.11 | 2026-05-29 | |||
| Claude Opus 4.8anthropic/claude-opus-4-8 | 1M | $5 | $25 | 2026-05-28 | |||
| Nano Banana 2google/gemini-3.1-flash-image | 131.072K | $0.5 | $60 | 2026-05-28 | |||
| Gemini 3.5 Flashgoogle/gemini-3.5-flash | 1.04858M | $1.5 | $9 | 2026-05-19 | |||
| Gemini 3.1 Flash Litegoogle/gemini-3.1-flash-lite | 1.04858M | $0.25 | $1.5 | 2026-05-07 | |||
| Qwen3.6 Flashalibaba/qwen3.6-flash | 1M | $0.188 | $1.125 | 2026-04-27 | |||
| GPT-5.5 Proopenai/gpt-5.5-pro | 1.05M | $30 | $180 | 2026-04-23 | |||
| GPT-5.5openai/gpt-5.5 | 1.05M | $5 | $30 | 2026-04-23 | |||
| Qwen3.6 27Balibaba/qwen3.6-27b | 262.144K | $0.6 | $3.6 | 2026-04-22 | |||
| MiMo-V2.5xiaomi/mimo-v2.5 | 1.04858M | $0.14 | $0.28 | 2026-04-22 | |||
| Kimi K2.6moonshotai/kimi-k2.6 | 262.144K | $0.95 | $4 | 2026-04-21 | |||
| Grok 4.3xai/grok-4.3 | 1M | $1.25 | $2.5 | 2026-04-17 | |||
| Qwen3.6 35B-A3Balibaba/qwen3.6-35b-a3b | 262.144K | $0.248 | $1.485 | 2026-04-17 | |||
| Grok Build 0.1xai/grok-build-0.1 | 256K | $1 | $2 | 2026-04-16 | |||
| Claude Opus 4.7anthropic/claude-opus-4-7 | 1M | $5 | $25 | 2026-04-16 | |||
| Muse Spark 1.1meta/muse-spark-1.1 | 1.04858M | $1.25 | $4.25 | 2026-04-08 |