DeepSeek V4.1 Flash model for reasoning and agentic coding
Models
Every model in the catalog with source-linked pricing, context limits, provider availability, and published benchmark results.
GPT-6 Astra is OpenAI's most capable model for complex reasoning, coding, computer use, research, and document creation.
Fast variant of GPT-6 Astra for low-latency assistance and high-volume workloads.
Muse Spark 1.3 is a multimodal reasoning model from Meta for long-running agentic, multi-agent, and coding workflows. It improves long-horizon agent collaboration, instruction following, and coding efficiency relative to Muse Spark 1.2.
2026-09-02 upgraded snapshot of Qwen3.8 Max with stronger coding, collaborative agents, and multimodal document understanding
Google's most intelligent Flash model, engineered for long-horizon software engineering, autonomous agents, and complex enterprise workflows
Claude model for demanding reasoning and long-horizon agentic work
A next-generation productivity model with significantly enhanced Agent and complex task execution capabilities.
Finance-enhanced model for financial research, multi-step investment workflows, and long-horizon planning and execution
Open-weight experimental preview of the Qwen4 architecture: hybrid-attention MoE (125B total, 6B active) with vision encoder for coding, agent tasks, and image and video understanding
Qwen vision-language model for visual reasoning, documents, and agent tasks
Native multimodal GLM model for efficient coding and long-horizon agent tasks
Experimental multimodal DeepSeek V4 Flash model for image understanding, coding, and agentic work
Mixture-of-experts coding-reasoning model for agentic software tasks, tool use, and image understanding
Flagship GLM model for long-horizon coding, agents, and complex project delivery
Dense 27B vision-language model for coding, agent tasks, and image and video understanding
High-efficiency Gemini model for agentic workflows, coding, and multimodal reasoning
High-efficiency Gemini model for agentic workflows, coding, and multimodal reasoning
DeepSeek V4 Pro snapshot with million-token context and support for thinking and non-thinking modes
Open-weight sparse MoE (2.4T total, 95B active), the open-weight twin of Qwen3.8 Max for coding, research, complex reasoning, and agentic workflows
xAI's frontier model for long-running agents, coding, knowledge work, and visual projects
Fast NVIDIA Nemotron MoE for reliable agentic tasks across enterprise workloads
Fast NVIDIA Nemotron MoE for reliable agentic tasks across enterprise workloads
Microsoft coding model with native vision support, optimized for fast and efficient software development
Muse Glimmer is a 30-billion-parameter open-weight multimodal model from Meta Superintelligence Labs, distilled from Muse Spark for always-on local agents, tool use, coding, and image understanding.
Upstage's flagship model, specialized for agentic use
Muse Spark 1.2 is a coding-focused update to Muse Spark 1.1 with improvements in code generation, complex debugging, codebase understanding, and end-to-end developer workflows.
2.4-trillion-parameter MoE flagship for coding, professional work, multimodal understanding, and long-horizon agentic workflows
Japanese-specialized reasoning model based on Kimi K2.6 and tuned for Japanese language, culture, and business workflows
Official DeepSeek V4 Flash release with enhanced agentic capabilities and integrated DSpark speculative decoding
Multimodal MoE reasoning model (276B total, 12B active) for text, image, and audio
Strongest Claude Opus model for coding, agents, and professional work
Fast Gemini model balancing multimodal reasoning, tool use, and cost
Agentic coding model from Poolside in the XS size class for local deployment
Fast Gemini model balancing multimodal reasoning, tool use, and cost
Fast Gemini model balancing multimodal reasoning, tool use, and cost
Preview Qwen flagship for million-token multimodal reasoning and long-horizon agentic workflows
Multimodal Kimi model with 1M context and toggleable max-effort thinking for long-horizon agent work
Lightweight multimodal Qwen model for high-throughput text, image, and video tasks
Multimodal MoE reasoning model (975B total, 41B active) for text, image, and audio
Frontier GPT-5.6 model for complex professional work, coding, and agentic workflows
Cost-efficient GPT-5.6 model for fast, high-volume workloads
Balanced GPT-5.6 model for capable, cost-efficient everyday work
xAI's Grok model for chat, coding, agentic tools, and lower hallucination risk
Realtime speech-to-speech model with configurable reasoning, tool use, and robust voice-agent behavior
Tencent Hy reasoning model for coding, instruction following, and agent tasks
Agentic coding model from Poolside in the XS size class for local deployment
Meituan LongCat-2.0, a reasoning model with tool calling and a 1M-token context window
Everyday Claude agent model for coding, planning, browsing, and general work
Fastest, most cost-efficient Gemini image model for high-volume 1K generation and editing
| Model | Creator | Inputs | Context | Input | Output | Released | Compare |
|---|---|---|---|---|---|---|---|
| DeepSeek V4.1 Flashdeepseek/deepseek-v4.1-flash | 1M | $0.15 | $0.6 | 2026-09-10 | |||
| GPT-6 Astraopenai/gpt-6-astra | 1.05M | $10 | $50 | 2026-09-04 | |||
| GPT-6 Astra (Fast)openai/gpt-6-astra-fast | 1.05M | $20 | $100 | 2026-09-04 | |||
| Muse Spark 1.3meta/muse-spark-1.3 | 1.04858M | $1.25 | $4.25 | 2026-09-02 | |||
| Qwen3.8 Max 0902alibaba/qwen3.8-max-0902 | 1M | $1.71 | $5.14 | 2026-09-02 | |||
| Gemini 3.8 Flashgoogle/gemini-3.8-flash | 1.04858M | $0.75 | $3.75 | 2026-09-02 | |||
| Claude Fable 5.1anthropic/claude-fable-5-1 | 1M | $10 | $50 | 2026-09-01 | |||
| Hy4 previewtencent/hy4-preview | 1.024M | $0.67 | $2 | 2026-08-28 | |||
| Ling 3.0 Flash Fininclusionai/ling-3.0-flash-fin | 262.144K | $0.06 | $0.18 | 2026-08-27 | |||
| Qwen3.8 Flash Nextalibaba/qwen3.8-flash-next | 262.144K | $0.12 | $0.4 | 2026-08-27 | |||
| Qwen3.8 Flashalibaba/qwen3.8-flash | 1M | $0.15 | $0.47 | 2026-08-26 | |||
| GLM-5.3-Flashzhipuai/glm-5.3-flash | 1M | $0.075 | $0.25 | 2026-08-26 | |||
| DeepSeek V4 Flash Vision Expdeepseek/deepseek-v4-flash-vision-exp | 1M | $0.15 | $0.6 | 2026-08-21 | |||
| Ornith 1.5 35B A3Bdeepreinforce/ornith-1.5-35b-a3b | 262.144K | $0.1 | $0.4 | 2026-08-18 | |||
| GLM-5.3zhipuai/glm-5.3 | 1M | $1.4 | $4.4 | 2026-08-14 | |||
| Qwen3.8 27Balibaba/qwen3.8-27b | 262.144K | $0.1 | $0.4 | 2026-08-14 | |||
| Gemini 3.7 Flashgoogle/gemini-3.7-flash | 1.04858M | $0.75 | $3.75 | 2026-08-13 | |||
| Gemini Flash Latestgoogle/gemini-flash-latest | 1.04858M | $0.75 | $3.75 | 2026-08-13 | |||
| DeepSeek V4 Pro 0813deepseek/deepseek-v4-pro-0813 | 1M | $0.442 | $0.884 | 2026-08-12 | |||
| Qwen3.8 2.4T A95Balibaba/qwen3.8-2.4t-a95b | 262.144K | $2 | $6 | 2026-08-12 | |||
| Grok 4.6xai/grok-4.6 | 500K | $2 | $6 | 2026-08-12 | |||
| Nemotron 3.5 Lightning 30B A3Bnvidia/nemotron-3.5-lightning-30b-a3b | 262.144K | — | — | 2026-08-11 | |||
| Nemotron 3.5 Lightning 30B A3Bnvidia/nemotron-3.5-lightning | 262.144K | $0.05 | $0.2 | 2026-08-11 | |||
| MAI-Code-1.1-Flashmicrosoft/mai-code-1.1-flash | 256K | $0.2 | $1.2 | 2026-08-11 | |||
| Muse Glimmer 30Bmeta/muse-glimmer-30b | 131.072K | $0.2 | $0.8 | 2026-08-10 | |||
| Solar Pro 4upstage/solar-pro4 | 524.288K | $0.3 | $1.2 | 2026-08-06 | |||
| Muse Spark 1.2meta/muse-spark-1.2 | 1.04858M | $1.25 | $4.25 | 2026-08-05 | |||
| Qwen3.8 Maxalibaba/qwen3.8-max | 1M | $2 | $6 | 2026-08-03 | |||
| Sakana Namazusakana/sakana-namazu | 262.144K | $0.95 | $4 | 2026-08-03 | |||
| DeepSeek V4 Flash 0731deepseek/deepseek-v4-flash-0731 | 1M | $0.035 | $0.07 | 2026-07-31 | |||
| Inkling Smallthinkingmachines/inkling-small | 1.04858M | $0.45 | $1.2 | 2026-07-30 | |||
| Claude Opus 5anthropic/claude-opus-5 | 1M | $5 | $25 | 2026-07-24 | |||
| Gemini Flash-Lite Latestgoogle/gemini-flash-lite-latest | 1.04858M | $0.3 | $2.5 | 2026-07-21 | |||
| Laguna S 2.1poolside/laguna-s-2.1 | 1.04858M | $0.09 | $0.18 | 2026-07-21 | |||
| Gemini 3.6 Flashgoogle/gemini-3.6-flash | 1.04858M | $0.75 | $3.75 | 2026-07-21 | |||
| Gemini 3.5 Flash Litegoogle/gemini-3.5-flash-lite | 1.04858M | $0.3 | $2.5 | 2026-07-21 | |||
| Qwen3.8 Max Previewalibaba/qwen3.8-max-preview | 1M | $2 | $6 | 2026-07-19 | |||
| Kimi K3moonshotai/kimi-k3 | 1.04858M | $3 | $15 | 2026-07-16 | |||
| Qwen3.7 Flashalibaba/qwen3.7-flash | 1M | $0.028 | $0.113 | 2026-07-15 | |||
| Inklingthinkingmachines/inkling | 1.04858M | $1.87 | $4.68 | 2026-07-15 | |||
| GPT-5.6 Solopenai/gpt-5.6-sol | 1.05M | $4 | $20 | 2026-07-09 | |||
| GPT-5.6 Lunaopenai/gpt-5.6-luna | 1.05M | $0.2 | $1.2 | 2026-07-09 | |||
| GPT-5.6 Terraopenai/gpt-5.6-terra | 1.05M | $2 | $12 | 2026-07-09 | |||
| Grok 4.5xai/grok-4.5 | 500K | $2 | $6 | 2026-07-08 | |||
| GPT-Realtime-2.1openai/gpt-realtime-2.1 | 128K | $4 | $24 | 2026-07-06 | |||
| Hy3tencent/hy3 | 256K | $0.066 | $0.26 | 2026-07-06 | |||
| Laguna XS 2.1poolside/laguna-xs-2.1 | 262.144K | $0.06 | $0.12 | 2026-07-02 | |||
| LongCat-2.0meituan/longcat-2.0 | 1M | $0.3 | $1.2 | 2026-06-30 | |||
| Claude Sonnet 5anthropic/claude-sonnet-5 | 1M | $2 | $10 | 2026-06-30 | |||
| Nano Banana 2 Litegoogle/gemini-3.1-flash-lite-image | 65.536K | $0.25 | $30 | 2026-06-30 |