GPT-6 Astra is OpenAI's most capable model for complex reasoning, coding, computer use, research, and document creation.
Models
Every model in the catalog with source-linked pricing, context limits, provider availability, and published benchmark results.
Fast variant of GPT-6 Astra for low-latency assistance and high-volume workloads.
Muse Spark 1.3 is a multimodal reasoning model from Meta for long-running agentic, multi-agent, and coding workflows. It improves long-horizon agent collaboration, instruction following, and coding efficiency relative to Muse Spark 1.2.
Google's most intelligent Flash model, engineered for long-horizon software engineering, autonomous agents, and complex enterprise workflows
2026-09-02 upgraded snapshot of Qwen3.8 Max with stronger coding, collaborative agents, and multimodal document understanding
Claude model for demanding reasoning and long-horizon agentic work
Finance-enhanced model for financial research, multi-step investment workflows, and long-horizon planning and execution
Qwen vision-language model for visual reasoning, documents, and agent tasks
Speech transcription model for accurate audio-to-text and captioning workflows
Experimental multimodal DeepSeek V4 Flash model for image understanding, coding, and agentic work
High-efficiency Gemini model for agentic workflows, coding, and multimodal reasoning
High-efficiency Gemini model for agentic workflows, coding, and multimodal reasoning
xAI's frontier model for long-running agents, coding, knowledge work, and visual projects
Microsoft coding model with native vision support, optimized for fast and efficient software development
Image model for prompt-driven generation, editing, and visual design workflows
Upstage's flagship model, specialized for agentic use
Muse Spark 1.2 is a coding-focused update to Muse Spark 1.1 with improvements in code generation, complex debugging, codebase understanding, and end-to-end developer workflows.
Japanese-specialized reasoning model based on Kimi K2.6 and tuned for Japanese language, culture, and business workflows
2.4-trillion-parameter MoE flagship for coding, professional work, multimodal understanding, and long-horizon agentic workflows
Strongest Claude Opus model for coding, agents, and professional work
Fast Gemini model balancing multimodal reasoning, tool use, and cost
Fast Gemini model balancing multimodal reasoning, tool use, and cost
Fast Gemini model balancing multimodal reasoning, tool use, and cost
Preview Qwen flagship for million-token multimodal reasoning and long-horizon agentic workflows
Lightweight multimodal Qwen model for high-throughput text, image, and video tasks
Cost-efficient GPT-5.6 model for fast, high-volume workloads
Frontier GPT-5.6 model for complex professional work, coding, and agentic workflows
Balanced GPT-5.6 model for capable, cost-efficient everyday work
xAI's Grok model for chat, coding, agentic tools, and lower hallucination risk
Realtime speech-to-speech model with configurable reasoning, tool use, and robust voice-agent behavior
Fastest, most cost-efficient Gemini image model for high-volume 1K generation and editing
Video generation and editing model for fast, conversational text- and image-to-video workflows
Everyday Claude agent model for coding, planning, browsing, and general work
Meituan LongCat-2.0, a reasoning model with tool calling and a 1M-token context window
ByteDance Seed model optimized for character-driven dialogue and consistent conversational behavior
Faster ByteDance Seed 2.1 model for multimodal reasoning and latency-sensitive agent workflows
Flagship ByteDance Seed 2.1 model for complex multimodal reasoning, coding, and agents
Rolling ByteDance Seed model for rapidly updated reasoning, coding, and agent capabilities
Quality-first multi-agent model for hard research, analysis, and competitions
Multi-agent model for routing expert agents across complex analytical tasks
Low-latency audio-to-audio model for real-time speech translation across 70+ languages
Restricted Claude model for advanced cybersecurity and biology research workflows
Claude model for creative writing, analysis, and controlled agent workflows
Microsoft coding model built for fast, efficient assistance in everyday developer workflows
Multimodal Qwen workhorse for long-context agents, visual inputs, and coding
Video model for image-to-video generation, editing, and extension workflows
Top Claude Opus tier for the hardest reasoning, coding, and long-horizon agents
Image model for prompt-driven generation, editing, and visual design workflows
Nano Banana Pro for higher-fidelity image generation and design-heavy edits
Qwen frontier model tuned for agent frameworks, coding assistants, and long tasks
| Model | Creator | Inputs | Context | Input | Output | Released | Compare |
|---|---|---|---|---|---|---|---|
| GPT-6 Astraopenai/gpt-6-astra | 1.05M | $10 | $50 | 2026-09-04 | |||
| GPT-6 Astra (Fast)openai/gpt-6-astra-fast | 1.05M | $20 | $100 | 2026-09-04 | |||
| Muse Spark 1.3meta/muse-spark-1.3 | 1.04858M | $1.25 | $4.25 | 2026-09-02 | |||
| Gemini 3.8 Flashgoogle/gemini-3.8-flash | 1.04858M | $0.75 | $3.75 | 2026-09-02 | |||
| Qwen3.8 Max 0902alibaba/qwen3.8-max-0902 | 1M | $1.71 | $5.14 | 2026-09-02 | |||
| Claude Fable 5.1anthropic/claude-fable-5-1 | 1M | $10 | $50 | 2026-09-01 | |||
| Ling 3.0 Flash Fininclusionai/ling-3.0-flash-fin | 262.144K | $0.06 | $0.18 | 2026-08-27 | |||
| Qwen3.8 Flashalibaba/qwen3.8-flash | 1M | $0.15 | $0.47 | 2026-08-26 | |||
| Gemini 3.5 Transcribe Livegoogle/gemini-3.5-transcribe-live | Not documented | — | — | 2026-08-26 | |||
| DeepSeek V4 Flash Vision Expdeepseek/deepseek-v4-flash-vision-exp | 1M | $0.15 | $0.6 | 2026-08-21 | |||
| Gemini 3.7 Flashgoogle/gemini-3.7-flash | 1.04858M | $0.75 | $3.75 | 2026-08-13 | |||
| Gemini Flash Latestgoogle/gemini-flash-latest | 1.04858M | $0.75 | $3.75 | 2026-08-13 | |||
| Grok 4.6xai/grok-4.6 | 500K | $2 | $6 | 2026-08-12 | |||
| MAI-Code-1.1-Flashmicrosoft/mai-code-1.1-flash | 256K | $0.2 | $1.2 | 2026-08-11 | |||
| Grok Imagine Image 2.0xai/grok-imagine-image-2.0 | 8K | — | — | 2026-08-07 | |||
| Solar Pro 4upstage/solar-pro4 | 524.288K | $0.3 | $1.2 | 2026-08-06 | |||
| Muse Spark 1.2meta/muse-spark-1.2 | 1.04858M | $1.25 | $4.25 | 2026-08-05 | |||
| Sakana Namazusakana/sakana-namazu | 262.144K | $0.95 | $4 | 2026-08-03 | |||
| Qwen3.8 Maxalibaba/qwen3.8-max | 1M | $2 | $6 | 2026-08-03 | |||
| Claude Opus 5anthropic/claude-opus-5 | 1M | $5 | $25 | 2026-07-24 | |||
| Gemini 3.5 Flash Litegoogle/gemini-3.5-flash-lite | 1.04858M | $0.3 | $2.5 | 2026-07-21 | |||
| Gemini Flash-Lite Latestgoogle/gemini-flash-lite-latest | 1.04858M | $0.3 | $2.5 | 2026-07-21 | |||
| Gemini 3.6 Flashgoogle/gemini-3.6-flash | 1.04858M | $0.75 | $3.75 | 2026-07-21 | |||
| Qwen3.8 Max Previewalibaba/qwen3.8-max-preview | 1M | $2 | $6 | 2026-07-19 | |||
| Qwen3.7 Flashalibaba/qwen3.7-flash | 1M | $0.028 | $0.113 | 2026-07-15 | |||
| GPT-5.6 Lunaopenai/gpt-5.6-luna | 1.05M | $0.2 | $1.2 | 2026-07-09 | |||
| GPT-5.6 Solopenai/gpt-5.6-sol | 1.05M | $4 | $20 | 2026-07-09 | |||
| GPT-5.6 Terraopenai/gpt-5.6-terra | 1.05M | $2 | $12 | 2026-07-09 | |||
| Grok 4.5xai/grok-4.5 | 500K | $2 | $6 | 2026-07-08 | |||
| GPT-Realtime-2.1openai/gpt-realtime-2.1 | 128K | $4 | $24 | 2026-07-06 | |||
| Nano Banana 2 Litegoogle/gemini-3.1-flash-lite-image | 65.536K | $0.25 | $30 | 2026-06-30 | |||
| Gemini Omni Flash Previewgoogle/gemini-omni-flash-preview | 1.04858M | $1.5 | $17.5 | 2026-06-30 | |||
| Claude Sonnet 5anthropic/claude-sonnet-5 | 1M | $2 | $10 | 2026-06-30 | |||
| LongCat-2.0meituan/longcat-2.0 | 1M | $0.3 | $1.2 | 2026-06-30 | |||
| Seed Characterbytedance-seed/seed-character | 256K | $0.119 | $0.297 | 2026-06-23 | |||
| Seed 2.1 Turbobytedance-seed/seed-2.1-turbo | 256K | $0.354 | $1.77 | 2026-06-23 | |||
| Seed 2.1 Probytedance-seed/seed-2.1-pro | 256K | $0.707 | $3.536 | 2026-06-23 | |||
| Seed Evolvingbytedance-seed/seed-evolving | 256K | $0.884 | $4.42 | 2026-06-23 | |||
| Fugu Ultrasakana/fugu-ultra | 1M | $5 | $30 | 2026-06-15 | |||
| Fugusakana/fugu | 1M | — | — | 2026-06-15 | |||
| Gemini 3.5 Live Translate Previewgoogle/gemini-3.5-live-translate-preview | 131.072K | $3.5 | $21 | 2026-06-09 | |||
| Claude Mythos 5anthropic/claude-mythos-5 | 1M | $10 | $50 | 2026-06-09 | |||
| Claude Fable 5anthropic/claude-fable-5 | 1M | $10 | $50 | 2026-06-09 | |||
| MAI-Code-1-Flashmicrosoft/mai-code-1-flash | 256K | $0.75 | $4.5 | 2026-06-02 | |||
| Qwen3.7 Plusalibaba/qwen3.7-plus | 1M | $0.5 | $3 | 2026-06-02 | |||
| Grok Imagine Video 1.5xai/grok-imagine-video-1.5 | 1.024K | — | — | 2026-05-30 | |||
| Claude Opus 4.8anthropic/claude-opus-4-8 | 1M | $5 | $25 | 2026-05-28 | |||
| Nano Banana 2google/gemini-3.1-flash-image | 131.072K | $0.5 | $60 | 2026-05-28 | |||
| Nano Banana Progoogle/gemini-3-pro-image | 65.536K | $2 | $120 | 2026-05-28 | |||
| Qwen3.7 Maxalibaba/qwen3.7-max | 1M | $2.5 | $7.5 | 2026-05-21 |