Video generation and editing model for fast, conversational text- and image-to-video workflows
Models
Every model in the catalog with source-linked pricing, context limits, provider availability, and published benchmark results.
Open coding-reasoning model for repository tasks and self-improving agents
Open coding-reasoning model for repository tasks and self-improving agents
Large coding-reasoning model for agentic software tasks and RL search
Large coding-reasoning model for agentic software tasks and RL search
Faster ByteDance Seed 2.1 model for multimodal reasoning and latency-sensitive agent workflows
Rolling ByteDance Seed model for rapidly updated reasoning, coding, and agent capabilities
ByteDance Seed model optimized for character-driven dialogue and consistent conversational behavior
Flagship ByteDance Seed 2.1 model for complex multimodal reasoning, coding, and agents
Quality-first multi-agent model for hard research, analysis, and competitions
Multi-agent model for routing expert agents across complex analytical tasks
Open flagship GLM for long-horizon coding agents and million-token context work
Lower-latency Kimi Code variant for interactive edits and coding-agent loops
Coding-focused Kimi model, stronger on long-horizon repo work with less overthinking
Claude model for creative writing, analysis, and controlled agent workflows
Restricted Claude model for advanced cybersecurity and biology research workflows
Cohere coding model for practical software engineering and agentic edits
Low-latency audio-to-audio model for real-time speech translation across 70+ languages
MiMo pro model for strong multimodal reasoning and agent execution
Largest Nemotron 3 model for maximum open-weight reasoning and agent accuracy
Safety model for policy screening, moderation, and risk-aware routing workflows
Microsoft coding model built for fast, efficient assistance in everyday developer workflows
Multimodal Qwen workhorse for long-context agents, visual inputs, and coding
MiniMax multimodal model for long-context coding, perception, and agent planning
Newer StepFun flash model for faster agents, coding, and multimodal prompts
Nano Banana Pro for higher-fidelity image generation and design-heavy edits
Image model for prompt-driven generation, editing, and visual design workflows
Top Claude Opus tier for the hardest reasoning, coding, and long-horizon agents
Qwen frontier model tuned for agent frameworks, coding assistants, and long tasks
Cohere's stronger command model for multilingual agents and enterprise workflows
Dense 1B-class open-source model for on-device and resource-constrained use, with native long-context support, Think / No Think chat modes, and tool calling
Fast Gemini model balancing multimodal reasoning, tool use, and cost
Low-latency Gemini model for high-volume multimodal and agent workloads
Compact GPT model for low-latency assistance and high-volume workloads
Balanced Mistral model for enterprise assistants, multilingual work, and tools
Balanced Mistral model for enterprise assistants, multilingual work, and tools
Poolside's open-weight model for agentic coding and long-horizon work
Open Nemotron omni model combining reasoning with text, vision, and audio
Agentic coding model from Poolside in the XS size class for local deployment
Qwen vision-language model for visual reasoning, documents, and agent tasks
Open MoE flagship with million-token context for coding and long agent runs
Fast DeepSeek V4 lane for economical reasoning, coding, and long-context work
Highest-accuracy GPT-5.5 tier for slower, precision-heavy reasoning and coding
DeepSeek V4 Pro initial snapshot with million-token context and support for thinking and non-thinking modes
Initial DeepSeek V4 Flash snapshot for economical reasoning, coding, and million-token agent workloads
Default frontier GPT for coding, computer use, research, and knowledge work
Open MiMo model for multimodal coding agents and long-context automation
Stronger MiMo Pro tier for multimodal reasoning and coding-agent execution
Qwen vision-language model for visual reasoning, documents, and agent tasks
Agentic model for autonomous multi-step research, synthesis, and cited reports
| Model | Creator | Inputs | Context | Input | Output | Released | Compare |
|---|---|---|---|---|---|---|---|
| Gemini Omni Flash Previewgoogle/gemini-omni-flash-preview | 1.04858M | $1.5 | $17.5 | 2026-06-30 | |||
| Ornith 1.0 9Bdeepreinforce/ornith-1.0-9b | 262.144K | — | — | 2026-06-25 | |||
| Ornith 1.0 31Bdeepreinforce/ornith-1.0-31b | 262.144K | — | — | 2026-06-25 | |||
| Ornith 1.0 397Bdeepreinforce/ornith-1.0-397b | 262.144K | — | — | 2026-06-25 | |||
| Ornith 1.0 35Bdeepreinforce/ornith-1.0-35b | 262.144K | — | — | 2026-06-25 | |||
| Seed 2.1 Turbobytedance-seed/seed-2.1-turbo | 256K | $0.354 | $1.77 | 2026-06-23 | |||
| Seed Evolvingbytedance-seed/seed-evolving | 256K | $0.884 | $4.42 | 2026-06-23 | |||
| Seed Characterbytedance-seed/seed-character | 256K | $0.119 | $0.297 | 2026-06-23 | |||
| Seed 2.1 Probytedance-seed/seed-2.1-pro | 256K | $0.707 | $3.536 | 2026-06-23 | |||
| Fugu Ultrasakana/fugu-ultra | 1M | $5 | $30 | 2026-06-15 | |||
| Fugusakana/fugu | 1M | — | — | 2026-06-15 | |||
| GLM-5.2zhipuai/glm-5.2 | 1M | $1.4 | $4.4 | 2026-06-13 | |||
| Kimi K2.7 Code Highspeedmoonshotai/kimi-k2.7-code-highspeed | 262.144K | $1.9 | $8 | 2026-06-12 | |||
| Kimi K2.7 Codemoonshotai/kimi-k2.7-code | 262.144K | $0.95 | $4 | 2026-06-12 | |||
| Claude Fable 5anthropic/claude-fable-5 | 1M | $10 | $50 | 2026-06-09 | |||
| Claude Mythos 5anthropic/claude-mythos-5 | 1M | $10 | $50 | 2026-06-09 | |||
| North Mini Codecohere/north-mini-code-1-0 | 256K | — | — | 2026-06-09 | |||
| Gemini 3.5 Live Translate Previewgoogle/gemini-3.5-live-translate-preview | 131.072K | $3.5 | $21 | 2026-06-09 | |||
| MiMo-V2.5-Pro-UltraSpeedxiaomi/mimo-v2.5-pro-ultraspeed | 1.04858M | $1.305 | $2.61 | 2026-06-08 | |||
| Nemotron 3 Ultra 550B A55Bnvidia/nemotron-3-ultra-550b-a55b | 1M | $0.5 | $2.5 | 2026-06-04 | |||
| Nemotron 3.5 Content Safetynvidia/nemotron-3.5-content-safety | 128K | $0.2 | $0.2 | 2026-06-04 | |||
| MAI-Code-1-Flashmicrosoft/mai-code-1-flash | 256K | $0.75 | $4.5 | 2026-06-02 | |||
| Qwen3.7 Plusalibaba/qwen3.7-plus | 1M | $0.5 | $3 | 2026-06-02 | |||
| MiniMax-M3minimax/MiniMax-M3 | 1.04858M | $0.3 | $1.2 | 2026-06-01 | |||
| Step 3.7 Flashstepfun/step-3.7-flash | 256K | $0.185 | $1.11 | 2026-05-29 | |||
| Nano Banana Progoogle/gemini-3-pro-image | 65.536K | $2 | $120 | 2026-05-28 | |||
| Nano Banana 2google/gemini-3.1-flash-image | 131.072K | $0.5 | $60 | 2026-05-28 | |||
| Claude Opus 4.8anthropic/claude-opus-4-8 | 1M | $5 | $25 | 2026-05-28 | |||
| Qwen3.7 Maxalibaba/qwen3.7-max | 1M | $2.5 | $7.5 | 2026-05-21 | |||
| Command A Pluscohere/command-a-plus-05-2026 | 128K | $2.5 | $10 | 2026-05-20 | |||
| MiniCPM5-1Bopenbmb/minicpm5-1b | 131.072K | $0.124 | $0.743 | 2026-05-19 | |||
| Gemini 3.5 Flashgoogle/gemini-3.5-flash | 1.04858M | $1.5 | $9 | 2026-05-19 | |||
| Gemini 3.1 Flash Litegoogle/gemini-3.1-flash-lite | 1.04858M | $0.25 | $1.5 | 2026-05-07 | |||
| GPT-5.5 Instantopenai/gpt-5.5-instant | 400K | $5 | $30 | 2026-05-05 | |||
| Mistral Medium (latest)mistral/mistral-medium-latest | 262.144K | $1.5 | $7.5 | 2026-04-29 | |||
| Mistral Medium 3.5mistral/mistral-medium-2604 | 262.144K | $1.5 | $7.5 | 2026-04-29 | |||
| Laguna M.1poolside/laguna-m.1 | 262.144K | — | — | 2026-04-28 | |||
| Nemotron 3 Nano Omni 30B A3B Reasoningnvidia/nemotron-3-nano-omni-30b-a3b-reasoning | 256K | $0.2 | $0.8 | 2026-04-28 | |||
| Laguna XS.2poolside/laguna-xs.2 | 262.144K | — | — | 2026-04-28 | |||
| Qwen3.6 Flashalibaba/qwen3.6-flash | 1M | $0.188 | $1.125 | 2026-04-27 | |||
| DeepSeek V4 Prodeepseek/deepseek-v4-pro | 1M | $0.435 | $0.87 | 2026-04-24 | |||
| DeepSeek V4 Flashdeepseek/deepseek-v4-flash | 1M | $0.15 | $0.6 | 2026-04-24 | |||
| GPT-5.5 Proopenai/gpt-5.5-pro | 1.05M | $30 | $180 | 2026-04-23 | |||
| DeepSeek V4 Pro 0423deepseek/deepseek-v4-pro-0423 | 1M | $1.32 | $3.96 | 2026-04-23 | |||
| DeepSeek V4 Flash 0423deepseek/deepseek-v4-flash-0423 | 1M | $0.139 | $0.278 | 2026-04-23 | |||
| GPT-5.5openai/gpt-5.5 | 1.05M | $5 | $30 | 2026-04-23 | |||
| MiMo-V2.5xiaomi/mimo-v2.5 | 1.04858M | $0.14 | $0.28 | 2026-04-22 | |||
| MiMo-V2.5-Proxiaomi/mimo-v2.5-pro | 1.04858M | $0.435 | $0.87 | 2026-04-22 | |||
| Qwen3.6 27Balibaba/qwen3.6-27b | 262.144K | $0.6 | $3.6 | 2026-04-22 | |||
| Gemini Deep Research Previewgoogle/deep-research-preview-04-2026 | 1.04858M | — | — | 2026-04-21 |