Video generation and editing model for fast, conversational text- and image-to-video workflows
Models
Every model in the catalog with source-linked pricing, context limits, provider availability, and published benchmark results.
Open coding-reasoning model for repository tasks and self-improving agents
Open coding-reasoning model for repository tasks and self-improving agents
Large coding-reasoning model for agentic software tasks and RL search
Large coding-reasoning model for agentic software tasks and RL search
ByteDance Seed model optimized for character-driven dialogue and consistent conversational behavior
Flagship ByteDance Seed 2.1 model for complex multimodal reasoning, coding, and agents
Rolling ByteDance Seed model for rapidly updated reasoning, coding, and agent capabilities
Faster ByteDance Seed 2.1 model for multimodal reasoning and latency-sensitive agent workflows
Quality-first multi-agent model for hard research, analysis, and competitions
Multi-agent model for routing expert agents across complex analytical tasks
Open flagship GLM for long-horizon coding agents and million-token context work
Coding-focused Kimi model, stronger on long-horizon repo work with less overthinking
Lower-latency Kimi Code variant for interactive edits and coding-agent loops
Claude model for creative writing, analysis, and controlled agent workflows
Restricted Claude model for advanced cybersecurity and biology research workflows
Cohere coding model for practical software engineering and agentic edits
Low-latency audio-to-audio model for real-time speech translation across 70+ languages
MiMo pro model for strong multimodal reasoning and agent execution
Safety model for policy screening, moderation, and risk-aware routing workflows
Largest Nemotron 3 model for maximum open-weight reasoning and agent accuracy
Multimodal Qwen workhorse for long-context agents, visual inputs, and coding
Microsoft coding model built for fast, efficient assistance in everyday developer workflows
MiniMax multimodal model for long-context coding, perception, and agent planning
Newer StepFun flash model for faster agents, coding, and multimodal prompts
Nano Banana Pro for higher-fidelity image generation and design-heavy edits
Top Claude Opus tier for the hardest reasoning, coding, and long-horizon agents
Image model for prompt-driven generation, editing, and visual design workflows
Qwen frontier model tuned for agent frameworks, coding assistants, and long tasks
Cohere's stronger command model for multilingual agents and enterprise workflows
Dense 1B-class open-source model for on-device and resource-constrained use, with native long-context support, Think / No Think chat modes, and tool calling
Fast Gemini model balancing multimodal reasoning, tool use, and cost
Low-latency Gemini model for high-volume multimodal and agent workloads
Compact GPT model for low-latency assistance and high-volume workloads
Balanced Mistral model for enterprise assistants, multilingual work, and tools
Balanced Mistral model for enterprise assistants, multilingual work, and tools
Poolside's open-weight model for agentic coding and long-horizon work
Open Nemotron omni model combining reasoning with text, vision, and audio
Agentic coding model from Poolside in the XS size class for local deployment
Qwen vision-language model for visual reasoning, documents, and agent tasks
Open MoE flagship with million-token context for coding and long agent runs
Fast DeepSeek V4 lane for economical reasoning, coding, and long-context work
Highest-accuracy GPT-5.5 tier for slower, precision-heavy reasoning and coding
Initial DeepSeek V4 Flash snapshot for economical reasoning, coding, and million-token agent workloads
DeepSeek V4 Pro initial snapshot with million-token context and support for thinking and non-thinking modes
Default frontier GPT for coding, computer use, research, and knowledge work
Stronger MiMo Pro tier for multimodal reasoning and coding-agent execution
Qwen vision-language model for visual reasoning, documents, and agent tasks
Multimodal embedding model mapping text, images, video, audio, and PDFs into a unified embedding space
Open MiMo model for multimodal coding agents and long-context automation
| Model | Creator | Inputs | Context | Input | Output | Released | Compare |
|---|---|---|---|---|---|---|---|
| Gemini Omni Flash Previewgoogle/gemini-omni-flash-preview | 1.04858M | $1.5 | $17.5 | 2026-06-30 | |||
| Ornith 1.0 9Bdeepreinforce/ornith-1.0-9b | 262.144K | — | — | 2026-06-25 | |||
| Ornith 1.0 31Bdeepreinforce/ornith-1.0-31b | 262.144K | — | — | 2026-06-25 | |||
| Ornith 1.0 397Bdeepreinforce/ornith-1.0-397b | 262.144K | — | — | 2026-06-25 | |||
| Ornith 1.0 35Bdeepreinforce/ornith-1.0-35b | 262.144K | — | — | 2026-06-25 | |||
| Seed Characterbytedance-seed/seed-character | 256K | $0.119 | $0.297 | 2026-06-23 | |||
| Seed 2.1 Probytedance-seed/seed-2.1-pro | 256K | $0.707 | $3.536 | 2026-06-23 | |||
| Seed Evolvingbytedance-seed/seed-evolving | 256K | $0.884 | $4.42 | 2026-06-23 | |||
| Seed 2.1 Turbobytedance-seed/seed-2.1-turbo | 256K | $0.354 | $1.77 | 2026-06-23 | |||
| Fugu Ultrasakana/fugu-ultra | 1M | $5 | $30 | 2026-06-15 | |||
| Fugusakana/fugu | 1M | — | — | 2026-06-15 | |||
| GLM-5.2zhipuai/glm-5.2 | 1M | $1.4 | $4.4 | 2026-06-13 | |||
| Kimi K2.7 Codemoonshotai/kimi-k2.7-code | 262.144K | $0.95 | $4 | 2026-06-12 | |||
| Kimi K2.7 Code Highspeedmoonshotai/kimi-k2.7-code-highspeed | 262.144K | $1.9 | $8 | 2026-06-12 | |||
| Claude Fable 5anthropic/claude-fable-5 | 1M | $10 | $50 | 2026-06-09 | |||
| Claude Mythos 5anthropic/claude-mythos-5 | 1M | $10 | $50 | 2026-06-09 | |||
| North Mini Codecohere/north-mini-code-1-0 | 256K | — | — | 2026-06-09 | |||
| Gemini 3.5 Live Translate Previewgoogle/gemini-3.5-live-translate-preview | 131.072K | $3.5 | $21 | 2026-06-09 | |||
| MiMo-V2.5-Pro-UltraSpeedxiaomi/mimo-v2.5-pro-ultraspeed | 1.04858M | $1.305 | $2.61 | 2026-06-08 | |||
| Nemotron 3.5 Content Safetynvidia/nemotron-3.5-content-safety | 128K | $0.2 | $0.2 | 2026-06-04 | |||
| Nemotron 3 Ultra 550B A55Bnvidia/nemotron-3-ultra-550b-a55b | 1M | $0.5 | $2.5 | 2026-06-04 | |||
| Qwen3.7 Plusalibaba/qwen3.7-plus | 1M | $0.5 | $3 | 2026-06-02 | |||
| MAI-Code-1-Flashmicrosoft/mai-code-1-flash | 256K | $0.75 | $4.5 | 2026-06-02 | |||
| MiniMax-M3minimax/MiniMax-M3 | 1.04858M | $0.3 | $1.2 | 2026-06-01 | |||
| Step 3.7 Flashstepfun/step-3.7-flash | 256K | $0.185 | $1.11 | 2026-05-29 | |||
| Nano Banana Progoogle/gemini-3-pro-image | 65.536K | $2 | $120 | 2026-05-28 | |||
| Claude Opus 4.8anthropic/claude-opus-4-8 | 1M | $5 | $25 | 2026-05-28 | |||
| Nano Banana 2google/gemini-3.1-flash-image | 131.072K | $0.5 | $60 | 2026-05-28 | |||
| Qwen3.7 Maxalibaba/qwen3.7-max | 1M | $2.5 | $7.5 | 2026-05-21 | |||
| Command A Pluscohere/command-a-plus-05-2026 | 128K | $2.5 | $10 | 2026-05-20 | |||
| MiniCPM5-1Bopenbmb/minicpm5-1b | 131.072K | $0.124 | $0.743 | 2026-05-19 | |||
| Gemini 3.5 Flashgoogle/gemini-3.5-flash | 1.04858M | $1.5 | $9 | 2026-05-19 | |||
| Gemini 3.1 Flash Litegoogle/gemini-3.1-flash-lite | 1.04858M | $0.25 | $1.5 | 2026-05-07 | |||
| GPT-5.5 Instantopenai/gpt-5.5-instant | 400K | $5 | $30 | 2026-05-05 | |||
| Mistral Medium (latest)mistral/mistral-medium-latest | 262.144K | $1.5 | $7.5 | 2026-04-29 | |||
| Mistral Medium 3.5mistral/mistral-medium-2604 | 262.144K | $1.5 | $7.5 | 2026-04-29 | |||
| Laguna M.1poolside/laguna-m.1 | 262.144K | — | — | 2026-04-28 | |||
| Nemotron 3 Nano Omni 30B A3B Reasoningnvidia/nemotron-3-nano-omni-30b-a3b-reasoning | 256K | $0.2 | $0.8 | 2026-04-28 | |||
| Laguna XS.2poolside/laguna-xs.2 | 262.144K | — | — | 2026-04-28 | |||
| Qwen3.6 Flashalibaba/qwen3.6-flash | 1M | $0.188 | $1.125 | 2026-04-27 | |||
| DeepSeek V4 Prodeepseek/deepseek-v4-pro | 1M | $0.435 | $0.87 | 2026-04-24 | |||
| DeepSeek V4 Flashdeepseek/deepseek-v4-flash | 1M | $0.15 | $0.6 | 2026-04-24 | |||
| GPT-5.5 Proopenai/gpt-5.5-pro | 1.05M | $30 | $180 | 2026-04-23 | |||
| DeepSeek V4 Flash 0423deepseek/deepseek-v4-flash-0423 | 1M | $0.139 | $0.278 | 2026-04-23 | |||
| DeepSeek V4 Pro 0423deepseek/deepseek-v4-pro-0423 | 1M | $1.32 | $3.96 | 2026-04-23 | |||
| GPT-5.5openai/gpt-5.5 | 1.05M | $5 | $30 | 2026-04-23 | |||
| MiMo-V2.5-Proxiaomi/mimo-v2.5-pro | 1.04858M | $0.435 | $0.87 | 2026-04-22 | |||
| Qwen3.6 27Balibaba/qwen3.6-27b | 262.144K | $0.6 | $3.6 | 2026-04-22 | |||
| Gemini Embedding 2google/gemini-embedding-2 | 8.192K | $0.2 | — | 2026-04-22 | |||
| MiMo-V2.5xiaomi/mimo-v2.5 | 1.04858M | $0.14 | $0.28 | 2026-04-22 |