Fastest, most cost-efficient Gemini image model for high-volume 1K generation and editing
Models
Every model in the catalog with source-linked pricing, context limits, provider availability, and published benchmark results.
Meituan LongCat-2.0, a reasoning model with tool calling and a 1M-token context window
Video generation and editing model for fast, conversational text- and image-to-video workflows
Large coding-reasoning model for agentic software tasks and RL search
Large coding-reasoning model for agentic software tasks and RL search
Open coding-reasoning model for repository tasks and self-improving agents
Open coding-reasoning model for repository tasks and self-improving agents
ByteDance Seed model optimized for character-driven dialogue and consistent conversational behavior
Flagship ByteDance Seed 2.1 model for complex multimodal reasoning, coding, and agents
Rolling ByteDance Seed model for rapidly updated reasoning, coding, and agent capabilities
Faster ByteDance Seed 2.1 model for multimodal reasoning and latency-sensitive agent workflows
Quality-first multi-agent model for hard research, analysis, and competitions
Multi-agent model for routing expert agents across complex analytical tasks
Open flagship GLM for long-horizon coding agents and million-token context work
Lower-latency Kimi Code variant for interactive edits and coding-agent loops
Coding-focused Kimi model, stronger on long-horizon repo work with less overthinking
Claude model for creative writing, analysis, and controlled agent workflows
Restricted Claude model for advanced cybersecurity and biology research workflows
Cohere coding model for practical software engineering and agentic edits
Low-latency audio-to-audio model for real-time speech translation across 70+ languages
MiMo pro model for strong multimodal reasoning and agent execution
Largest Nemotron 3 model for maximum open-weight reasoning and agent accuracy
Safety model for policy screening, moderation, and risk-aware routing workflows
Microsoft coding model built for fast, efficient assistance in everyday developer workflows
Multimodal Qwen workhorse for long-context agents, visual inputs, and coding
MiniMax multimodal model for long-context coding, perception, and agent planning
Video model for image-to-video generation, editing, and extension workflows
Newer StepFun flash model for faster agents, coding, and multimodal prompts
Top Claude Opus tier for the hardest reasoning, coding, and long-horizon agents
Nano Banana Pro for higher-fidelity image generation and design-heavy edits
Image model for prompt-driven generation, editing, and visual design workflows
Qwen frontier model tuned for agent frameworks, coding assistants, and long tasks
Cohere's stronger command model for multilingual agents and enterprise workflows
Fast Gemini model balancing multimodal reasoning, tool use, and cost
Dense 1B-class open-source model for on-device and resource-constrained use, with native long-context support, Think / No Think chat modes, and tool calling
Streaming speech-to-text model for low-latency transcript deltas from live audio
Low-latency Gemini model for high-volume multimodal and agent workloads
Compact GPT model for low-latency assistance and high-volume workloads
Balanced Mistral model for enterprise assistants, multilingual work, and tools
Balanced Mistral model for enterprise assistants, multilingual work, and tools
Poolside's open-weight model for agentic coding and long-horizon work
Open Nemotron omni model combining reasoning with text, vision, and audio
Agentic coding model from Poolside in the XS size class for local deployment
Qwen vision-language model for visual reasoning, documents, and agent tasks
Fast DeepSeek V4 lane for economical reasoning, coding, and long-context work
Open MoE flagship with million-token context for coding and long agent runs
Default frontier GPT for coding, computer use, research, and knowledge work
Initial DeepSeek V4 Flash snapshot for economical reasoning, coding, and million-token agent workloads
Highest-accuracy GPT-5.5 tier for slower, precision-heavy reasoning and coding
DeepSeek V4 Pro initial snapshot with million-token context and support for thinking and non-thinking modes
| Model | Creator | Inputs | Context | Input | Output | Released | Compare |
|---|---|---|---|---|---|---|---|
| Nano Banana 2 Litegoogle/gemini-3.1-flash-lite-image | 65.536K | $0.25 | $30 | 2026-06-30 | |||
| LongCat-2.0meituan/longcat-2.0 | 1M | $0.3 | $1.2 | 2026-06-30 | |||
| Gemini Omni Flash Previewgoogle/gemini-omni-flash-preview | 1.04858M | $1.5 | $17.5 | 2026-06-30 | |||
| Ornith 1.0 397Bdeepreinforce/ornith-1.0-397b | 262.144K | — | — | 2026-06-25 | |||
| Ornith 1.0 35Bdeepreinforce/ornith-1.0-35b | 262.144K | — | — | 2026-06-25 | |||
| Ornith 1.0 9Bdeepreinforce/ornith-1.0-9b | 262.144K | — | — | 2026-06-25 | |||
| Ornith 1.0 31Bdeepreinforce/ornith-1.0-31b | 262.144K | — | — | 2026-06-25 | |||
| Seed Characterbytedance-seed/seed-character | 256K | $0.119 | $0.297 | 2026-06-23 | |||
| Seed 2.1 Probytedance-seed/seed-2.1-pro | 256K | $0.707 | $3.536 | 2026-06-23 | |||
| Seed Evolvingbytedance-seed/seed-evolving | 256K | $0.884 | $4.42 | 2026-06-23 | |||
| Seed 2.1 Turbobytedance-seed/seed-2.1-turbo | 256K | $0.354 | $1.77 | 2026-06-23 | |||
| Fugu Ultrasakana/fugu-ultra | 1M | $5 | $30 | 2026-06-15 | |||
| Fugusakana/fugu | 1M | — | — | 2026-06-15 | |||
| GLM-5.2zhipuai/glm-5.2 | 1M | $1.4 | $4.4 | 2026-06-13 | |||
| Kimi K2.7 Code Highspeedmoonshotai/kimi-k2.7-code-highspeed | 262.144K | $1.9 | $8 | 2026-06-12 | |||
| Kimi K2.7 Codemoonshotai/kimi-k2.7-code | 262.144K | $0.95 | $4 | 2026-06-12 | |||
| Claude Fable 5anthropic/claude-fable-5 | 1M | $10 | $50 | 2026-06-09 | |||
| Claude Mythos 5anthropic/claude-mythos-5 | 1M | $10 | $50 | 2026-06-09 | |||
| North Mini Codecohere/north-mini-code-1-0 | 256K | — | — | 2026-06-09 | |||
| Gemini 3.5 Live Translate Previewgoogle/gemini-3.5-live-translate-preview | 131.072K | $3.5 | $21 | 2026-06-09 | |||
| MiMo-V2.5-Pro-UltraSpeedxiaomi/mimo-v2.5-pro-ultraspeed | 1.04858M | $1.305 | $2.61 | 2026-06-08 | |||
| Nemotron 3 Ultra 550B A55Bnvidia/nemotron-3-ultra-550b-a55b | 1M | $0.5 | $2.5 | 2026-06-04 | |||
| Nemotron 3.5 Content Safetynvidia/nemotron-3.5-content-safety | 128K | $0.2 | $0.2 | 2026-06-04 | |||
| MAI-Code-1-Flashmicrosoft/mai-code-1-flash | 256K | $0.75 | $4.5 | 2026-06-02 | |||
| Qwen3.7 Plusalibaba/qwen3.7-plus | 1M | $0.5 | $3 | 2026-06-02 | |||
| MiniMax-M3minimax/MiniMax-M3 | 1.04858M | $0.3 | $1.2 | 2026-06-01 | |||
| Grok Imagine Video 1.5xai/grok-imagine-video-1.5 | 1.024K | — | — | 2026-05-30 | |||
| Step 3.7 Flashstepfun/step-3.7-flash | 256K | $0.185 | $1.11 | 2026-05-29 | |||
| Claude Opus 4.8anthropic/claude-opus-4-8 | 1M | $5 | $25 | 2026-05-28 | |||
| Nano Banana Progoogle/gemini-3-pro-image | 65.536K | $2 | $120 | 2026-05-28 | |||
| Nano Banana 2google/gemini-3.1-flash-image | 131.072K | $0.5 | $60 | 2026-05-28 | |||
| Qwen3.7 Maxalibaba/qwen3.7-max | 1M | $2.5 | $7.5 | 2026-05-21 | |||
| Command A Pluscohere/command-a-plus-05-2026 | 128K | $2.5 | $10 | 2026-05-20 | |||
| Gemini 3.5 Flashgoogle/gemini-3.5-flash | 1.04858M | $1.5 | $9 | 2026-05-19 | |||
| MiniCPM5-1Bopenbmb/minicpm5-1b | 131.072K | $0.124 | $0.743 | 2026-05-19 | |||
| GPT Realtime Whisperopenai/gpt-realtime-whisper | Not documented | — | — | 2026-05-07 | |||
| Gemini 3.1 Flash Litegoogle/gemini-3.1-flash-lite | 1.04858M | $0.25 | $1.5 | 2026-05-07 | |||
| GPT-5.5 Instantopenai/gpt-5.5-instant | 400K | $5 | $30 | 2026-05-05 | |||
| Mistral Medium 3.5mistral/mistral-medium-2604 | 262.144K | $1.5 | $7.5 | 2026-04-29 | |||
| Mistral Medium (latest)mistral/mistral-medium-latest | 262.144K | $1.5 | $7.5 | 2026-04-29 | |||
| Laguna M.1poolside/laguna-m.1 | 262.144K | — | — | 2026-04-28 | |||
| Nemotron 3 Nano Omni 30B A3B Reasoningnvidia/nemotron-3-nano-omni-30b-a3b-reasoning | 256K | $0.2 | $0.8 | 2026-04-28 | |||
| Laguna XS.2poolside/laguna-xs.2 | 262.144K | — | — | 2026-04-28 | |||
| Qwen3.6 Flashalibaba/qwen3.6-flash | 1M | $0.188 | $1.125 | 2026-04-27 | |||
| DeepSeek V4 Flashdeepseek/deepseek-v4-flash | 1M | $0.15 | $0.6 | 2026-04-24 | |||
| DeepSeek V4 Prodeepseek/deepseek-v4-pro | 1M | $0.435 | $0.87 | 2026-04-24 | |||
| GPT-5.5openai/gpt-5.5 | 1.05M | $5 | $30 | 2026-04-23 | |||
| DeepSeek V4 Flash 0423deepseek/deepseek-v4-flash-0423 | 1M | $0.139 | $0.278 | 2026-04-23 | |||
| GPT-5.5 Proopenai/gpt-5.5-pro | 1.05M | $30 | $180 | 2026-04-23 | |||
| DeepSeek V4 Pro 0423deepseek/deepseek-v4-pro-0423 | 1M | $1.32 | $3.96 | 2026-04-23 |