Faster ByteDance Seed 2.1 model for multimodal reasoning and latency-sensitive agent workflows
Models
Every model in the catalog with source-linked pricing, context limits, provider availability, and published benchmark results.
Rolling ByteDance Seed model for rapidly updated reasoning, coding, and agent capabilities
Flagship ByteDance Seed 2.1 model for complex multimodal reasoning, coding, and agents
ByteDance Seed model optimized for character-driven dialogue and consistent conversational behavior
Multi-agent model for routing expert agents across complex analytical tasks
Quality-first multi-agent model for hard research, analysis, and competitions
Open flagship GLM for long-horizon coding agents and million-token context work
Lower-latency Kimi Code variant for interactive edits and coding-agent loops
Coding-focused Kimi model, stronger on long-horizon repo work with less overthinking
Cohere coding model for practical software engineering and agentic edits
Restricted Claude model for advanced cybersecurity and biology research workflows
Claude model for creative writing, analysis, and controlled agent workflows
MiMo pro model for strong multimodal reasoning and agent execution
Largest Nemotron 3 model for maximum open-weight reasoning and agent accuracy
Microsoft coding model built for fast, efficient assistance in everyday developer workflows
Multimodal Qwen workhorse for long-context agents, visual inputs, and coding
MiniMax multimodal model for long-context coding, perception, and agent planning
Newer StepFun flash model for faster agents, coding, and multimodal prompts
Top Claude Opus tier for the hardest reasoning, coding, and long-horizon agents
Nano Banana Pro for higher-fidelity image generation and design-heavy edits
Image model for prompt-driven generation, editing, and visual design workflows
Qwen frontier model tuned for agent frameworks, coding assistants, and long tasks
Cohere's stronger command model for multilingual agents and enterprise workflows
Fast Gemini model balancing multimodal reasoning, tool use, and cost
Dense 1B-class open-source model for on-device and resource-constrained use, with native long-context support, Think / No Think chat modes, and tool calling
Low-latency Gemini model for high-volume multimodal and agent workloads
Compact GPT model for low-latency assistance and high-volume workloads
Balanced Mistral model for enterprise assistants, multilingual work, and tools
Balanced Mistral model for enterprise assistants, multilingual work, and tools
Poolside's open-weight model for agentic coding and long-horizon work
Open Nemotron omni model combining reasoning with text, vision, and audio
Agentic coding model from Poolside in the XS size class for local deployment
Qwen vision-language model for visual reasoning, documents, and agent tasks
Fast DeepSeek V4 lane for economical reasoning, coding, and long-context work
Open MoE flagship with million-token context for coding and long agent runs
Highest-accuracy GPT-5.5 tier for slower, precision-heavy reasoning and coding
Default frontier GPT for coding, computer use, research, and knowledge work
DeepSeek V4 Pro initial snapshot with million-token context and support for thinking and non-thinking modes
Initial DeepSeek V4 Flash snapshot for economical reasoning, coding, and million-token agent workloads
Stronger MiMo Pro tier for multimodal reasoning and coding-agent execution
Qwen vision-language model for visual reasoning, documents, and agent tasks
Open MiMo model for multimodal coding agents and long-context automation
Maximum-comprehensiveness agentic researcher for multi-step investigation, synthesis, and cited reports
Agentic model for autonomous multi-step research, synthesis, and cited reports
Multimodal Kimi workhorse for agent loops, coding tasks, and visual context
Tencent Hy reasoning model for coding, instruction following, and agent tasks
Flagship Qwen model for complex reasoning, coding, and agentic workflows
Open multimodal Qwen MoE for local agents that need vision, audio, and code
xAI's default Grok for chat, coding, agentic tools, and lower hallucination risk
Fast Grok coding model tuned for agentic engineering and iterative edits
| Model | Creator | Inputs | Context | Input | Output | Released | Compare |
|---|---|---|---|---|---|---|---|
| Seed 2.1 Turbobytedance-seed/seed-2.1-turbo | 256K | $0.354 | $1.77 | 2026-06-23 | |||
| Seed Evolvingbytedance-seed/seed-evolving | 256K | $0.884 | $4.42 | 2026-06-23 | |||
| Seed 2.1 Probytedance-seed/seed-2.1-pro | 256K | $0.707 | $3.536 | 2026-06-23 | |||
| Seed Characterbytedance-seed/seed-character | 256K | $0.119 | $0.297 | 2026-06-23 | |||
| Fugusakana/fugu | 1M | — | — | 2026-06-15 | |||
| Fugu Ultrasakana/fugu-ultra | 1M | $5 | $30 | 2026-06-15 | |||
| GLM-5.2zhipuai/glm-5.2 | 1M | $1.4 | $4.4 | 2026-06-13 | |||
| Kimi K2.7 Code Highspeedmoonshotai/kimi-k2.7-code-highspeed | 262.144K | $1.9 | $8 | 2026-06-12 | |||
| Kimi K2.7 Codemoonshotai/kimi-k2.7-code | 262.144K | $0.95 | $4 | 2026-06-12 | |||
| North Mini Codecohere/north-mini-code-1-0 | 256K | — | — | 2026-06-09 | |||
| Claude Mythos 5anthropic/claude-mythos-5 | 1M | $10 | $50 | 2026-06-09 | |||
| Claude Fable 5anthropic/claude-fable-5 | 1M | $10 | $50 | 2026-06-09 | |||
| MiMo-V2.5-Pro-UltraSpeedxiaomi/mimo-v2.5-pro-ultraspeed | 1.04858M | $1.305 | $2.61 | 2026-06-08 | |||
| Nemotron 3 Ultra 550B A55Bnvidia/nemotron-3-ultra-550b-a55b | 1M | $0.5 | $2.5 | 2026-06-04 | |||
| MAI-Code-1-Flashmicrosoft/mai-code-1-flash | 256K | $0.75 | $4.5 | 2026-06-02 | |||
| Qwen3.7 Plusalibaba/qwen3.7-plus | 1M | $0.5 | $3 | 2026-06-02 | |||
| MiniMax-M3minimax/MiniMax-M3 | 1.04858M | $0.3 | $1.2 | 2026-06-01 | |||
| Step 3.7 Flashstepfun/step-3.7-flash | 256K | $0.185 | $1.11 | 2026-05-29 | |||
| Claude Opus 4.8anthropic/claude-opus-4-8 | 1M | $5 | $25 | 2026-05-28 | |||
| Nano Banana Progoogle/gemini-3-pro-image | 65.536K | $2 | $120 | 2026-05-28 | |||
| Nano Banana 2google/gemini-3.1-flash-image | 131.072K | $0.5 | $60 | 2026-05-28 | |||
| Qwen3.7 Maxalibaba/qwen3.7-max | 1M | $2.5 | $7.5 | 2026-05-21 | |||
| Command A Pluscohere/command-a-plus-05-2026 | 128K | $2.5 | $10 | 2026-05-20 | |||
| Gemini 3.5 Flashgoogle/gemini-3.5-flash | 1.04858M | $1.5 | $9 | 2026-05-19 | |||
| MiniCPM5-1Bopenbmb/minicpm5-1b | 131.072K | $0.124 | $0.743 | 2026-05-19 | |||
| Gemini 3.1 Flash Litegoogle/gemini-3.1-flash-lite | 1.04858M | $0.25 | $1.5 | 2026-05-07 | |||
| GPT-5.5 Instantopenai/gpt-5.5-instant | 400K | $5 | $30 | 2026-05-05 | |||
| Mistral Medium 3.5mistral/mistral-medium-2604 | 262.144K | $1.5 | $7.5 | 2026-04-29 | |||
| Mistral Medium (latest)mistral/mistral-medium-latest | 262.144K | $1.5 | $7.5 | 2026-04-29 | |||
| Laguna M.1poolside/laguna-m.1 | 262.144K | — | — | 2026-04-28 | |||
| Nemotron 3 Nano Omni 30B A3B Reasoningnvidia/nemotron-3-nano-omni-30b-a3b-reasoning | 256K | $0.2 | $0.8 | 2026-04-28 | |||
| Laguna XS.2poolside/laguna-xs.2 | 262.144K | — | — | 2026-04-28 | |||
| Qwen3.6 Flashalibaba/qwen3.6-flash | 1M | $0.188 | $1.125 | 2026-04-27 | |||
| DeepSeek V4 Flashdeepseek/deepseek-v4-flash | 1M | $0.15 | $0.6 | 2026-04-24 | |||
| DeepSeek V4 Prodeepseek/deepseek-v4-pro | 1M | $0.435 | $0.87 | 2026-04-24 | |||
| GPT-5.5 Proopenai/gpt-5.5-pro | 1.05M | $30 | $180 | 2026-04-23 | |||
| GPT-5.5openai/gpt-5.5 | 1.05M | $5 | $30 | 2026-04-23 | |||
| DeepSeek V4 Pro 0423deepseek/deepseek-v4-pro-0423 | 1M | $1.32 | $3.96 | 2026-04-23 | |||
| DeepSeek V4 Flash 0423deepseek/deepseek-v4-flash-0423 | 1M | $0.139 | $0.278 | 2026-04-23 | |||
| MiMo-V2.5-Proxiaomi/mimo-v2.5-pro | 1.04858M | $0.435 | $0.87 | 2026-04-22 | |||
| Qwen3.6 27Balibaba/qwen3.6-27b | 262.144K | $0.6 | $3.6 | 2026-04-22 | |||
| MiMo-V2.5xiaomi/mimo-v2.5 | 1.04858M | $0.14 | $0.28 | 2026-04-22 | |||
| Deep Research Max Previewgoogle/deep-research-max-preview-04-2026 | 1.04858M | — | — | 2026-04-21 | |||
| Gemini Deep Research Previewgoogle/deep-research-preview-04-2026 | 1.04858M | — | — | 2026-04-21 | |||
| Kimi K2.6moonshotai/kimi-k2.6 | 262.144K | $0.95 | $4 | 2026-04-21 | |||
| Hy3 previewtencent/hy3-preview | 256K | $0.066 | $0.26 | 2026-04-20 | |||
| Qwen3.6 Max Previewalibaba/qwen3.6-max-preview | 262.144K | $1.3 | $7.8 | 2026-04-20 | |||
| Qwen3.6 35B-A3Balibaba/qwen3.6-35b-a3b | 262.144K | $0.248 | $1.485 | 2026-04-17 | |||
| Grok 4.3xai/grok-4.3 | 1M | $1.25 | $2.5 | 2026-04-17 | |||
| Grok Build 0.1xai/grok-build-0.1 | 256K | $1 | $2 | 2026-04-16 |