Large open Qwen multimodal MoE for visual agents and long technical tasks
Models
Every model in the catalog with source-linked pricing, context limits, provider availability, and published benchmark results.
Seed-2.0-mini targets latency-sensitive, high-concurrency, and cost-sensitive scenarios, emphasizing fast response and flexible inference deployment. It delivers performance comparable to ByteDance-Seed-1.6, supports 256k context, four reasoning effort modes (minimal/low/medium/high), multimodal understanding,...
Seed 2.0 Code is a model from ByteDance Seed optimized for agentic coding. It is suited for frontend development, multilingual programming tasks, and coding-agent workflows in tools such as Claude...
Seed-2.0-Lite is a versatile, cost‑efficient enterprise workhorse that delivers strong multimodal and agent capabilities while offering noticeably lower latency, making it a practical default choice for most production workloads across...
Prior MiniMax coding model for agent workflows, office edits, and automation
General GLM flagship for coding, analysis, and tool-heavy engineering workflows
GPT-5.3-Codex is OpenAI’s most advanced agentic coding model, combining the frontier software engineering performance of GPT-5.2-Codex with the broader reasoning and professional knowledge capabilities of GPT-5.2. It achieves state-of-the-art results...
High-end Claude for difficult coding, planning, and slower expert reasoning
Open-weight Qwen coding model for agents, repository edits, and multi-turn tool use
Step 3.5 Flash is StepFun's most capable open-source foundation model. Built on a sparse Mixture of Experts (MoE) architecture, it selectively activates only 11B of its 196B parameters per token....
MiniMax M2 variant tuned for conversational and character-driven agent interactions
Budget GLM lane for fast coding help, routing, and everyday automation
Flagship model for demanding analysis, coding, and production agent workflows
Kimi K2.5 is Moonshot AI's native multimodal model, delivering state-of-the-art visual coding capability and a self-directed agent swarm paradigm. Built on Kimi K2 with continued pretraining over approximately 15T mixed...
Earlier MiniMax agent model for practical coding and productivity tasks
Mature GLM model for dependable coding, reasoning, and structured agent tasks
Gemini 3 Flash Preview is a high speed, high value thinking model designed for agentic workflows, multi turn chat, and coding assistance. It delivers near Pro level reasoning and tool...
NVIDIA Nemotron 3 Nano 30B A3B is a small language MoE model with highest compute efficiency and accuracy for developers to build specialized agentic AI systems. The model is fully...
GPT-5.2-Codex is an upgraded version of GPT-5.1-Codex optimized for software engineering and coding workflows. It is designed for both interactive development sessions and long, independent execution of complex engineering tasks....
GPT-5.2 is the latest frontier-grade model in the GPT-5 series, offering stronger agentic and long context perfomance compared to GPT-5.1. It uses adaptive reasoning to allocate computation dynamically, responding quickly...
Chat-tuned GPT model for conversational assistance, writing, and tool workflows
GPT-5.2 Pro is OpenAI’s most advanced model, offering major improvements in agentic coding and long context performance over GPT-5 Pro. It is optimized for complex tasks that require step-by-step reasoning,...
Mistral's coding-agent model for repository work, terminal tasks, and software fixes
GLM vision model for visual reasoning, documents, and multimodal agents
Multimodal reasoning model for visual analysis, planning, and tool use
DeepSeek-V3.2 is a large language model designed to harmonize high computational efficiency with strong reasoning and agentic tool-use performance. It introduces DeepSeek Sparse Attention (DSA), a fine-grained sparse attention mechanism...
DeepSeek-V3 is the latest model from the DeepSeek team, building upon the instruction following and coding abilities of the previous versions. Pre-trained on nearly 15 trillion tokens, the reported evaluations...
Flagship Claude model for deep reasoning, coding, and long-horizon agents
Nano Banana Pro is Google’s most advanced image-generation and editing model, built on Gemini 3 Pro. It extends the original Nano Banana with significantly improved multimodal reasoning, real-world grounding, and...
GPT-5.1 is the latest frontier-grade model in the GPT-5 series, offering stronger general-purpose reasoning, improved instruction adherence, and a more natural conversational style compared to GPT-5. It uses adaptive reasoning...
GPT-5.1-Codex is a specialized version of GPT-5.1 optimized for software engineering and coding workflows. It is designed for both interactive development sessions and long, independent execution of complex engineering tasks....
GPT-5.1-Codex-Mini is a smaller and faster version of GPT-5.1-Codex
Chat-tuned GPT-5.1 for polished assistants, writing, and product conversations
GPT-5.1-Codex-Max is OpenAI’s latest agentic coding model, designed for long-running, high-context software development tasks. It is based on an updated version of the 5.1 reasoning stack and trained on agentic...
Kimi K2 Thinking is Moonshot AI’s most advanced open reasoning model to date, extending the K2 series into agentic, long-horizon reasoning. Built on the trillion-parameter Mixture-of-Experts (MoE) architecture introduced in...
gpt-oss-safeguard-20b is a safety reasoning model from OpenAI built upon gpt-oss-20b. This open-weight, 21B-parameter Mixture-of-Experts (MoE) model offers lower latency for safety tasks like content classification, LLM filtering, and trust...
Efficient open MiniMax model built for coding agents and tool-heavy workflows
Fast Claude lane for lightweight agents, office tasks, and responsive chat
ByteDance Seed model for long-context reasoning, instruction following, and tool-assisted tasks
GPT-5 Pro is OpenAI’s most advanced model, offering major improvements in reasoning, code quality, and user experience. It is optimized for complex tasks that require step-by-step reasoning, instruction following, and...
Late GLM-4 workhorse for coding agents, reasoning, and structured tasks
Balanced Claude model for coding, analysis, agent workflows, and cost control
Qwen vision-language instruct model for visual reasoning, documents, and agent tasks
Qwen vision-language thinking model for visual reasoning, documents, and agent tasks
Flagship Qwen3 model for coding agents, complex reasoning, and tool use
Coding-optimized GPT model for repository edits, reviews, and agentic software work
Qwen instruction model for multilingual chat, reasoning, and tool use
Efficient Qwen thinking model for local reasoning, math, and coding agents
Low-latency ByteDance Seed model for high-throughput chat, extraction, and lightweight tool use
Gemini 2.5 Flash Image, a.k.a. "Nano Banana," is now generally available. It is a state of the art image generation model with contextual understanding. It is capable of image generation,...
| Model | Creator | Inputs | Context | Input | Output | Released | Compare |
|---|---|---|---|---|---|---|---|
| Qwen3.5 397B-A17Balibaba/qwen3.5-397b-a17b | 262.144K | $0.6 | $3.6 | 2026-02-15 | |||
| ByteDance Seed: Seed-2.0-Minibytedance-seed/seed-2.0-mini | 262.144K | $0.1 | $0.4 | 2026-02-14 | |||
| ByteDance Seed: Seed-2.0-Codebytedance-seed/seed-2.0-code | 262.144K | $0.5 | $3 | 2026-02-14 | |||
| ByteDance Seed: Seed-2.0-Litebytedance-seed/seed-2.0-lite | 262.144K | $0.25 | $2 | 2026-02-14 | |||
| MiniMax-M2.5minimax/MiniMax-M2.5 | 204.8K | $0.3 | $1.2 | 2026-02-12 | |||
| GLM-5zhipuai/glm-5 | 204.8K | $1 | $3.2 | 2026-02-12 | |||
| OpenAI: GPT-5.3-Codexopenai/gpt-5.3-codex | 400K | $1.75 | $14 | 2026-02-05 | |||
| Claude Opus 4.6anthropic/claude-opus-4-6 | 1M | $5 | $25 | 2026-02-05 | |||
| Qwen3 Coder Nextalibaba/qwen3-coder-next | 262.144K | $0.108 | $0.675 | 2026-02-03 | |||
| StepFun: Step 3.5 Flashstepfun/step-3.5-flash | 262.144K | $0.1 | $0.3 | 2026-01-29 | |||
| MiniMax-M2 Herminimax/MiniMax-M2-Her | 65.536K | $0.3 | $1.2 | 2026-01-23 | |||
| GLM-4.7-Flashzhipuai/glm-4.7-flash | 200K | $0.06 | $0.4 | 2026-01-19 | |||
| Solar Pro 3upstage/solar-pro3 | 131.072K | $0.15 | $0.6 | 2026-01 | |||
| MoonshotAI: Kimi K2.5moonshotai/kimi-k2.5 | 262.144K | $0.45 | $2.25 | 2026-01 | |||
| MiniMax-M2.1minimax/MiniMax-M2.1 | 204.8K | $0.3 | $1.2 | 2025-12-23 | |||
| GLM-4.7zhipuai/glm-4.7 | 204.8K | $0.6 | $2.2 | 2025-12-22 | |||
| Google: Gemini 3 Flash Previewgoogle/gemini-3-flash-preview | 1.04858M | $0.5 | $3 | 2025-12-17 | |||
| NVIDIA: Nemotron 3 Nano 30B A3Bnvidia/nemotron-3-nano-30b-a3b | 262.144K | $0.05 | $0.2 | 2025-12-15 | |||
| OpenAI: GPT-5.2-Codexopenai/gpt-5.2-codex | 400K | $1.75 | $14 | 2025-12-11 | |||
| OpenAI: GPT-5.2openai/gpt-5.2 | 400K | $1.75 | $14 | 2025-12-11 | |||
| GPT-5.2 Chatopenai/gpt-5.2-chat-latest | 128K | $1.75 | $14 | 2025-12-11 | |||
| OpenAI: GPT-5.2 Proopenai/gpt-5.2-pro | 400K | $21 | $168 | 2025-12-11 | |||
| Devstral 2mistral/devstral-2512 | 262.144K | $0.4 | $2 | 2025-12-09 | |||
| GLM-4.6Vzhipuai/glm-4.6v | 128K | $0.3 | $0.9 | 2025-12-08 | |||
| Nova 2 Liteamazon/nova-2-lite | 1M | $0.3 | $2.5 | 2025-12-02 | |||
| DeepSeek: DeepSeek V3.2deepseek/deepseek-v3.2 | 163.84K | $0.269 | $0.4 | 2025-12-01 | |||
| DeepSeek: DeepSeek V3deepseek/deepseek-chat | 128K | $0.257 | $1.029 | 2025-12-01 | |||
| Claude Opus 4.5 (latest)anthropic/claude-opus-4-5 | 200K | $5 | $25 | 2025-11-24 | |||
| Google: Nano Banana Pro (Gemini 3 Pro Image Preview)google/gemini-3-pro-image-preview | 65.536K | $2 | $12 | 2025-11-20 | |||
| OpenAI: GPT-5.1openai/gpt-5.1 | 400K | $1.25 | $10 | 2025-11-13 | |||
| OpenAI: GPT-5.1-Codexopenai/gpt-5.1-codex | 400K | $1.25 | $10 | 2025-11-13 | |||
| OpenAI: GPT-5.1-Codex-Miniopenai/gpt-5.1-codex-mini | 400K | $0.25 | $2 | 2025-11-13 | |||
| GPT-5.1 Chatopenai/gpt-5.1-chat-latest | 128K | $1.25 | $10 | 2025-11-13 | |||
| OpenAI: GPT-5.1-Codex-Maxopenai/gpt-5.1-codex-max | 400K | $1.25 | $10 | 2025-11-13 | |||
| MoonshotAI: Kimi K2 Thinkingmoonshotai/kimi-k2-thinking | 262.144K | $0.6 | $2.5 | 2025-11-06 | |||
| OpenAI: gpt-oss-safeguard-20bopenai/gpt-oss-safeguard-20b | 131.072K | $0.075 | $0.3 | 2025-10-29 | |||
| MiniMax-M2minimax/MiniMax-M2 | 204.8K | $0.3 | $1.2 | 2025-10-27 | |||
| Claude Haiku 4.5 (latest)anthropic/claude-haiku-4-5 | 200K | $1 | $5 | 2025-10-15 | |||
| Seed 1.6bytedance-seed/seed-1-6 | 256K | $0.119 | $1.187 | 2025-10-15 | |||
| OpenAI: GPT-5 Proopenai/gpt-5-pro | 400K | $15 | $120 | 2025-10-06 | |||
| GLM-4.6zhipuai/glm-4.6 | 204.8K | $0.6 | $2.2 | 2025-09-30 | |||
| Claude Sonnet 4.5 (latest)anthropic/claude-sonnet-4-5 | 200K | $3 | $15 | 2025-09-29 | |||
| Qwen3 VL 235B A22B Instructalibaba/qwen3-vl-235b-a22b-instruct | 131.072K | $0.2 | $0.88 | 2025-09-23 | |||
| Qwen3 VL 235B A22B Thinkingalibaba/qwen3-vl-235b-a22b-thinking | 131.072K | $0.4 | $4 | 2025-09-23 | |||
| Qwen3 Maxalibaba/qwen3-max | 262.144K | $1.2 | $6 | 2025-09-23 | |||
| GPT-5-Codexopenai/gpt-5-codex | 400K | $1.1 | $9 | 2025-09-15 | |||
| Qwen3-Next 80B-A3B Instructalibaba/qwen3-next-80b-a3b-instruct | 131.072K | $0.5 | $2 | 2025-09 | |||
| Qwen3-Next 80B-A3B (Thinking)alibaba/qwen3-next-80b-a3b-thinking | 131.072K | $0.5 | $6 | 2025-09 | |||
| Seed 1.6 Flashbytedance-seed/seed-1-6-flash | 256K | $0.022 | $0.223 | 2025-08-28 | |||
| Google: Nano Banana (Gemini 2.5 Flash Image)google/gemini-2.5-flash-image | 32.768K | $0.3 | $2.5 | 2025-08-26 |