Earlier MiniMax agent model for practical coding and productivity tasks
Models
Every model in the catalog with source-linked pricing, context limits, provider availability, and published benchmark results.
Mature GLM model for dependable coding, reasoning, and structured agent tasks
Small Nemotron 3 MoE for efficient coding, math, and long-context agents
Mistral's coding-agent model for repository work, terminal tasks, and software fixes
GLM vision model for visual reasoning, documents, and multimodal agents
DeepSeek chat model for instruction following, coding, and analysis
Hybrid-reasoning DeepSeek model with thinking and non-thinking modes, sparse attention, and tool-use
Thinking Kimi model for slower research passes, planning, and hard technical questions
Safety model for policy screening, moderation, and risk-aware routing workflows
Efficient open MiniMax model built for coding agents and tool-heavy workflows
Late GLM-4 workhorse for coding agents, reasoning, and structured tasks
Qwen vision-language instruct model for visual reasoning, documents, and agent tasks
Qwen vision-language thinking model for visual reasoning, documents, and agent tasks
Qwen instruction model for multilingual chat, reasoning, and tool use
Efficient Qwen thinking model for local reasoning, math, and coding agents
GLM vision model for visual reasoning, documents, and multimodal agents
Open GPT reasoning model for self-hosted agents and controllable deployments
Open GPT reasoning model for self-hosted agents and controllable deployments
Lighter GLM-4.5 variant for fast coding assistance and cheaper agents
Hybrid-reasoning GLM release that made the 4.5 line broadly useful
Nemotron model for efficient reasoning, coding, and specialized AI agents
Sparse MoE Qwen model with 3B active parameters for efficient chat and reasoning
Dense open Qwen model for self-hosted chat, reasoning, and coding
Smaller Qwen coder for efficient local agents and repo-level fixes
Large open Qwen MoE for multilingual reasoning, coding, and tool use
March 2025 checkpoint of DeepSeek-V3 with improved reasoning and coding
Efficient multimodal model for instruction following, coding, reasoning, and function calling
Cohere command model for multilingual enterprise agents, tools, and chat
Open multimodal Gemma instruction model for efficient text generation and image understanding
Open multimodal Gemma instruction model for multilingual text generation and image understanding
Largest open Gemma 3 instruction model for multilingual text generation and visual understanding
Classic open reasoning model for transparent math, coding, and deliberate problem solving
Popular open Llama workhorse for multilingual chat, coding, and self-hosting
Cohere retrieval model for long-context chat and enterprise RAG workflows
Mistral's largest general model for enterprise agents, coding, and multilingual reasoning
Cohere's RAG workhorse for long-context enterprise search and tool use
Cohere retrieval model for long-context chat and enterprise RAG workflows
Compact open Llama model for lightweight chat, drafting, and self-hosting
Open Llama instruction model for multilingual chat, reasoning, and coding
Efficient Mistral-NVIDIA open model for multilingual chat and local deployment
| Model | Creator | Inputs | Context | Input | Output | Released | Compare |
|---|---|---|---|---|---|---|---|
| MiniMax-M2.1minimax/MiniMax-M2.1 | 204.8K | $0.3 | $1.2 | 2025-12-23 | |||
| GLM-4.7zhipuai/glm-4.7 | 204.8K | $0.6 | $2.2 | 2025-12-22 | |||
| Nemotron 3 Nano 30B A3Bnvidia/nemotron-3-nano-30b-a3b | 262.144K | $0.05 | $0.2 | 2025-12-15 | |||
| Devstral 2mistral/devstral-2512 | 262.144K | $0.4 | $2 | 2025-12-09 | |||
| GLM-4.6Vzhipuai/glm-4.6v | 128K | $0.3 | $0.9 | 2025-12-08 | |||
| DeepSeek Chatdeepseek/deepseek-chat | 1M | $0.147 | $0.295 | 2025-12-01 | |||
| DeepSeek V3.2deepseek/deepseek-v3.2 | 128K | $0.18 | $0.35 | 2025-12-01 | |||
| Kimi K2 Thinkingmoonshotai/kimi-k2-thinking | 262.144K | $0.4 | $2.5 | 2025-11-06 | |||
| GPT OSS Safeguard 20Bopenai/gpt-oss-safeguard-20b | 131.072K | $0.07 | $0.2 | 2025-10-29 | |||
| MiniMax-M2minimax/MiniMax-M2 | 204.8K | $0.3 | $1.2 | 2025-10-27 | |||
| GLM-4.6zhipuai/glm-4.6 | 204.8K | $0.6 | $2.2 | 2025-09-30 | |||
| Qwen3 VL 235B A22B Instructalibaba/qwen3-vl-235b-a22b-instruct | 131.072K | $0.2 | $0.88 | 2025-09-23 | |||
| Qwen3 VL 235B A22B Thinkingalibaba/qwen3-vl-235b-a22b-thinking | 131.072K | $0.4 | $4 | 2025-09-23 | |||
| Qwen3-Next 80B-A3B Instructalibaba/qwen3-next-80b-a3b-instruct | 131.072K | $0.5 | $2 | 2025-09 | |||
| Qwen3-Next 80B-A3B (Thinking)alibaba/qwen3-next-80b-a3b-thinking | 131.072K | $0.5 | $6 | 2025-09 | |||
| GLM-4.5Vzhipuai/glm-4.5v | 64K | $0.6 | $1.8 | 2025-08-11 | |||
| GPT OSS 20Bopenai/gpt-oss-20b | 131.072K | $0.02 | $0.1 | 2025-08-05 | |||
| GPT OSS 120Bopenai/gpt-oss-120b | 131.072K | $0.03 | $0.17 | 2025-08-05 | |||
| GLM-4.5-Airzhipuai/glm-4.5-air | 131.072K | $0.2 | $1.1 | 2025-07-28 | |||
| GLM-4.5zhipuai/glm-4.5 | 131.072K | $0.6 | $2.2 | 2025-07-28 | |||
| Llama 3.3 Nemotron Super 49B v1.5nvidia/llama-3.3-nemotron-super-49b-v1.5 | 131.072K | $0.4 | $0.4 | 2025-07-25 | |||
| Qwen3 30B A3Balibaba/qwen3-30b-a3b | 131.072K | $0.08 | $0.29 | 2025-04-28 | |||
| Qwen3 32Balibaba/qwen3-32b | 131.072K | $0.7 | $2.8 | 2025-04 | |||
| Qwen3-Coder 30B-A3B Instructalibaba/qwen3-coder-30b-a3b-instruct | 262.144K | $0.45 | $2.25 | 2025-04 | |||
| Qwen3 235B-A22Balibaba/qwen3-235b-a22b | 131.072K | $0.7 | $2.8 | 2025-04 | |||
| DeepSeek V3 0324deepseek/deepseek-v3-0324 | 163.84K | $0.2 | $0.8 | 2025-03-24 | |||
| Mistral Small 3.1 24Bmistral/mistral-small-3-1-24b-instruct-2503 | 128K | $0.106 | $0.318 | 2025-03-17 | |||
| Command Acohere/command-a-03-2025 | 256K | $2.5 | $10 | 2025-03-13 | |||
| Gemma 3 4B ITgoogle/gemma-3-4b-it | 131.072K | $0.04 | $0.08 | 2025-03-12 | |||
| Gemma 3 12B ITgoogle/gemma-3-12b-it | 131.072K | $0.05 | $0.1 | 2025-03-12 | |||
| Gemma 3 27B ITgoogle/gemma-3-27b-it | 131.072K | $0.08 | $0.16 | 2025-03-12 | |||
| DeepSeek-R1deepseek/deepseek-r1 | 128K | $0.7 | $2.5 | 2025-01-20 | |||
| Llama-3.3-70B-Instructmeta/llama-3.3-70b-instruct | 128K | $0.1 | $0.32 | 2024-12-06 | |||
| Command R7Bcohere/command-r7b-12-2024 | 128K | $0.037 | $0.15 | 2024-12-02 | |||
| Mistral Large 3mistral/mistral-large-2512 | 262.144K | $0.5 | $1.5 | 2024-11-01 | |||
| Command R+cohere/command-r-plus-08-2024 | 128K | $2.5 | $10 | 2024-08-30 | |||
| Command Rcohere/command-r-08-2024 | 128K | $0.15 | $0.6 | 2024-08-30 | |||
| Llama-3.1-8B-Instructmeta/llama-3.1-8b-instruct | 128K | $0.02 | $0.04 | 2024-07-23 | |||
| Llama-3.1-70B-Instructmeta/llama-3.1-70b-instruct | 128K | $0.4 | $0.4 | 2024-07-23 | |||
| Mistral Nemomistral/mistral-nemo | 128K | $0.15 | $0.15 | 2024-07-01 |