Qwen vision-language thinking model for visual reasoning, documents, and agent tasks
Models
Every model in the catalog with source-linked pricing, context limits, provider availability, and published benchmark results.
Qwen vision-language model for visual reasoning, documents, and agent tasks
Flagship Qwen3 model for coding agents, complex reasoning, and tool use
Open multimodal reasoning model for transparent analysis of text and images
Coding-optimized GPT model for repository edits, reviews, and agentic software work
Fully open 70B multilingual LLM supporting 1800+ languages with 65K context. Trained on 15T tokens of compliant open data. Apache 2.0, EU AI Act compliant.
Flagship Indian-language reasoning model for enterprise multilingual applications
Efficient Qwen thinking model for local reasoning, math, and coding agents
Low-latency ByteDance Seed model for high-throughput chat, extraction, and lightweight tool use
Nano Banana image model for fast generation, edits, and character-consistent assets
Hybrid-reasoning DeepSeek model with thinking and non-thinking modes
Cohere reasoning model for multilingual enterprise agents, tools, and complex workflows
Compact Nemotron model for efficient reasoning and deployable AI agents
ByteDance Seed multimodal model for image understanding, visual reasoning, and tool-assisted tasks
GLM vision model for visual reasoning, documents, and multimodal agents
Tiny GPT-5 lane for routing, extraction, classification, and bulk jobs
Chat-tuned GPT model for conversational assistance, writing, and tool workflows
Original GPT-5 workhorse for reasoning, coding, writing, and tool workflows
Small GPT-5 for responsive agents, coding help, and everyday automation
Flagship Claude model for deep reasoning, coding, and long-horizon agents
Open GPT reasoning model for self-hosted agents and controllable deployments
Open GPT reasoning model for self-hosted agents and controllable deployments
Flagship Claude model for deep reasoning, coding, and long-horizon agents
Lighter GLM-4.5 variant for fast coding assistance and cheaper agents
Efficient GLM model for fast reasoning, coding, and agent workflows
Efficient Qwen model for fast chat, extraction, and high-volume workloads
Hybrid-reasoning GLM release that made the 4.5 line broadly useful
Nemotron model for efficient reasoning, coding, and specialized AI agents
Efficient Mistral model for fast chat, extraction, and production assistants
Fast Gemini workhorse for multimodal apps where latency and price matter
Google's proven reasoning model for coding, math, and multimodal analysis
Lean Gemini 2.5 lane for cheap multimodal traffic and quick agents
Open Mistral reasoning model for transparent step-by-step problem solving
High-effort o3 tier for difficult technical reasoning and careful answers
Flagship Claude model for deep reasoning, coding, and long-horizon agents
Balanced Claude model for coding, analysis, agent workflows, and cost control
Flagship Claude model for deep reasoning, coding, and long-horizon agents
Balanced Claude model for coding, analysis, agent workflows, and cost control
Flagship model for demanding analysis, coding, and production agent workflows
Sparse MoE Qwen model with 3B active parameters for efficient chat and reasoning
Reasoning model for deliberate analysis, multi-step problem solving, and tool use
Deliberate o-series reasoner for hard math, coding, and multi-step analysis
Fast o-series model for compact reasoning, coding, and tool use
Mistral vision-language model for image understanding and multimodal chat
Flagship Nemotron model for high-throughput reasoning and complex agents
Nemotron model for efficient reasoning, coding, and specialized AI agents
Large open Qwen MoE for multilingual reasoning, coding, and tool use
Dense open Qwen model for self-hosted chat, reasoning, and coding
O-series reasoning model for hard analysis, math, coding, and planning
Mistral reasoning model for transparent analysis, math, and complex decisions
| Model | Creator | Inputs | Context | Input | Output | Released | Compare |
|---|---|---|---|---|---|---|---|
| Qwen3 VL 235B A22B Thinkingalibaba/qwen3-vl-235b-a22b-thinking | 131.072K | $0.4 | $4 | 2025-09-23 | |||
| Qwen3-VL Plusalibaba/qwen3-vl-plus | 262.144K | $0.2 | $1.6 | 2025-09-23 | |||
| Qwen3 Maxalibaba/qwen3-max | 262.144K | $1.2 | $6 | 2025-09-23 | |||
| Magistral Small 1.2mistral/magistral-small-2509 | 131.072K | $0.5 | $1.5 | 2025-09-18 | |||
| GPT-5-Codexopenai/gpt-5-codex | 400K | $1.1 | $9 | 2025-09-15 | |||
| Apertus 70Bswiss-ai/apertus-70b | 65.536K | $0.46 | $2.42 | 2025-09-02 | |||
| Sarvam 105Bsarvam/sarvam-105b | 131.072K | $0.04 | $0.16 | 2025-09-01 | |||
| Qwen3-Next 80B-A3B (Thinking)alibaba/qwen3-next-80b-a3b-thinking | 131.072K | $0.5 | $6 | 2025-09 | |||
| Seed 1.6 Flashbytedance-seed/seed-1-6-flash | 256K | $0.022 | $0.223 | 2025-08-28 | |||
| Nano Bananagoogle/gemini-2.5-flash-image | 32.768K | $0.3 | $30 | 2025-08-26 | |||
| DeepSeek-V3.1deepseek/deepseek-v3.1 | 131.072K | $0.19 | $0.71 | 2025-08-21 | |||
| Command A Reasoningcohere/command-a-reasoning-08-2025 | 256K | $2.5 | $10 | 2025-08-21 | |||
| Nemotron Nano 9B v2nvidia/nemotron-nano-9b-v2 | 131.072K | $0.06 | $0.23 | 2025-08-18 | |||
| Seed 1.6 Visionbytedance-seed/seed-1-6-vision | 256K | $0.119 | $1.187 | 2025-08-15 | |||
| GLM-4.5Vzhipuai/glm-4.5v | 64K | $0.6 | $1.8 | 2025-08-11 | |||
| GPT-5 Nanoopenai/gpt-5-nano | 400K | $0.05 | $0.4 | 2025-08-07 | |||
| GPT-5 Chat (latest)openai/gpt-5-chat-latest | 400K | $1.25 | $10 | 2025-08-07 | |||
| GPT-5openai/gpt-5 | 400K | $1.25 | $10 | 2025-08-07 | |||
| GPT-5 Miniopenai/gpt-5-mini | 400K | $0.25 | $2 | 2025-08-07 | |||
| Claude Opus 4.1 (latest)anthropic/claude-opus-4-1 | 200K | $15 | $75 | 2025-08-05 | |||
| GPT OSS 20Bopenai/gpt-oss-20b | 131.072K | $0.02 | $0.1 | 2025-08-05 | |||
| GPT OSS 120Bopenai/gpt-oss-120b | 131.072K | $0.03 | $0.17 | 2025-08-05 | |||
| Claude Opus 4.1anthropic/claude-opus-4-1-20250805 | 200K | $15 | $75 | 2025-08-05 | |||
| GLM-4.5-Airzhipuai/glm-4.5-air | 131.072K | $0.2 | $1.1 | 2025-07-28 | |||
| GLM-4.5-Flashzhipuai/glm-4.5-flash | 131.072K | — | — | 2025-07-28 | |||
| Qwen Flashalibaba/qwen-flash | 1M | $0.05 | $0.4 | 2025-07-28 | |||
| GLM-4.5zhipuai/glm-4.5 | 131.072K | $0.6 | $2.2 | 2025-07-28 | |||
| Llama 3.3 Nemotron Super 49B v1.5nvidia/llama-3.3-nemotron-super-49b-v1.5 | 131.072K | $0.4 | $0.4 | 2025-07-25 | |||
| Mistral Small 3.2mistral/mistral-small-2506 | 128K | $0.1 | $0.3 | 2025-06-20 | |||
| Gemini 2.5 Flashgoogle/gemini-2.5-flash | 1.04858M | $0.3 | $2.5 | 2025-06-17 | |||
| Gemini 2.5 Progoogle/gemini-2.5-pro | 1.04858M | $1.25 | $10 | 2025-06-17 | |||
| Gemini 2.5 Flash-Litegoogle/gemini-2.5-flash-lite | 1.04858M | $0.1 | $0.4 | 2025-06-17 | |||
| Magistral Smallmistral/magistral-small-2506 | 131.072K | $0.5 | $1.5 | 2025-06-10 | |||
| o3-proopenai/o3-pro | 200K | $20 | $80 | 2025-06-10 | |||
| Claude Opus 4 (latest)anthropic/claude-opus-4-0 | 200K | — | — | 2025-05-22 | |||
| Claude Sonnet 4anthropic/claude-sonnet-4-20250514 | 200K | $3 | $15 | 2025-05-22 | |||
| Claude Opus 4anthropic/claude-opus-4-20250514 | 200K | $15 | $75 | 2025-05-22 | |||
| Claude Sonnet 4 (latest)anthropic/claude-sonnet-4-0 | 200K | $2.898 | $14.493 | 2025-05-22 | |||
| Solar Pro 2upstage/solar-pro2 | 65.536K | $0.15 | $0.6 | 2025-05-20 | |||
| Qwen3 30B A3Balibaba/qwen3-30b-a3b | 131.072K | $0.08 | $0.29 | 2025-04-28 | |||
| Palmyra X5writer/palmyra-x5 | 1M | $0.6 | $6 | 2025-04-28 | |||
| o3openai/o3 | 200K | $2 | $8 | 2025-04-16 | |||
| o4-miniopenai/o4-mini | 200K | $1.1 | $4.4 | 2025-04-16 | |||
| Pixtral Large (25.02)mistral/pixtral-large-2502 | 128K | $1.993 | $5.978 | 2025-04-08 | |||
| Llama 3.1 Nemotron Ultra 253Bnvidia/llama-3.1-nemotron-ultra-253b | 128K | — | — | 2025-04-07 | |||
| Llama 3.3 Nemotron Super 49B v1nvidia/llama-3.3-nemotron-super-49b-v1 | 131.072K | — | — | 2025-04-07 | |||
| Qwen3 235B-A22Balibaba/qwen3-235b-a22b | 131.072K | $0.7 | $2.8 | 2025-04 | |||
| Qwen3 32Balibaba/qwen3-32b | 131.072K | $0.7 | $2.8 | 2025-04 | |||
| o1-proopenai/o1-pro | 200K | $150 | $600 | 2025-03-19 | |||
| Magistral Medium (latest)mistral/magistral-medium-latest | 128K | $2 | $5 | 2025-03-17 |