DeepSeek-V3.2 is a large language model designed to harmonize high computational efficiency with strong reasoning and agentic tool-use performance. It introduces DeepSeek Sparse Attention (DSA), a fine-grained sparse attention mechanism...
Models
Every model in the catalog with source-linked pricing, context limits, provider availability, and published benchmark results.
Flagship Claude model for deep reasoning, coding, and long-horizon agents
Nano Banana Pro is Google’s most advanced image-generation and editing model, built on Gemini 3 Pro. It extends the original Nano Banana with significantly improved multimodal reasoning, real-world grounding, and...
xAI's fast agentic tool-calling model with a 2M context window and built-in reasoning
xAI's fast agentic tool-calling model with a 2M context window; non-reasoning variant for low-latency responses
Preview Gemini flagship for complex reasoning, coding, and rich multimodal prompts
GPT-5.1-Codex-Mini is a smaller and faster version of GPT-5.1-Codex
Chat-tuned GPT-5.1 for polished assistants, writing, and product conversations
GPT-5.1-Codex is a specialized version of GPT-5.1 optimized for software engineering and coding workflows. It is designed for both interactive development sessions and long, independent execution of complex engineering tasks....
GPT-5.1 is the latest frontier-grade model in the GPT-5 series, offering stronger general-purpose reasoning, improved instruction adherence, and a more natural conversational style compared to GPT-5. It uses adaptive reasoning...
GPT-5.1-Codex-Max is OpenAI’s latest agentic coding model, designed for long-running, high-context software development tasks. It is based on an updated version of the 5.1 reasoning stack and trained on agentic...
Kimi reasoning model for long-horizon research, planning, and tool use
Kimi K2 Thinking is Moonshot AI’s most advanced open reasoning model to date, extending the K2 series into agentic, long-horizon reasoning. Built on the trillion-parameter Mixture-of-Experts (MoE) architecture introduced in...
Flagship Claude model for deep reasoning, coding, and long-horizon agents
gpt-oss-safeguard-20b is a safety reasoning model from OpenAI built upon gpt-oss-20b. This open-weight, 21B-parameter Mixture-of-Experts (MoE) model offers lower latency for safety tasks like content classification, LLM filtering, and trust...
Safety model for policy screening, moderation, and risk-aware routing workflows
Safety model for policy screening, moderation, and risk-aware routing workflows
Nemotron multimodal model for visual reasoning and agentic AI workflows
Efficient open MiniMax model built for coding agents and tool-heavy workflows
ByteDance Seed model for long-context reasoning, instruction following, and tool-assisted tasks
Fast Claude model for responsive assistance, classification, and lightweight agents
Fast Claude lane for lightweight agents, office tasks, and responsive chat
Specialized Gemini 2.5 model for browser-control agents that automate UI tasks
GPT-5 Pro is OpenAI’s most advanced model, offering major improvements in reasoning, code quality, and user experience. It is optimized for complex tasks that require step-by-step reasoning, instruction following, and...
Compact open-weight hybrid Granite model for lightweight enterprise chat and tool calling
Open-weight hybrid model for enterprise chat, coding, retrieval-augmented generation, and tool-calling workloads
Speech generation model for controllable voice, narration, and audio delivery
Speech generation model for controllable voice, narration, and audio delivery
Late GLM-4 workhorse for coding agents, reasoning, and structured tasks
Balanced Claude model for coding, analysis, agent workflows, and cost control
Balanced Claude model for coding, analysis, agent workflows, and cost control
Flagship Qwen3 model for coding agents, complex reasoning, and tool use
Gemma 3 27B tuned by AI Singapore for Southeast Asian languages and instruction following
Qwen vision-language thinking model for visual reasoning, documents, and agent tasks
Qwen vision-language model for visual reasoning, documents, and agent tasks
Qwen vision-language instruct model for visual reasoning, documents, and agent tasks
Open multimodal reasoning model for transparent analysis of text and images
Coding-optimized GPT model for repository edits, reviews, and agentic software work
Fully open 70B multilingual LLM supporting 1800+ languages with 65K context. Trained on 15T tokens of compliant open data. Apache 2.0, EU AI Act compliant.
Fully open 8B multilingual LLM supporting 1800+ languages with 65K context. Trained on compliant open data. Apache 2.0, EU AI Act compliant.
Flagship Indian-language reasoning model for enterprise multilingual applications
Qwen instruction model for multilingual chat, reasoning, and tool use
Efficient Qwen thinking model for local reasoning, math, and coding agents
Low-latency ByteDance Seed model for high-throughput chat, extraction, and lightweight tool use
Gemini 2.5 Flash Image, a.k.a. "Nano Banana," is now generally available. It is a state of the art image generation model with contextual understanding. It is capable of image generation,...
Cohere reasoning model for multilingual enterprise agents, tools, and complex workflows
Hybrid-reasoning DeepSeek model with thinking and non-thinking modes
Compact Nemotron model for efficient reasoning and deployable AI agents
ByteDance Seed multimodal model for image understanding, visual reasoning, and tool-assisted tasks
GLM vision model for visual reasoning, documents, and multimodal agents
| Model | Creator | Inputs | Context | Input | Output | Released | Compare |
|---|---|---|---|---|---|---|---|
| DeepSeek: DeepSeek V3.2deepseek/deepseek-v3.2 | 163.84K | $0.269 | $0.4 | 2025-12-01 | |||
| Claude Opus 4.5 (latest)anthropic/claude-opus-4-5 | 200K | $5 | $25 | 2025-11-24 | |||
| Google: Nano Banana Pro (Gemini 3 Pro Image Preview)google/gemini-3-pro-image-preview | 65.536K | $2 | $12 | 2025-11-20 | |||
| Grok 4.1 Fast (Reasoning)xai/grok-4.1-fast-reasoning | 2M | $0.2 | $0.5 | 2025-11-19 | |||
| Grok 4.1 Fastxai/grok-4.1-fast | 2M | $0.2 | $0.5 | 2025-11-19 | |||
| Gemini 3 Pro Previewgoogle/gemini-3-pro-preview | 1.04858M | $0.57 | $3.43 | 2025-11-18 | |||
| OpenAI: GPT-5.1-Codex-Miniopenai/gpt-5.1-codex-mini | 400K | $0.25 | $2 | 2025-11-13 | |||
| GPT-5.1 Chatopenai/gpt-5.1-chat-latest | 128K | $1.25 | $10 | 2025-11-13 | |||
| OpenAI: GPT-5.1-Codexopenai/gpt-5.1-codex | 400K | $1.25 | $10 | 2025-11-13 | |||
| OpenAI: GPT-5.1openai/gpt-5.1 | 400K | $1.25 | $10 | 2025-11-13 | |||
| OpenAI: GPT-5.1-Codex-Maxopenai/gpt-5.1-codex-max | 400K | $1.25 | $10 | 2025-11-13 | |||
| Kimi K2 Thinking Turbomoonshotai/kimi-k2-thinking-turbo | 262.144K | $1.15 | $8 | 2025-11-06 | |||
| MoonshotAI: Kimi K2 Thinkingmoonshotai/kimi-k2-thinking | 262.144K | $0.6 | $2.5 | 2025-11-06 | |||
| Claude Opus 4.5anthropic/claude-opus-4-5-20251101 | 200K | $5 | $25 | 2025-11-01 | |||
| OpenAI: gpt-oss-safeguard-20bopenai/gpt-oss-safeguard-20b | 131.072K | $0.075 | $0.3 | 2025-10-29 | |||
| GPT OSS Safeguard 120Bopenai/gpt-oss-safeguard-120b | 131.072K | $0.15 | $0.6 | 2025-10-29 | |||
| Llama 3.1 Nemotron Safety Guard 8B v3nvidia/llama-3.1-nemotron-safety-guard-8b-v3 | 128K | — | — | 2025-10-28 | |||
| Nemotron Nano 12B v2 VLnvidia/nemotron-nano-12b-v2-vl | 128K | $0.2 | $0.6 | 2025-10-28 | |||
| MiniMax-M2minimax/MiniMax-M2 | 204.8K | $0.3 | $1.2 | 2025-10-27 | |||
| Seed 1.6bytedance-seed/seed-1-6 | 256K | $0.119 | $1.187 | 2025-10-15 | |||
| Claude Haiku 4.5anthropic/claude-haiku-4-5-20251001 | 200K | $1 | $5 | 2025-10-15 | |||
| Claude Haiku 4.5 (latest)anthropic/claude-haiku-4-5 | 200K | $1 | $5 | 2025-10-15 | |||
| Gemini 2.5 Computer Use Previewgoogle/gemini-2.5-computer-use-preview-10-2025 | 128K | $1.25 | $10 | 2025-10-07 | |||
| OpenAI: GPT-5 Proopenai/gpt-5-pro | 400K | $15 | $120 | 2025-10-06 | |||
| Granite-4.0-H-Microibm/granite-4-h-micro | 131.072K | — | — | 2025-10-02 | |||
| Granite-4.0-H-Smallibm/granite-4-h-small | 131.072K | $0.064 | $0.265 | 2025-10-02 | |||
| Gemini 2.5 Flash TTSgoogle/gemini-2.5-flash-tts | 32.768K | $0.5 | $10 | 2025-09-30 | |||
| Gemini 2.5 Pro TTSgoogle/gemini-2.5-pro-tts | 32.768K | $1 | $20 | 2025-09-30 | |||
| GLM-4.6zhipuai/glm-4.6 | 204.8K | $0.6 | $2.2 | 2025-09-30 | |||
| Claude Sonnet 4.5 (latest)anthropic/claude-sonnet-4-5 | 200K | $3 | $15 | 2025-09-29 | |||
| Claude Sonnet 4.5anthropic/claude-sonnet-4-5-20250929 | 200K | $3 | $15 | 2025-09-29 | |||
| Qwen3 Maxalibaba/qwen3-max | 262.144K | $1.2 | $6 | 2025-09-23 | |||
| Gemma-SEA-LION-v4-27B-ITaisingapore/gemma-sea-lion-v4-27b-it | 128K | — | — | 2025-09-23 | |||
| Qwen3 VL 235B A22B Thinkingalibaba/qwen3-vl-235b-a22b-thinking | 131.072K | $0.4 | $4 | 2025-09-23 | |||
| Qwen3-VL Plusalibaba/qwen3-vl-plus | 262.144K | $0.2 | $1.6 | 2025-09-23 | |||
| Qwen3 VL 235B A22B Instructalibaba/qwen3-vl-235b-a22b-instruct | 131.072K | $0.2 | $0.88 | 2025-09-23 | |||
| Magistral Small 1.2mistral/magistral-small-2509 | 131.072K | $0.5 | $1.5 | 2025-09-18 | |||
| GPT-5-Codexopenai/gpt-5-codex | 400K | $1.1 | $9 | 2025-09-15 | |||
| Apertus 70Bswiss-ai/apertus-70b | 65.536K | $0.46 | $2.42 | 2025-09-02 | |||
| Apertus 8Bswiss-ai/apertus-8b | 65.536K | — | — | 2025-09-02 | |||
| Sarvam 105Bsarvam/sarvam-105b | 131.072K | $0.04 | $0.16 | 2025-09-01 | |||
| Qwen3-Next 80B-A3B Instructalibaba/qwen3-next-80b-a3b-instruct | 131.072K | $0.5 | $2 | 2025-09 | |||
| Qwen3-Next 80B-A3B (Thinking)alibaba/qwen3-next-80b-a3b-thinking | 131.072K | $0.5 | $6 | 2025-09 | |||
| Seed 1.6 Flashbytedance-seed/seed-1-6-flash | 256K | $0.022 | $0.223 | 2025-08-28 | |||
| Google: Nano Banana (Gemini 2.5 Flash Image)google/gemini-2.5-flash-image | 32.768K | $0.3 | $2.5 | 2025-08-26 | |||
| Command A Reasoningcohere/command-a-reasoning-08-2025 | 256K | $2.5 | $10 | 2025-08-21 | |||
| DeepSeek-V3.1deepseek/deepseek-v3.1 | 131.072K | $0.19 | $0.71 | 2025-08-21 | |||
| Nemotron Nano 9B v2nvidia/nemotron-nano-9b-v2 | 131.072K | $0.06 | $0.23 | 2025-08-18 | |||
| Seed 1.6 Visionbytedance-seed/seed-1-6-vision | 256K | $0.119 | $1.187 | 2025-08-15 | |||
| GLM-4.5Vzhipuai/glm-4.5v | 64K | $0.6 | $1.8 | 2025-08-11 |