Experimental chat-tuned 6B MoE model with 1B active parameters for low-resource chat and instruction following
Models
Every model in the catalog with source-linked pricing, context limits, provider availability, and published benchmark results.
DeepSeek reasoning model for multi-step analysis, math, coding, and tools
DeepSeek chat model for instruction following, coding, and analysis
Reasoning-tuned 26B MoE model with 3B active parameters for agents, tools, and multi-step workloads
Flagship Claude model for deep reasoning, coding, and long-horizon agents
Nano Banana Pro for higher-fidelity image generation and design-heavy edits
xAI's fast agentic tool-calling model with a 2M context window; non-reasoning variant for low-latency responses
xAI's fast agentic tool-calling model with a 2M context window and built-in reasoning
Preview Gemini flagship for complex reasoning, coding, and rich multimodal prompts
Coding-optimized GPT model for repository edits, reviews, and agentic software work
Sharper GPT-5 generation for coding, product work, and tool-assisted tasks
Coding-optimized GPT model for repository edits, reviews, and agentic software work
Codex GPT for repository edits, code review, and practical software agents
Chat-tuned GPT-5.1 for polished assistants, writing, and product conversations
Thinking Kimi model for slower research passes, planning, and hard technical questions
Kimi reasoning model for long-horizon research, planning, and tool use
Flagship Claude model for deep reasoning, coding, and long-horizon agents
Safety model for policy screening, moderation, and risk-aware routing workflows
Safety model for policy screening, moderation, and risk-aware routing workflows
Nemotron multimodal model for visual reasoning and agentic AI workflows
Safety model for policy screening, moderation, and risk-aware routing workflows
Efficient open MiniMax model built for coding agents and tool-heavy workflows
Fast Claude lane for lightweight agents, office tasks, and responsive chat
ByteDance Seed model for long-context reasoning, instruction following, and tool-assisted tasks
Fast Claude model for responsive assistance, classification, and lightweight agents
Specialized Gemini 2.5 model for browser-control agents that automate UI tasks
Higher-accuracy GPT-5 tier for tough analysis, coding reviews, and planning
Open-weight hybrid model for enterprise chat, coding, retrieval-augmented generation, and tool-calling workloads
Compact open-weight hybrid Granite model for lightweight enterprise chat and tool calling
Late GLM-4 workhorse for coding agents, reasoning, and structured tasks
Speech generation model for controllable voice, narration, and audio delivery
Speech generation model for controllable voice, narration, and audio delivery
Balanced Claude model for coding, analysis, agent workflows, and cost control
Balanced Claude model for coding, analysis, agent workflows, and cost control
Qwen vision-language thinking model for visual reasoning, documents, and agent tasks
Qwen vision-language model for visual reasoning, documents, and agent tasks
Flagship Qwen3 model for coding agents, complex reasoning, and tool use
Gemma 3 27B tuned by AI Singapore for Southeast Asian languages and instruction following
Qwen vision-language instruct model for visual reasoning, documents, and agent tasks
Open multimodal reasoning model for transparent analysis of text and images
Coding-optimized GPT model for repository edits, reviews, and agentic software work
Fully open 70B multilingual LLM supporting 1800+ languages with 65K context. Trained on 15T tokens of compliant open data. Apache 2.0, EU AI Act compliant.
Fully open 8B multilingual LLM supporting 1800+ languages with 65K context. Trained on compliant open data. Apache 2.0, EU AI Act compliant.
Flagship Indian-language reasoning model for enterprise multilingual applications
Efficient Qwen thinking model for local reasoning, math, and coding agents
Qwen instruction model for multilingual chat, reasoning, and tool use
Low-latency ByteDance Seed model for high-throughput chat, extraction, and lightweight tool use
Nano Banana image model for fast generation, edits, and character-consistent assets
Hybrid-reasoning DeepSeek model with thinking and non-thinking modes
Cohere reasoning model for multilingual enterprise agents, tools, and complex workflows
| Model | Creator | Inputs | Context | Input | Output | Released | Compare |
|---|---|---|---|---|---|---|---|
| Trinity Nano Previewarcee-ai/trinity-nano-preview | 131.072K | — | — | 2025-12-01 | |||
| DeepSeek Reasonerdeepseek/deepseek-reasoner | 1M | $0.147 | $0.295 | 2025-12-01 | |||
| DeepSeek Chatdeepseek/deepseek-chat | 1M | $0.147 | $0.295 | 2025-12-01 | |||
| Trinity Miniarcee-ai/trinity-mini | 131.072K | $0.045 | $0.15 | 2025-12-01 | |||
| Claude Opus 4.5 (latest)anthropic/claude-opus-4-5 | 200K | $5 | $25 | 2025-11-24 | |||
| Nano Banana Pro Previewgoogle/gemini-3-pro-image-preview | 65.536K | $1 | $6 | 2025-11-20 | |||
| Grok 4.1 Fastxai/grok-4.1-fast | 2M | $0.2 | $0.5 | 2025-11-19 | |||
| Grok 4.1 Fast (Reasoning)xai/grok-4.1-fast-reasoning | 2M | $0.2 | $0.5 | 2025-11-19 | |||
| Gemini 3 Pro Previewgoogle/gemini-3-pro-preview | 1.04858M | $0.57 | $3.43 | 2025-11-18 | |||
| GPT-5.1 Codex Maxopenai/gpt-5.1-codex-max | 400K | $1.1 | $9 | 2025-11-13 | |||
| GPT-5.1openai/gpt-5.1 | 400K | $1.25 | $10 | 2025-11-13 | |||
| GPT-5.1 Codex miniopenai/gpt-5.1-codex-mini | 400K | $0.22 | $1.8 | 2025-11-13 | |||
| GPT-5.1 Codexopenai/gpt-5.1-codex | 400K | $1.07 | $8.5 | 2025-11-13 | |||
| GPT-5.1 Chatopenai/gpt-5.1-chat-latest | 128K | $1.25 | $10 | 2025-11-13 | |||
| Kimi K2 Thinkingmoonshotai/kimi-k2-thinking | 262.144K | $0.4 | $2.5 | 2025-11-06 | |||
| Kimi K2 Thinking Turbomoonshotai/kimi-k2-thinking-turbo | 262.144K | $1.15 | $8 | 2025-11-06 | |||
| Claude Opus 4.5anthropic/claude-opus-4-5-20251101 | 200K | $5 | $25 | 2025-11-01 | |||
| GPT OSS Safeguard 120Bopenai/gpt-oss-safeguard-120b | 131.072K | $0.15 | $0.6 | 2025-10-29 | |||
| GPT OSS Safeguard 20Bopenai/gpt-oss-safeguard-20b | 131.072K | $0.07 | $0.2 | 2025-10-29 | |||
| Nemotron Nano 12B v2 VLnvidia/nemotron-nano-12b-v2-vl | 128K | $0.2 | $0.6 | 2025-10-28 | |||
| Llama 3.1 Nemotron Safety Guard 8B v3nvidia/llama-3.1-nemotron-safety-guard-8b-v3 | 128K | — | — | 2025-10-28 | |||
| MiniMax-M2minimax/MiniMax-M2 | 204.8K | $0.3 | $1.2 | 2025-10-27 | |||
| Claude Haiku 4.5 (latest)anthropic/claude-haiku-4-5 | 200K | $1 | $5 | 2025-10-15 | |||
| Seed 1.6bytedance-seed/seed-1-6 | 256K | $0.119 | $1.187 | 2025-10-15 | |||
| Claude Haiku 4.5anthropic/claude-haiku-4-5-20251001 | 200K | $1 | $5 | 2025-10-15 | |||
| Gemini 2.5 Computer Use Previewgoogle/gemini-2.5-computer-use-preview-10-2025 | 128K | $1.25 | $10 | 2025-10-07 | |||
| GPT-5 Proopenai/gpt-5-pro | 400K | $15 | $120 | 2025-10-06 | |||
| Granite-4.0-H-Smallibm/granite-4-h-small | 131.072K | $0.064 | $0.265 | 2025-10-02 | |||
| Granite-4.0-H-Microibm/granite-4-h-micro | 131.072K | — | — | 2025-10-02 | |||
| GLM-4.6zhipuai/glm-4.6 | 204.8K | $0.6 | $2.2 | 2025-09-30 | |||
| Gemini 2.5 Pro TTSgoogle/gemini-2.5-pro-tts | 32.768K | $1 | $20 | 2025-09-30 | |||
| Gemini 2.5 Flash TTSgoogle/gemini-2.5-flash-tts | 32.768K | $0.5 | $10 | 2025-09-30 | |||
| Claude Sonnet 4.5anthropic/claude-sonnet-4-5-20250929 | 200K | $3 | $15 | 2025-09-29 | |||
| Claude Sonnet 4.5 (latest)anthropic/claude-sonnet-4-5 | 200K | $3 | $15 | 2025-09-29 | |||
| Qwen3 VL 235B A22B Thinkingalibaba/qwen3-vl-235b-a22b-thinking | 131.072K | $0.4 | $4 | 2025-09-23 | |||
| Qwen3-VL Plusalibaba/qwen3-vl-plus | 262.144K | $0.2 | $1.6 | 2025-09-23 | |||
| Qwen3 Maxalibaba/qwen3-max | 262.144K | $1.2 | $6 | 2025-09-23 | |||
| Gemma-SEA-LION-v4-27B-ITaisingapore/gemma-sea-lion-v4-27b-it | 128K | — | — | 2025-09-23 | |||
| Qwen3 VL 235B A22B Instructalibaba/qwen3-vl-235b-a22b-instruct | 131.072K | $0.2 | $0.88 | 2025-09-23 | |||
| Magistral Small 1.2mistral/magistral-small-2509 | 131.072K | $0.5 | $1.5 | 2025-09-18 | |||
| GPT-5-Codexopenai/gpt-5-codex | 400K | $1.1 | $9 | 2025-09-15 | |||
| Apertus 70Bswiss-ai/apertus-70b | 65.536K | $0.46 | $2.42 | 2025-09-02 | |||
| Apertus 8Bswiss-ai/apertus-8b | 65.536K | — | — | 2025-09-02 | |||
| Sarvam 105Bsarvam/sarvam-105b | 131.072K | $0.04 | $0.16 | 2025-09-01 | |||
| Qwen3-Next 80B-A3B (Thinking)alibaba/qwen3-next-80b-a3b-thinking | 131.072K | $0.5 | $6 | 2025-09 | |||
| Qwen3-Next 80B-A3B Instructalibaba/qwen3-next-80b-a3b-instruct | 131.072K | $0.5 | $2 | 2025-09 | |||
| Seed 1.6 Flashbytedance-seed/seed-1-6-flash | 256K | $0.022 | $0.223 | 2025-08-28 | |||
| Nano Bananagoogle/gemini-2.5-flash-image | 32.768K | $0.3 | $30 | 2025-08-26 | |||
| DeepSeek-V3.1deepseek/deepseek-v3.1 | 131.072K | $0.19 | $0.71 | 2025-08-21 | |||
| Command A Reasoningcohere/command-a-reasoning-08-2025 | 256K | $2.5 | $10 | 2025-08-21 |