Open Qwen coding heavyweight for repository reasoning and agentic engineering
Models
Every model in the catalog with source-linked pricing, context limits, provider availability, and published benchmark results.
March 2025 checkpoint of DeepSeek-V3 with improved reasoning and coding
Efficient multimodal model for instruction following, coding, reasoning, and function calling
Cohere command model for multilingual enterprise agents, tools, and chat
Largest open Gemma 3 instruction model for multilingual text generation and visual understanding
Open multimodal Gemma instruction model for multilingual text generation and image understanding
Open multimodal Gemma instruction model for efficient text generation and image understanding
Open reasoning model from the Qwen team for math, coding, and step-by-step problem solving
Compact open multilingual vision model for OCR and visual question answering
Open multilingual vision model for OCR, visual reasoning, and image question answering
Open Command R model optimized for Arabic enterprise chat, RAG, and cultural knowledge
ALLaM-2-7b instruction tuned model by SDAIA
R1 reasoning distilled into Qwen 2.5 32B for efficient open-weight step-by-step problem solving
Classic open reasoning model for transparent math, coding, and deliberate problem solving
Open DeepSeek MoE chat model for coding, math, and general reasoning
Compact Microsoft instruction model tuned for efficient coding assistance, reasoning, and low-latency agent tasks
Popular open Llama workhorse for multilingual chat, coding, and self-hosting
Cohere retrieval model for long-context chat and enterprise RAG workflows
Flagship Mistral model for advanced reasoning, coding, and multilingual work
Tiny open Qwen code model for lightweight completion and on-device coding
Open coding-focused Qwen model for code generation, repair, and repository reasoning
Mistral's larger vision model for document-heavy image understanding and chat
Flagship Mistral model for advanced reasoning, coding, and multilingual work
Open multilingual model optimized for generation across 23 languages
Compact open multilingual model optimized for generation across 23 languages
Efficient open Mistral edge model for on-device chat and function calling
Compact Mistral model for edge, latency-sensitive, and cost-efficient workloads
Open Whisper checkpoint for robust multilingual transcription and captioning
Speech transcription model for accurate audio-to-text and captioning workflows
Small open Llama base model for lightweight text generation and self-hosting
Open multimodal Llama model for image understanding, captioning, and visual QA
Compact open Llama base model for lightweight and on-device use
Mistral vision-language model for image understanding and multimodal chat
Qwen vision-language model for visual reasoning, documents, and agent tasks
Cohere's RAG workhorse for long-context enterprise search and tool use
Cohere retrieval model for long-context chat and enterprise RAG workflows
Compact Nemotron model for efficient reasoning and deployable AI agents
Llama 3.1-based safety classifier for moderating prompts and model responses
Open Llama instruction model for multilingual chat, reasoning, and coding
Compact open Llama model for lightweight chat, drafting, and self-hosting
Efficient Mistral-NVIDIA open model for multilingual chat and local deployment
Open Mistral code model for fill-in-the-middle and 80+ programming languages
Mistral code model for completions, refactors, and developer IDE workflows
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
| Model | Creator | Inputs | Context | Input | Output | Released | Compare |
|---|---|---|---|---|---|---|---|
| Qwen3-Coder 480B-A35B Instructalibaba/qwen3-coder-480b-a35b-instruct | 262.144K | $1.5 | $7.5 | 2025-04 | |||
| DeepSeek V3 0324deepseek/deepseek-v3-0324 | 163.84K | $0.2 | $0.8 | 2025-03-24 | |||
| Mistral Small 3.1 24Bmistral/mistral-small-3-1-24b-instruct-2503 | 128K | $0.106 | $0.318 | 2025-03-17 | |||
| Command Acohere/command-a-03-2025 | 256K | $2.5 | $10 | 2025-03-13 | |||
| Gemma 3 27B ITgoogle/gemma-3-27b-it | 131.072K | $0.08 | $0.16 | 2025-03-12 | |||
| Gemma 3 12B ITgoogle/gemma-3-12b-it | 131.072K | $0.05 | $0.1 | 2025-03-12 | |||
| Gemma 3 4B ITgoogle/gemma-3-4b-it | 131.072K | $0.04 | $0.08 | 2025-03-12 | |||
| QwQ 32Balibaba/qwq-32b | 131.072K | $0.287 | $0.861 | 2025-03-05 | |||
| Aya Vision 8Bcohere/c4ai-aya-vision-8b | 16K | — | — | 2025-03-04 | |||
| Aya Vision 32Bcohere/c4ai-aya-vision-32b | 16K | — | — | 2025-03-04 | |||
| Command R7B Arabiccohere/command-r7b-arabic-02-2025 | 128K | $0.037 | $0.15 | 2025-02-27 | |||
| ALLaM-2-7bsdaia/allam-2-7b | 4.096K | — | — | 2025-01-23 | |||
| DeepSeek-R1-Distill-Qwen-32Bdeepseek/deepseek-r1-distill-qwen-32b | 131.072K | $0.3 | $0.3 | 2025-01-20 | |||
| DeepSeek-R1deepseek/deepseek-r1 | 128K | $0.7 | $2.5 | 2025-01-20 | |||
| DeepSeek-V3deepseek/deepseek-v3 | 131.072K | $0.27 | $1.12 | 2024-12-26 | |||
| Phi-4-minimicrosoft/phi-4-mini | 128K | $0.075 | $0.3 | 2024-12-11 | |||
| Llama-3.3-70B-Instructmeta/llama-3.3-70b-instruct | 128K | $0.1 | $0.32 | 2024-12-06 | |||
| Command R7Bcohere/command-r7b-12-2024 | 128K | $0.037 | $0.15 | 2024-12-02 | |||
| Mistral Large 2.1mistral/mistral-large-2411 | 131.072K | $2 | $6 | 2024-11-18 | |||
| Qwen2.5-Coder-0.5Balibaba/qwen2.5-coder-0.5b | 32.768K | $0.1 | $0.1 | 2024-11-12 | |||
| Qwen2.5-Coder-32B-Instructalibaba/qwen2.5-coder-32b-instruct | 131.072K | $0.06 | $0.2 | 2024-11-12 | |||
| Pixtral Large (latest)mistral/pixtral-large-latest | 128K | $2 | $6 | 2024-11-01 | |||
| Mistral Large (latest)mistral/mistral-large-latest | 262.144K | $0.5 | $1.5 | 2024-11-01 | |||
| Aya Expanse 32Bcohere/c4ai-aya-expanse-32b | 128K | — | — | 2024-10-24 | |||
| Aya Expanse 8Bcohere/c4ai-aya-expanse-8b | 8K | — | — | 2024-10-24 | |||
| Ministral 8B Instructmistral/ministral-8b-instruct-2410 | 131.072K | $0.15 | $0.15 | 2024-10-16 | |||
| Ministral 3Bmistral/ministral-3b | 128K | $0.04 | $0.04 | 2024-10-16 | |||
| Whisper 3 Largeopenai/whisper-large-v3 | 448 | $0.002 | $0.002 | 2024-10-01 | |||
| Whisper Large v3 Turboopenai/whisper-large-v3-turbo | 448 | $0.002 | $0.002 | 2024-10-01 | |||
| Llama-3.2-3Bmeta/llama-3.2-3b | 131.072K | $0.1 | $0.1 | 2024-09-25 | |||
| Llama-3.2-11B-Vision-Instructmeta/llama-3.2-11b-vision-instruct | 128K | $0.055 | $0.055 | 2024-09-25 | |||
| Llama-3.2-1Bmeta/llama-3.2-1b | 131.072K | $0.1 | $0.1 | 2024-09-25 | |||
| Pixtral 12Bmistral/pixtral-12b | 128K | $0.15 | $0.15 | 2024-09-01 | |||
| Qwen2.5-VL 72B Instructalibaba/qwen2-5-vl-72b-instruct | 131.072K | $2.8 | $8.4 | 2024-09 | |||
| Command R+cohere/command-r-plus-08-2024 | 128K | $2.5 | $10 | 2024-08-30 | |||
| Command Rcohere/command-r-08-2024 | 128K | $0.15 | $0.6 | 2024-08-30 | |||
| Nemotron Mini 4B Instructnvidia/nemotron-mini-4b-instruct | 128K | — | — | 2024-08-21 | |||
| Llama-Guard-3-8Bmeta/llama-guard-3-8b | 128K | — | — | 2024-07-23 | |||
| Llama-3.1-70B-Instructmeta/llama-3.1-70b-instruct | 128K | $0.4 | $0.4 | 2024-07-23 | |||
| Llama-3.1-8B-Instructmeta/llama-3.1-8b-instruct | 128K | $0.02 | $0.04 | 2024-07-23 | |||
| Mistral Nemomistral/mistral-nemo | 128K | $0.15 | $0.15 | 2024-07-01 | |||
| Codestral-22B-v0.1mistral/codestral-22b-v0.1 | 32.768K | $0.3 | $0.9 | 2024-05-29 | |||
| Codestral (latest)mistral/codestral-latest | 256K | $0.3 | $0.9 | 2024-05-29 | |||
| Qwen/Qwen3.8-27BQwen/Qwen3.8-27B | Not documented | — | — | — | |||
| google/gemma-4-12B-it-qat-w4a16-ctgoogle/gemma-4-12B-it-qat-w4a16-ct | Not documented | — | — | — | |||
| meta-llama/Llama-3.2-1Bmeta-llama/Llama-3.2-1B | Not documented | — | — | — | |||
| google/gemma-4-31B-it-qat-w4a16-ctgoogle/gemma-4-31B-it-qat-w4a16-ct | Not documented | — | — | — | |||
| google/gemma-4-E4B-it-qat-mobile-ctgoogle/gemma-4-E4B-it-qat-mobile-ct | Not documented | — | — | — | |||
| meta-llama/Llama-Guard-3-11B-Visionmeta-llama/Llama-Guard-3-11B-Vision | Not documented | — | — | — | |||
| google/gemma-4-E2B-it-qat-mobile-ctgoogle/gemma-4-E2B-it-qat-mobile-ct | Not documented | — | — | — |