Sparse MoE Qwen model with 3B active parameters for efficient chat and reasoning
8 providers
Every model in the catalog with source-linked pricing, context limits, provider availability, and published benchmark results.
Sparse MoE Qwen model with 3B active parameters for efficient chat and reasoning
Compact Mistral model for edge, latency-sensitive, and cost-efficient workloads
Cohere's RAG workhorse for long-context enterprise search and tool use
Efficient Mistral-NVIDIA open model for multilingual chat and local deployment
| Model | Creator | Inputs | Context | Input | Output | Released | Compare |
|---|---|---|---|---|---|---|---|
| Qwen3 30B A3Balibaba/qwen3-30b-a3b | 131.072K | $0.08 | $0.29 | 2025-04-28 | |||
| Ministral 3Bmistral/ministral-3b | 128K | $0.04 | $0.04 | 2024-10-16 | |||
| Command R+cohere/command-r-plus-08-2024 | 128K | $2.5 | $10 | 2024-08-30 | |||
| Mistral Nemomistral/mistral-nemo | 128K | $0.15 | $0.15 | 2024-07-01 |