Compact Mistral model for edge, latency-sensitive, and cost-efficient workloads
5 providers
Every model in the catalog with source-linked pricing, context limits, provider availability, and published benchmark results.
Compact Mistral model for edge, latency-sensitive, and cost-efficient workloads
Efficient Mistral-NVIDIA open model for multilingual chat and local deployment
| Model | Creator | Inputs | Context | Input | Output | Released | Compare |
|---|---|---|---|---|---|---|---|
| Ministral 3Bmistral/ministral-3b | 128K | $0.04 | $0.04 | 2024-10-16 | |||
| Mistral Nemomistral/mistral-nemo | 128K | $0.15 | $0.15 | 2024-07-01 |