Providers
Provider-specific identifiers, limits, and listed prices per million tokens. Every row links back to the provider's own documentation.
The model record exists, but no source-linked provider offer is available yet.
Submit a sourceMistral model for multilingual chat, reasoning, and tool-assisted workflows
Provider-specific identifiers, limits, and listed prices per million tokens. Every row links back to the provider's own documentation.
The model record exists, but no source-linked provider offer is available yet.
Submit a sourceRecorded from the source catalog and provider listings.
nvidia/mistral-nemotronEvery result stays attached to its source, version, metric and harness. Scores from different versions are never merged, and the profile below plots each benchmark against its own population rather than on a shared scale.
ModelBench does not infer quality from price, context size, or model name. When a source publishes a comparable result, it appears here with its version and link.
Read the methodologyField-level changes detected between successful source imports.
This record has not changed within the retained import history.
Every figure on this page traces back to one of these records.
Fetch the complete source-linked model record. No key, no account, no rate-limited tier.
GET https://model.kyssta.lol/api/v1/models/nvidia/mistral-nemotroncurl "https://model.kyssta.lol/api/v1/models/nvidia/mistral-nemotron"Answered directly from the stored record — nothing here is generated beyond the catalog's own fields.
Mistral model for multilingual chat, reasoning, and tool-assisted workflows. It is published by NVIDIA and catalogued here from Models.dev.
Mistral Nemotron accepts up to 128K tokens of context and returns up to 8.192K output tokens.
Yes. The weights are published.
The catalog records a release date of 2025-06-11, last verified Sep 11, 2026.
NVIDIA-Nemotron-Nano-9B-v2 is a large language model (LLM) trained from scratch by NVIDIA, and designed as a unified model for both reasoning and non-reasoning tasks. It responds to user queries and...
NVIDIA: Nemotron 3 Super (free)NVIDIA Nemotron 3 Super is a 120B-parameter open hybrid MoE model, activating just 12B parameters for maximum compute efficiency and accuracy in complex multi-agent applications. Built on a hybrid Mamba-Transformer...
Llama 3.3 Nemotron Super 49B v1.5Nemotron model for efficient reasoning, coding, and specialized AI agents
NVIDIA: Nemotron 3 Nano Omni (free)NVIDIA Nemotron™ 3 Nano Omni is a 30B-A3B open multimodal model designed to function as a perception and context sub-agent in enterprise agent systems. It accepts text, image, video, and...