Providers
Provider-specific identifiers, limits, and listed prices per million tokens. Every row links back to the provider's own documentation.
The model record exists, but no source-linked provider offer is available yet.
Submit a sourceDeepSeek R1 Distill Llama 70B is a distilled large language model based on [Llama-3.3-70B-Instruct](/meta-llama/llama-3.3-70b-instruct), using outputs from [DeepSeek R1](/deepseek/deepseek-r1). The model combines advanced distillation techniques to achieve high performance across...
Provider-specific identifiers, limits, and listed prices per million tokens. Every row links back to the provider's own documentation.
The model record exists, but no source-linked provider offer is available yet.
Submit a sourceRecorded from the source catalog and provider listings.
deepseek/deepseek-r1-distill-llama-70bEvery result stays attached to its source, version, metric and harness. Scores from different versions are never merged, and the profile below plots each benchmark against its own population rather than on a shared scale.
ModelBench does not infer quality from price, context size, or model name. When a source publishes a comparable result, it appears here with its version and link.
Read the methodologyField-level changes detected between successful source imports.
Every figure on this page traces back to one of these records.
Fetch the complete source-linked model record. No key, no account, no rate-limited tier.
GET https://model.kyssta.lol/api/v1/models/deepseek/deepseek-r1-distill-llama-70bcurl "https://model.kyssta.lol/api/v1/models/deepseek/deepseek-r1-distill-llama-70b"Answered directly from the stored record — nothing here is generated beyond the catalog's own fields.
DeepSeek R1 Distill Llama 70B is a distilled large language model based on [Llama-3.3-70B-Instruct](/meta-llama/llama-3.3-70b-instruct), using outputs from [DeepSeek R1](/deepseek/deepseek-r1). The model combines advanced distillation techniques to achieve high performance across. It is published by DeepSeek and catalogued here from OpenRouter.
DeepSeek: R1 Distill Llama 70B accepts up to 8.192K tokens of context and returns up to 7.372K output tokens.
No provider description available.
DeepSeek V4 Flash 0423Initial DeepSeek V4 Flash snapshot for economical reasoning, coding, and million-token agent workloads
DeepSeek V4 FlashFast DeepSeek V4 lane for economical reasoning, coding, and long-context work
DeepSeek V4 Pro 0423DeepSeek V4 Pro initial snapshot with million-token context and support for thinking and non-thinking modes