DeepSeek-V3 is the latest model from the DeepSeek team, building upon the instruction following and coding abilities of the previous versions. Pre-trained on nearly 15 trillion tokens, the reported evaluations...
8 providers
Every model in the catalog with source-linked pricing, context limits, provider availability, and published benchmark results.
DeepSeek-V3 is the latest model from the DeepSeek team, building upon the instruction following and coding abilities of the previous versions. Pre-trained on nearly 15 trillion tokens, the reported evaluations...
Large open Qwen MoE for multilingual reasoning, coding, and tool use
DeepSeek R1 is here: Performance on par with [OpenAI o1](/openai/o1), but open-sourced and with fully open reasoning tokens. It's 671B parameters in size, with 37B active in an inference pass....
Dense open Qwen model for self-hosted chat, reasoning, and coding
| Model | Creator | Score | Inputs | Context | Input | Output | Released | Compare |
|---|---|---|---|---|---|---|---|---|
| DeepSeek: DeepSeek V3deepseek/deepseek-chat | 70.2 | 128K | $0.257 | $1.029 | 2025-12-01 | |||
| Qwen3 235B-A22Balibaba/qwen3-235b-a22b | 59.6 | 131.072K | $0.7 | $2.8 | 2025-04 | |||
| DeepSeek: R1deepseek/deepseek-r1 | 56.9 | 64K | $0.7 | $2.5 | 2025-01-20 | |||
| Qwen3 32Balibaba/qwen3-32b | 40.0 | 131.072K | $0.7 | $2.8 | 2025-04 |