No provider description is available for this model yet.
Models
Every model in the catalog with source-linked pricing, context limits, provider availability, and published benchmark results.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
Qwen 3.6 Plus builds on a hybrid architecture that combines efficient linear attention with sparse mixture-of-experts routing, enabling strong scalability and high-performance inference. Compared to the 3.5 series, it delivers...
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
Hy-MT2-30B-A3B is Tencent's flagship translation model in the Hy-MT2 family. It supports 33 language pairs and five Chinese dialect and minority-language pairs, with workflows for structured, delimiter-based, contextual, glossary-based, and...
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
Gemma 4 31B Instruct is Google DeepMind's 30.7B dense multimodal model supporting text and image input with text output. Features a 256K token context window, configurable thinking/reasoning mode, native function...
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
| Model | Creator | Inputs | Context | Input | Output | Released | Compare |
|---|---|---|---|---|---|---|---|
| meta-llama/Meta-Llama-3.1-70B-Instructdeepinfra/meta-llama/meta-llama-3.1-70b-instruct | 131.072K | $0.4 | $0.4 | — | |||
| meta-llama/Meta-Llama-3.1-70B-Instruct-Turbodeepinfra/meta-llama/meta-llama-3.1-70b-instruct-turbo | 131.072K | $0.1 | $0.28 | — | |||
| meta-llama/Meta-Llama-3.1-8B-Instructdeepinfra/meta-llama/meta-llama-3.1-8b-instruct | 131.072K | $0.03 | $0.05 | — | |||
| meta-llama/Meta-Llama-3.1-8B-Instruct-Turbodeepinfra/meta-llama/meta-llama-3.1-8b-instruct-turbo | 131.072K | $0.02 | $0.03 | — | |||
| apac.anthropic.claude-sonnet-4-20250514-v1:0bedrock_converse/apac.anthropic.claude-sonnet-4-20250514-v1:0 | 1M | $3 | $15 | — | |||
| microsoft/phi-4deepinfra/microsoft/phi-4 | 16.384K | $0.07 | $0.14 | — | |||
| mistralai/Mistral-Nemo-Instruct-2407deepinfra/mistralai/mistral-nemo-instruct-2407 | 131.072K | $0.02 | $0.04 | — | |||
| mistralai/Mistral-Small-24B-Instruct-2501deepinfra/mistralai/mistral-small-24b-instruct-2501 | 32.768K | $0.05 | $0.08 | — | |||
| mistralai/Mistral-Small-3.2-24B-Instruct-2506deepinfra/mistralai/mistral-small-3.2-24b-instruct-2506 | 128K | $0.075 | $0.2 | — | |||
| mistralai/Mixtral-8x7B-Instruct-v0.1deepinfra/mistralai/mixtral-8x7b-instruct-v0.1 | 32.768K | $0.4 | $0.4 | — | |||
| us/o1-preview-2024-09-12azure/us/o1-preview-2024-09-12 | 128K | $16.5 | $66 | — | |||
| moonshotai/Kimi-K2-Instruct-0905deepinfra/moonshotai/kimi-k2-instruct-0905 | 262.144K | $0.5 | $2 | — | |||
| Qwen: Qwen3.6 Plusqwen/qwen3.6-plus | 1M | $0.325 | $1.95 | — | |||
| nvidia/Llama-3.3-Nemotron-Super-49B-v1.5deepinfra/nvidia/llama-3.3-nemotron-super-49b-v1.5 | 131.072K | $0.1 | $0.4 | — | |||
| nvidia/NVIDIA-Nemotron-Nano-9B-v2deepinfra/nvidia/nvidia-nemotron-nano-9b-v2 | 131.072K | $0.04 | $0.16 | — | |||
| openai/gpt-oss-120bdeepinfra/openai/gpt-oss-120b | 131.072K | $0.05 | $0.45 | — | |||
| apac.anthropic.claude-3-sonnet-20240229-v1:0bedrock/apac.anthropic.claude-3-sonnet-20240229-v1:0 | 200K | $3 | $15 | — | |||
| zai-org/GLM-4.5deepinfra/zai-org/glm-4.5 | 131.072K | $0.4 | $1.6 | — | |||
| deepseek-coderdeepseek/deepseek-coder | 128K | $0.14 | $0.28 | — | |||
| Tencent: Hy-MT2-30B-A3Btencent/hy-mt2-30b-a3b | 8.192K | $0.074 | $0.295 | — | |||
| deepseek.v3-v1:0bedrock_converse/deepseek.v3-v1:0 | 163.84K | $0.58 | $1.68 | — | |||
| deepseek.v3.2bedrock_converse/deepseek.v3.2 | 163.84K | $0.62 | $1.85 | — | |||
| @cf/qwen/qwen2.5-coder-32b-instructcloudflare/@cf/qwen/qwen2.5-coder-32b-instruct | 32.768K | $0.66 | $1 | — | |||
| deepseek-v3-2-251201volcengine/deepseek-v3-2-251201 | 98.304K | — | — | — | |||
| glm-4-7-251222volcengine/glm-4-7-251222 | 204.8K | — | — | — | |||
| kimi-k2-thinking-251104volcengine/kimi-k2-thinking-251104 | 229.376K | — | — | — | |||
| Google: Gemma 4 31B (batch)google/gemma-4-31b-it:batch | 262.144K | $0.39 | $0.97 | — | |||
| eu.amazon.nova-micro-v1:0bedrock_converse/eu.amazon.nova-micro-v1:0 | 128K | $0.046 | $0.184 | — | |||
| zai-glm-4.7cerebras/zai-glm-4.7 | 128K | $2.25 | $2.75 | — | |||
| eu.anthropic.claude-3-5-haiku-20241022-v1:0bedrock/eu.anthropic.claude-3-5-haiku-20241022-v1:0 | 200K | $0.25 | $1.25 | — | |||
| mistral-7bsnowflake/mistral-7b | 32K | — | — | — | |||
| eu.anthropic.claude-3-5-sonnet-20240620-v1:0bedrock/eu.anthropic.claude-3-5-sonnet-20240620-v1:0 | 200K | $3 | $15 | — | |||
| gpt-4.1-2025-04-14github_copilot/gpt-4.1-2025-04-14 | 128K | — | — | — | |||
| eu.anthropic.claude-3-7-sonnet-20250219-v1:0bedrock/eu.anthropic.claude-3-7-sonnet-20250219-v1:0 | 200K | $3 | $15 | — | |||
| eu.anthropic.claude-3-haiku-20240307-v1:0bedrock/eu.anthropic.claude-3-haiku-20240307-v1:0 | 200K | $0.25 | $1.25 | — | |||
| eu.anthropic.claude-3-opus-20240229-v1:0bedrock/eu.anthropic.claude-3-opus-20240229-v1:0 | 200K | $15 | $75 | — | |||
| eu.anthropic.claude-3-sonnet-20240229-v1:0bedrock/eu.anthropic.claude-3-sonnet-20240229-v1:0 | 200K | $3 | $15 | — | |||
| eu.anthropic.claude-opus-4-1-20250805-v1:0bedrock_converse/eu.anthropic.claude-opus-4-1-20250805-v1:0 | 200K | $15 | $75 | — | |||
| kimi-k3-usfireworks_ai/kimi-k3-us | 1.04858M | $3.3 | $16.5 | — | |||
| accounts/fireworks/models/deepseek-coder-v2-instructfireworks_ai/accounts/fireworks/models/deepseek-coder-v2-instruct | 65.536K | $1.2 | $1.2 | — | |||
| eu.anthropic.claude-sonnet-4-20250514-v1:0bedrock_converse/eu.anthropic.claude-sonnet-4-20250514-v1:0 | 1M | $3 | $15 | — | |||
| eu.anthropic.claude-sonnet-4-5-20250929-v1:0bedrock_converse/eu.anthropic.claude-sonnet-4-5-20250929-v1:0 | 200K | $3.3 | $16.5 | — | |||
| eu.meta.llama3-2-1b-instruct-v1:0bedrock/eu.meta.llama3-2-1b-instruct-v1:0 | 128K | $0.13 | $0.13 | — | |||
| eu.meta.llama3-2-3b-instruct-v1:0bedrock/eu.meta.llama3-2-3b-instruct-v1:0 | 128K | $0.19 | $0.19 | — | |||
| eu.mistral.pixtral-large-2502-v1:0bedrock_converse/eu.mistral.pixtral-large-2502-v1:0 | 128K | $2 | $6 | — | |||
| featherless-ai/Qwerky-72Bfeatherless_ai/featherless-ai/qwerky-72b | 32.768K | — | — | — | |||
| featherless-ai/Qwerky-QwQ-32Bfeatherless_ai/featherless-ai/qwerky-qwq-32b | 32.768K | — | — | — | |||
| accounts/fireworks/models/deepseek-r1fireworks_ai/accounts/fireworks/models/deepseek-r1 | 128K | $3 | $8 | — | |||
| accounts/fireworks/models/deepseek-r1-0528fireworks_ai/accounts/fireworks/models/deepseek-r1-0528 | 160K | $3 | $8 | — | |||
| us/o1-mini-2024-09-12azure/us/o1-mini-2024-09-12 | 128K | $1.21 | $4.84 | — |