No provider description is available for this model yet.
Models
Every model in the catalog with source-linked pricing, context limits, provider availability, and published benchmark results.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
NVIDIA Nemotron 3 Ultra is an open frontier-reasoning and orchestration model from NVIDIA, with 55B active parameters out of 550B total (MoE). Built on a hybrid Transformer-Mamba mixture-of-experts architecture, it...
Virtuoso‑Large is Arcee's top‑tier general‑purpose LLM at 72 B parameters, tuned to tackle cross‑domain reasoning, creative writing and enterprise QA. Unlike many 70 B peers, it retains the 128 k...
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
NVIDIA Nemotron 3.5 Lightning is an open mixture-of-experts model from NVIDIA, with 3B active parameters out of 30B total. It is suited for high-throughput agentic workloads and specialized tasks that...
No provider description is available for this model yet.
OpenAI o3-mini is a cost-efficient language model optimized for STEM reasoning tasks, particularly excelling in science, mathematics, and coding. This model supports the `reasoning_effort` parameter, which can be set to...
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
This model always redirects to the latest model in the GPT Terra family.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
Muse Spark 1.3 Contributor is the cost-efficient contributor tier of Meta’s multimodal reasoning model for experimentation, learning, and early-stage agentic, multi-agent, and coding workflows. It is designed to track information...
Claude Opus 4.8 is Anthropic's most capable generally available model in the Opus family. It supports text, image, and file inputs with text output, with reasoning support and a 1M-token...
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
This model always redirects to the latest GLM model from Z.ai.
Qwen3.8 2.4T A95B is an open-weight sparse mixture-of-experts model from Qwen and the open-weight variant of [Qwen3.8 Max](/qwen/qwen3.8-max), with 95 billion active parameters out of 2.4 trillion total. It is...
| Model | Creator | Inputs | Context | Input | Output | Released | Compare |
|---|---|---|---|---|---|---|---|
| google/gemma-4-26B-A4B-itdeepinfra/google/gemma-4-26b-a4b-it | 262.144K | $0.07 | $0.34 | — | |||
| google/gemini-3.1-prodeepinfra/google/gemini-3.1-pro | 1M | $2 | $12 | — | |||
| XiaomiMiMo/MiMo-V2.5-Prodeepinfra/xiaomimimo/mimo-v2.5-pro | 1.04858M | $1 | $3 | — | |||
| anthropic/claude-haiku-4-5deepinfra/anthropic/claude-haiku-4-5 | 200K | $1 | $5 | — | |||
| deepseek-ai/DeepSeek-V4-Flashdeepinfra/deepseek-ai/deepseek-v4-flash | 1.04858M | $0.09 | $0.18 | — | |||
| openai/gpt-oss-120b-Ultradeepinfra/openai/gpt-oss-120b-ultra | 131.072K | $0.2 | $0.95 | — | |||
| Qwen/Qwen3.5-9Bdeepinfra/qwen/qwen3.5-9b | 262.144K | $0.1 | $0.15 | — | |||
| MiniMaxAI/MiniMax-M2.7-Turbodeepinfra/minimaxai/minimax-m2.7-turbo | 196.608K | $0.38 | $1.7 | — | |||
| zai-org/GLM-4.7deepinfra/zai-org/glm-4.7 | 202.752K | $0.4 | $1.75 | — | |||
| google/gemma-4-31B-itdeepinfra/google/gemma-4-31b-it | 262.144K | $0.13 | $0.38 | — | |||
| NVIDIA: Nemotron 3 Ultra (batch)nvidia/nemotron-3-ultra-550b-a55b:batch | 512.288K | $0.6 | $3.6 | — | |||
| Arcee AI: Virtuoso Largearcee-ai/virtuoso-large | 131.072K | $0.75 | $1.2 | — | |||
| anthropic.claude-fable-5-1bedrock_converse/anthropic.claude-fable-5-1 | 1M | $10 | $50 | — | |||
| global.anthropic.claude-fable-5-1bedrock_converse/global.anthropic.claude-fable-5-1 | 1M | $10 | $50 | — | |||
| us.anthropic.claude-fable-5-1bedrock_converse/us.anthropic.claude-fable-5-1 | 1M | $11 | $55 | — | |||
| NVIDIA: Nemotron 3.5 Lightning (free)nvidia/nemotron-3.5-lightning:free | 1M | Free | Free | — | |||
| eu.anthropic.claude-fable-5-1bedrock_converse/eu.anthropic.claude-fable-5-1 | 1M | $11 | $55 | — | |||
| OpenAI: o3 Mini (batch)openai/o3-mini:batch | 200K | $0.55 | $2.2 | — | |||
| claude-fable-5-1azure_ai/claude-fable-5-1 | 1M | $10 | $50 | — | |||
| deepseek-v4-flash-0731azure_ai/deepseek-v4-flash-0731 | 1M | $0.19 | $0.51 | — | |||
| deepseek-v4-flashqwencloud/deepseek-v4-flash | 1M | $0.2 | $0.4 | — | |||
| deepseek-v4-flash-0731qwencloud/deepseek-v4-flash-0731 | 1M | $0.2 | $0.4 | — | |||
| kimi-k3moonshot/kimi-k3 | 1.04858M | $3 | $15 | — | |||
| deepseek-v4-proqwencloud/deepseek-v4-pro | 1M | $2.4 | $4.8 | — | |||
| glm-5.1qwencloud/glm-5.1 | 202.745K | $1.4 | $4.4 | — | |||
| glm-5.2qwencloud/glm-5.2 | 1.04858M | $1.4 | $4.4 | — | |||
| OpenAI: GPT Terra Latest~openai/gpt-terra-latest | 1.05M | $2 | $12 | — | |||
| kimi-k2.7-codeqwencloud/kimi-k2.7-code | 229.376K | $0.95 | $4 | — | |||
| qwen-coderqwencloud/qwen-coder | 1M | $0.3 | $1.5 | — | |||
| qwen-flashqwencloud/qwen-flash | 997.952K | — | — | — | |||
| Meta: Muse Spark 1.3 Contributormeta/muse-spark-1.3-contributor | 1.04858M | $0.1 | $0.2 | — | |||
| Anthropic: Claude Opus 4.8 (batch)anthropic/claude-opus-4.8:batch | 1M | $2.5 | $12.5 | — | |||
| qwen-flash-2025-07-28qwencloud/qwen-flash-2025-07-28 | 997.952K | — | — | — | |||
| qwen-maxqwencloud/qwen-max | 30.72K | $1.6 | $6.4 | — | |||
| qwen-plusqwencloud/qwen-plus | 129.024K | $0.4 | $1.2 | — | |||
| qwen-plus-2025-01-25qwencloud/qwen-plus-2025-01-25 | 129.024K | $0.4 | $1.2 | — | |||
| qwen-plus-2025-04-28qwencloud/qwen-plus-2025-04-28 | 129.024K | $0.4 | $1.2 | — | |||
| qwen-plus-2025-07-14qwencloud/qwen-plus-2025-07-14 | 129.024K | $0.4 | $1.2 | — | |||
| qwen-plus-2025-07-28qwencloud/qwen-plus-2025-07-28 | 997.952K | — | — | — | |||
| qwen-plus-2025-09-11qwencloud/qwen-plus-2025-09-11 | 997.952K | — | — | — | |||
| qwen-plus-latestqwencloud/qwen-plus-latest | 997.952K | — | — | — | |||
| qwen-turboqwencloud/qwen-turbo | 129.024K | $0.05 | $0.2 | — | |||
| qwen-turbo-2024-11-01qwencloud/qwen-turbo-2024-11-01 | 1M | $0.05 | $0.2 | — | |||
| qwen-turbo-2025-04-28qwencloud/qwen-turbo-2025-04-28 | 1M | $0.05 | $0.2 | — | |||
| qwen-turbo-latestqwencloud/qwen-turbo-latest | 1M | $0.05 | $0.2 | — | |||
| qwen3-30b-a3bqwencloud/qwen3-30b-a3b | 129.024K | — | — | — | |||
| qwen3-coder-flashqwencloud/qwen3-coder-flash | 997.952K | — | — | — | |||
| qwen3-coder-flash-2025-07-28qwencloud/qwen3-coder-flash-2025-07-28 | 997.952K | — | — | — | |||
| Z.ai: GLM Latest~z-ai/glm-latest | 1.04858M | $0.9 | $3 | — | |||
| Qwen: Qwen3.8 2.4T A95Bqwen/qwen3.8-2.4t-a95b | 1M | $2 | $6 | — |