No provider description is available for this model yet.
Models
Every model in the catalog with source-linked pricing, context limits, provider availability, and published benchmark results.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
OpenAI o4-mini-high is the same model as [o4-mini](/openai/o4-mini) with reasoning_effort set to high. OpenAI o4-mini is a compact reasoning model in the o-series, optimized for fast, cost-efficient performance while retaining...
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
GPT-5-Nano is the smallest and fastest variant in the GPT-5 system, optimized for developer tools, rapid interactions, and ultra-low latency environments. While limited in reasoning depth compared to its larger...
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
| Model | Creator | Inputs | Context | Input | Output | Released | Compare |
|---|---|---|---|---|---|---|---|
| inclusionAI/Ling-3.0-flashdeepinfra/inclusionai/ling-3.0-flash | 131.072K | $0.06 | $0.18 | — | |||
| stepfun-ai/Step-3.7-Flashdeepinfra/stepfun-ai/step-3.7-flash | 262.144K | $0.2 | $1.15 | — | |||
| Qwen/Qwen3.5-35B-A3Bdeepinfra/qwen/qwen3.5-35b-a3b | 262.144K | $0.14 | $1 | — | |||
| ByteDance/Seed-1.8deepinfra/bytedance/seed-1.8 | 256K | $0.25 | $2 | — | |||
| OpenAI: o4 Mini High (batch)openai/o4-mini-high:batch | 200K | $0.55 | $2.2 | — | |||
| tencent/Hy3deepinfra/tencent/hy3 | 262.144K | $0.14 | $0.58 | — | |||
| ByteDance/Seed-2.0-codedeepinfra/bytedance/seed-2.0-code | 256K | $0.5 | $3 | — | |||
| ByteDance/Seed-2.0-prodeepinfra/bytedance/seed-2.0-pro | 256K | $0.5 | $3 | — | |||
| zai-org/GLM-5deepinfra/zai-org/glm-5 | 202.752K | $0.6 | $2.08 | — | |||
| nvidia/Nemotron-3-Nano-30B-A3Bdeepinfra/nvidia/nemotron-3-nano-30b-a3b | 262.144K | $0.05 | $0.2 | — | |||
| moonshotai/Kimi-K2.7-Codedeepinfra/moonshotai/kimi-k2.7-code | 262.144K | $0.68 | $3.4 | — | |||
| anthropic/claude-sonnet-5deepinfra/anthropic/claude-sonnet-5 | 1M | $2 | $10 | — | |||
| mistral-vibe-cli-fastmistral/mistral-vibe-cli-fast | 262.144K | $0.15 | $0.6 | — | |||
| OpenAI: GPT-5 Nano (batch)openai/gpt-5-nano:batch | 400K | $0.025 | $0.2 | — | |||
| Qwen/Qwen3.5-397B-A17Bdeepinfra/qwen/qwen3.5-397b-a17b | 262.144K | $0.45 | $3 | — | |||
| mistral-code-latestmistral/mistral-code-latest | 128K | $0.3 | $0.9 | — | |||
| mistral-code-fim-latestmistral/mistral-code-fim-latest | 128K | $0.3 | $0.9 | — | |||
| mistral-code-agent-latestmistral/mistral-code-agent-latest | 256K | $0.4 | $2 | — | |||
| labs-leanstral-1-5-1mistral/labs-leanstral-1-5-1 | 262.144K | — | — | — | |||
| deepseek-ai/DeepSeek-V4-Flash-0731deepinfra/deepseek-ai/deepseek-v4-flash-0731 | 1.04858M | $0.08 | $0.18 | — |