No provider description is available for this model yet.
Models
Every model in the catalog with source-linked pricing, context limits, provider availability, and published benchmark results.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
Gemini 3.1 Flash Lite is Google’s GA high-efficiency multimodal model optimized for low-latency, high-volume workloads. It supports text, image, video, audio, and PDF inputs, and is designed for lightweight agentic...
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
GLM-5.1 delivers a major leap in coding capability, with particularly significant gains in handling long-horizon tasks. Unlike previous models built around minute-level interactions, GLM-5.1 can work independently and continuously on...
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
Qwen3-VL-235B-A22B Instruct is an open-weight multimodal model that unifies strong text generation with visual understanding across images and video. The Instruct model targets general vision-language use (VQA, document parsing, chart/table...
No provider description is available for this model yet.
| Model | Creator | Inputs | Context | Input | Output | Released | Compare |
|---|---|---|---|---|---|---|---|
| qwen/qwen3.8-27bgroq/qwen/qwen3.8-27b | 131.042K | $0.8 | $4 | — | |||
| mistral-medium-3.5mistral/mistral-medium-3.5 | 262.144K | $1.5 | $7.5 | — | |||
| mistral-vibe-cli-latestmistral/mistral-vibe-cli-latest | 262.144K | $1.5 | $7.5 | — | |||
| accounts/fireworks/models/qwen-v2p5-14b-instructfireworks_ai/accounts/fireworks/models/qwen-v2p5-14b-instruct | 32.768K | $0.2 | $0.2 | — | |||
| anthropic/claude-opus-5deepinfra/anthropic/claude-opus-5 | 1M | $5 | $25 | — | |||
| thinkingmachines/Inklingdeepinfra/thinkingmachines/inkling | 524.288K | $0.95 | $4.05 | — | |||
| moonshotai/Kimi-K2.6deepinfra/moonshotai/kimi-k2.6 | 262.144K | $0.75 | $3.5 | — | |||
| deepseek-ai/DeepSeek-V4-Pro-0813deepinfra/deepseek-ai/deepseek-v4-pro-0813 | 1.04858M | $1.3 | $2.6 | — | |||
| Qwen/Qwen3.7-Maxdeepinfra/qwen/qwen3.7-max | 256K | $2.5 | $7.5 | — | |||
| ByteDance/Seed-2.0-minideepinfra/bytedance/seed-2.0-mini | 256K | $0.1 | $0.4 | — | |||
| mistral-vibe-cli-with-toolsmistral/mistral-vibe-cli-with-tools | 262.144K | $1.5 | $7.5 | — | |||
| Qwen/Qwen3.8-2.4T-A95Bdeepinfra/qwen/qwen3.8-2.4t-a95b | 262.144K | $2 | $6 | — | |||
| accounts/fireworks/models/qwen-qwq-32b-previewfireworks_ai/accounts/fireworks/models/qwen-qwq-32b-preview | 32.768K | $0.9 | $0.9 | — | |||
| MiniMaxAI/MiniMax-M3deepinfra/minimaxai/minimax-m3 | 524.288K | $0.28 | $1.1 | — | |||
| google/gemini-3.1-flash-litedeepinfra/google/gemini-3.1-flash-lite | 1M | $0.25 | $1.5 | — | |||
| google/gemini-3.7-flashdeepinfra/google/gemini-3.7-flash | 1M | $0.75 | $3.75 | — | |||
| inclusionAI/Ling-3.0-flashdeepinfra/inclusionai/ling-3.0-flash | 131.072K | $0.06 | $0.18 | — | |||
| stepfun-ai/Step-3.7-Flashdeepinfra/stepfun-ai/step-3.7-flash | 262.144K | $0.2 | $1.15 | — | |||
| Qwen/Qwen3.5-35B-A3Bdeepinfra/qwen/qwen3.5-35b-a3b | 262.144K | $0.14 | $1 | — | |||
| ByteDance/Seed-1.8deepinfra/bytedance/seed-1.8 | 256K | $0.25 | $2 | — | |||
| qwen3-max-2026-01-23dashscope/qwen3-max-2026-01-23 | 258.048K | — | — | — | |||
| tencent/Hy3deepinfra/tencent/hy3 | 262.144K | $0.14 | $0.58 | — | |||
| ByteDance/Seed-2.0-codedeepinfra/bytedance/seed-2.0-code | 256K | $0.5 | $3 | — | |||
| ByteDance/Seed-2.0-prodeepinfra/bytedance/seed-2.0-pro | 256K | $0.5 | $3 | — | |||
| zai-org/GLM-5deepinfra/zai-org/glm-5 | 202.752K | $0.6 | $2.08 | — | |||
| nvidia/Nemotron-3-Nano-30B-A3Bdeepinfra/nvidia/nemotron-3-nano-30b-a3b | 262.144K | $0.05 | $0.2 | — | |||
| moonshotai/Kimi-K2.7-Codedeepinfra/moonshotai/kimi-k2.7-code | 262.144K | $0.68 | $3.4 | — | |||
| anthropic/claude-sonnet-5deepinfra/anthropic/claude-sonnet-5 | 1M | $2 | $10 | — | |||
| mistral-vibe-cli-fastmistral/mistral-vibe-cli-fast | 262.144K | $0.15 | $0.6 | — | |||
| Google: Gemini 3.1 Flash Lite (batch)google/gemini-3.1-flash-lite:batch | 1.04858M | $0.125 | $0.75 | — | |||
| @cf/nvidia/nemotron-3-120b-a12bcloudflare/@cf/nvidia/nemotron-3-120b-a12b | 256K | $0.5 | $1.5 | — | |||
| Qwen/Qwen3.5-397B-A17Bdeepinfra/qwen/qwen3.5-397b-a17b | 262.144K | $0.45 | $3 | — | |||
| mistral-code-latestmistral/mistral-code-latest | 128K | $0.3 | $0.9 | — | |||
| mistral-code-fim-latestmistral/mistral-code-fim-latest | 128K | $0.3 | $0.9 | — | |||
| mistral-code-agent-latestmistral/mistral-code-agent-latest | 256K | $0.4 | $2 | — | |||
| Z.ai: GLM 5.1z-ai/glm-5.1 | 200K | $0.966 | $3.036 | — | |||
| labs-leanstral-1-5-1mistral/labs-leanstral-1-5-1 | 262.144K | — | — | — | |||
| deepseek-ai/DeepSeek-V4-Flash-0731deepinfra/deepseek-ai/deepseek-v4-flash-0731 | 1.04858M | $0.08 | $0.18 | — | |||
| google/gemma-4-E4B-itdeepinfra/google/gemma-4-e4b-it | 131.072K | $0.02 | $0.1 | — | |||
| deepseek-ai/DeepSeek-V3.2deepinfra/deepseek-ai/deepseek-v3.2 | 163.84K | $0.26 | $0.38 | — | |||
| Qwen/Qwen3.8-Maxdeepinfra/qwen/qwen3.8-max | 256K | $1.65 | $4.951 | — | |||
| anthropic/claude-fable-5deepinfra/anthropic/claude-fable-5 | 1M | $10 | $50 | — | |||
| accounts/fireworks/models/phi-3-vision-128k-instructfireworks_ai/accounts/fireworks/models/phi-3-vision-128k-instruct | 32.064K | $0.2 | $0.2 | — | |||
| nvidia/NVIDIA-Nemotron-3-Ultra-550B-A55Bdeepinfra/nvidia/nvidia-nemotron-3-ultra-550b-a55b | 262.144K | $0.5 | $2.2 | — | |||
| Qwen/Qwen3.5-122B-A10Bdeepinfra/qwen/qwen3.5-122b-a10b | 262.144K | $0.29 | $2.4 | — | |||
| zai-org/GLM-5.1deepinfra/zai-org/glm-5.1 | 202.752K | $1.05 | $3.5 | — | |||
| accounts/fireworks/models/glm-5p3fireworks_ai/accounts/fireworks/models/glm-5p3 | 1.04858M | $1.4 | $4.4 | — | |||
| deepseek-ai/DeepSeek-V4-Prodeepinfra/deepseek-ai/deepseek-v4-pro | 1.04858M | $1.3 | $2.6 | — | |||
| Qwen: Qwen3 VL 235B A22B Instructqwen/qwen3-vl-235b-a22b-instruct | 131.072K | $0.21 | $1.9 | — | |||
| accounts/fireworks/models/phi-3-mini-128k-instructfireworks_ai/accounts/fireworks/models/phi-3-mini-128k-instruct | 131.072K | $0.1 | $0.1 | — |