No provider description is available for this model yet.
Models
Every model in the catalog with source-linked pricing, context limits, provider availability, and published benchmark results.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
GLM-5V-Turbo is Z.ai’s first native multimodal agent foundation model, built for vision-based coding and agent-driven tasks. It natively handles image, video, and text inputs, excels at long-horizon planning, complex coding,...
Seed 2.1 Turbo is a multimodal model from ByteDance Seed for coding and long-horizon agent workflows. It is suited for end-to-end software delivery, multi-step task execution, and understanding visual and...
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
Gemini 2.5 Pro is Google’s state-of-the-art AI model designed for advanced reasoning, coding, mathematics, and scientific tasks. It employs “thinking” capabilities, enabling it to reason through responses with enhanced accuracy...
Llama 3.2 3B is a 3-billion-parameter multilingual large language model, optimized for advanced natural language processing tasks like dialogue generation, reasoning, and summarization. Designed with the latest transformer architecture, it...
Gemini 2.5 Flash-Lite is a lightweight reasoning model in the Gemini 2.5 family, optimized for ultra-low latency and cost efficiency. It offers improved throughput, faster token generation, and better performance...
No provider description is available for this model yet.
No provider description is available for this model yet.
Gemini 3.6 Flash is a high-efficiency model from Google for coding, agentic workflows, and web and app development. It is designed to produce polished outputs with fewer unnecessary edits and...
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
Qwen3.5 Plus (April 2026) is a large-scale multimodal language model from Alibaba. It accepts text, image, and video input and produces text output, with a 1M token context window. This...
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
| Model | Creator | Inputs | Context | Input | Output | Released | Compare |
|---|---|---|---|---|---|---|---|
| global/grok-3azure_ai/global/grok-3 | 131.072K | $3 | $15 | — | |||
| global/grok-3-miniazure_ai/global/grok-3-mini | 131.072K | $0.25 | $1.27 | — | |||
| gpt-5.4-miniazure_ai/gpt-5.4-mini | 272K | $0.75 | $4.5 | — | |||
| grok-3-miniazure_ai/grok-3-mini | 131.072K | $0.25 | $1.27 | — | |||
| gpt-4.1github_copilot/gpt-4.1 | 128K | — | — | — | |||
| grok-4azure_ai/grok-4 | 131.072K | $3 | $15 | — | |||
| grok-4-fast-non-reasoningazure_ai/grok-4-fast-non-reasoning | 131.072K | $0.2 | $0.5 | — | |||
| grok-4-fast-reasoningazure_ai/grok-4-fast-reasoning | 131.072K | $0.2 | $0.5 | — | |||
| us-gov-east-1/nvidia.nemotron-super-3-120bbedrock/us-gov-east-1/nvidia.nemotron-super-3-120b | 256K | $0.18 | $0.78 | — | |||
| Z.ai: GLM 5V Turboz-ai/glm-5v-turbo | 202.752K | $1.2 | $4 | — | |||
| ByteDance Seed: Seed 2.1 Turbobytedance-seed/seed-2-1-turbo | 262.144K | $0.5 | $2.5 | — | |||
| grok-4-1-fast-reasoningazure_ai/grok-4-1-fast-reasoning | 131.072K | $0.2 | $0.5 | — | |||
| grok-code-fast-1azure_ai/grok-code-fast-1 | 131.072K | $0.2 | $1.5 | — | |||
| kimi-k2.5azure_ai/kimi-k2.5 | 262.144K | $0.6 | $3 | — | |||
| kimi-k2.6azure_ai/kimi-k2.6 | 262.144K | $0.95 | $4 | — | |||
| ministral-3bazure_ai/ministral-3b | 128K | $0.04 | $0.04 | — | |||
| us-east-1/minimax.minimax-m2.1bedrock/us-east-1/minimax.minimax-m2.1 | 196K | $0.3 | $1.2 | — | |||
| mistral-large-2407azure_ai/mistral-large-2407 | 128K | $2 | $6 | — | |||
| us-gov-east-1/nvidia.nemotron-nano-12b-v2bedrock/us-gov-east-1/nvidia.nemotron-nano-12b-v2 | 128K | $0.24 | $0.72 | — | |||
| mistral-large-3azure_ai/mistral-large-3 | 256K | $0.5 | $1.5 | — | |||
| mistral-medium-2505azure_ai/mistral-medium-2505 | 131.072K | $0.4 | $2 | — | |||
| mistral-nemoazure_ai/mistral-nemo | 131.072K | $0.15 | $0.15 | — | |||
| gpt-5.4-2026-03-05azure_ai/gpt-5.4-2026-03-05 | 1.05M | $2.5 | $15 | — | |||
| Google: Gemini 2.5 Pro Preview 05-06google/gemini-2.5-pro-preview-05-06 | 1.04858M | $1.25 | $10 | — | |||
| Meta: Llama 3.2 3B Instruct (free)meta-llama/llama-3.2-3b-instruct:free | 131.072K | Free | Free | — | |||
| Google: Gemini 2.5 Flash Lite (batch)google/gemini-2.5-flash-lite:batch | 1.04858M | $0.05 | $0.2 | — | |||
| ap-northeast-1/deepseek.v3.2bedrock/ap-northeast-1/deepseek.v3.2 | 163.84K | $0.74 | $2.22 | — | |||
| ap-northeast-1/minimax.minimax-m2.1bedrock/ap-northeast-1/minimax.minimax-m2.1 | 196K | $0.36 | $1.44 | — | |||
| Google: Gemini 3.6 Flash (batch)google/gemini-3.6-flash:batch | 1.04858M | $0.375 | $1.875 | — | |||
| us-gov-east-1/anthropic.claude-opus-5bedrock/us-gov-east-1/anthropic.claude-opus-5 | 1M | $6 | $30 | — | |||
| ap-northeast-1/minimax.minimax-m2.5bedrock/ap-northeast-1/minimax.minimax-m2.5 | 1M | $0.36 | $1.44 | — | |||
| ap-northeast-1/moonshotai.kimi-k2-thinkingbedrock/ap-northeast-1/moonshotai.kimi-k2-thinking | 262.144K | $0.73 | $3.03 | — | |||
| Qwen: Qwen3.5 Plus 2026-04-20qwen/qwen3.5-plus-20260420 | 1M | $0.3 | $1.8 | — | |||
| us-gov-east-1/nvidia.nemotron-nano-3-30bbedrock/us-gov-east-1/nvidia.nemotron-nano-3-30b | 262.144K | $0.072 | $0.288 | — | |||
| us-gov-west-1/anthropic.claude-opus-4-8bedrock/us-gov-west-1/anthropic.claude-opus-4-8 | 1M | $6 | $30 | — | |||
| moonshotai.kimi-k2-thinkingbedrock/moonshotai.kimi-k2-thinking | 262.144K | $0.73 | $3.03 | — | |||
| accounts/fireworks/models/nemotron-3-ultra-nvfp4fireworks_ai/accounts/fireworks/models/nemotron-3-ultra-nvfp4 | 262.144K | $0.6 | $2.4 | — | |||
| us-gov-west-1/anthropic.claude-sonnet-5bedrock/us-gov-west-1/anthropic.claude-sonnet-5 | 1M | $2.4 | $12 | — | |||
| us-gov-west-1/openai.gpt-oss-120b-1:0bedrock/us-gov-west-1/openai.gpt-oss-120b-1:0 | 128K | $0.18 | $0.72 | — | |||
| ap-south-1/minimax.minimax-m2.1bedrock/ap-south-1/minimax.minimax-m2.1 | 196K | $0.36 | $1.44 | — | |||
| claude-4-opus-20250514anthropic/claude-4-opus-20250514 | 200K | $15 | $75 | — | |||
| claude-4-sonnet-20250514anthropic/claude-4-sonnet-20250514 | 1M | $3 | $15 | — | |||
| grok-4.3azure_ai/grok-4.3 | 200K | $1.25 | $2.5 | — | |||
| claude-opus-4-6-20260205anthropic/claude-opus-4-6-20260205 | 1M | $5 | $25 | — | |||
| claude-opus-4-7-20260416anthropic/claude-opus-4-7-20260416 | 1M | $5 | $25 | — | |||
| @cf/google/gemma-4-26b-a4b-itcloudflare/@cf/google/gemma-4-26b-a4b-it | 256K | $0.1 | $0.3 | — | |||
| @cf/mistralai/mistral-small-3.1-24b-instructcloudflare/@cf/mistralai/mistral-small-3.1-24b-instruct | 128K | $0.351 | $0.555 | — | |||
| @cf/meta/llama-3.2-11b-vision-instructcloudflare/@cf/meta/llama-3.2-11b-vision-instruct | 128K | $0.048 | $0.676 | — | |||
| @cf/openai/gpt-oss-20bcloudflare/@cf/openai/gpt-oss-20b | 128K | $0.2 | $0.3 | — | |||
| us-gov-west-1/openai.gpt-oss-20b-1:0bedrock/us-gov-west-1/openai.gpt-oss-20b-1:0 | 128K | $0.084 | $0.36 | — |