No provider description is available for this model yet.
Models
Every model in the catalog with source-linked pricing, context limits, provider availability, and published benchmark results.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
Sakana Namazu is a Japanese-specialized reasoning model from Sakana AI, based on Kimi K2.6 with additional training for Japanese language and business contexts. It is suited for Japanese instruction following,...
No provider description is available for this model yet.
GPT-5.4 is OpenAI’s latest frontier model, unifying the Codex and GPT lines into a single system. It features a 1M+ token context window (922K input, 128K output) with support for...
No provider description is available for this model yet.
Qwen3.8 2.4T A95B is an open-weight sparse mixture-of-experts model from Qwen and the open-weight variant of [Qwen3.8 Max](/qwen/qwen3.8-max), with 95 billion active parameters out of 2.4 trillion total. It is...
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
Qwen 3.6 Plus builds on a hybrid architecture that combines efficient linear attention with sparse mixture-of-experts routing, enabling strong scalability and high-performance inference. Compared to the 3.5 series, it delivers...
No provider description is available for this model yet.
No provider description is available for this model yet.
GPT-3.5 Turbo is OpenAI's fastest model. It can understand and generate natural language or code, and is optimized for chat and traditional completion tasks. Training data up to Sep 2021.
No provider description is available for this model yet.
MoonshotAI: Kimi K2.7 Code is a coding-focused model in Moonshot AI's Kimi K2 family, built to complete end-to-end programming tasks reliably over long contexts. It uses a native multimodal mixture-of-experts...
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
GLM-4.5V is a vision-language foundation model for multimodal agent applications. Built on a Mixture-of-Experts (MoE) architecture with 106B parameters and 12B activated parameters, it achieves state-of-the-art results in video understanding,...
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
| Model | Creator | Inputs | Context | Input | Output | Released | Compare |
|---|---|---|---|---|---|---|---|
| claude-haiku-4-5azure_ai/claude-haiku-4-5 | 200K | $1 | $5 | — | |||
| claude-opus-4-5azure_ai/claude-opus-4-5 | 200K | $5 | $25 | — | |||
| @cf/mistral/mistral-7b-instruct-v0.2-loracloudflare/@cf/mistral/mistral-7b-instruct-v0.2-lora | 15K | — | — | — | |||
| openai/gpt-5.6-terra-proopenrouter/openai/gpt-5.6-terra-pro | 1.05M | $2 | $12 | — | |||
| claude-opus-4-6azure_ai/claude-opus-4-6 | 1M | $5 | $25 | — | |||
| accounts/fireworks/models/qwen2-vl-2b-instructfireworks_ai/accounts/fireworks/models/qwen2-vl-2b-instruct | 32.768K | $0.1 | $0.1 | — | |||
| au.anthropic.claude-opus-4-6-v1bedrock_converse/au.anthropic.claude-opus-4-6-v1 | 1M | $5.5 | $27.5 | — | |||
| openai/gpt-6-astra-proopenrouter/openai/gpt-6-astra-pro | 1.05M | $10 | $50 | — | |||
| claude-fable-5azure_ai/claude-fable-5 | 1M | $10 | $50 | — | |||
| claude-opus-5azure_ai/claude-opus-5 | 1M | $5 | $25 | — | |||
| qwen/qwen3.8-max-0902openrouter/qwen/qwen3.8-max-0902 | 1M | $2 | $6 | — | |||
| claude-opus-4-8azure_ai/claude-opus-4-8 | 1M | $5 | $25 | — | |||
| google/gemini-3-flash-previewgmi/google/gemini-3-flash-preview | 1.04858M | $0.5 | $3 | — | |||
| claude-opus-4-1azure_ai/claude-opus-4-1 | 200K | $15 | $75 | — | |||
| openai/gpt-chat-latestopenrouter/openai/gpt-chat-latest | 400K | $5 | $30 | — | |||
| @cf/moonshotai/kimi-k2.7-codecloudflare/@cf/moonshotai/kimi-k2.7-code | 262.144K | $0.95 | $4 | — | |||
| Sakana: Namazusakana/namazu | 262.144K | $0.95 | $4 | — | |||
| claude-sonnet-4-5azure_ai/claude-sonnet-4-5 | 200K | $3 | $15 | — | |||
| OpenAI: GPT-5.4 (batch)openai/gpt-5.4:batch | 1.05M | $1.25 | $7.5 | — | |||
| google/gemini-3-pro-previewgmi/google/gemini-3-pro-preview | 1.04858M | $2 | $12 | — | |||
| Qwen: Qwen3.8 2.4T A95B (batch)qwen/qwen3.8-2.4t-a95b:batch | 1.01M | $2 | $6 | — | |||
| claude-sonnet-5azure_ai/claude-sonnet-5 | 1M | $2 | $10 | — | |||
| claude-sonnet-4-6azure_ai/claude-sonnet-4-6 | 1M | $3 | $15 | — | |||
| computer-use-previewazure/computer-use-preview | 8.192K | $3 | $12 | — | |||
| Qwen: Qwen3.6 Plusqwen/qwen3.6-plus | 1M | $0.325 | $1.95 | — | |||
| gemini-3.8-flashvertex_ai/gemini-3.8-flash | 1.04858M | $0.75 | $3.75 | — | |||
| accounts/fireworks/models/qwen2-7b-instructfireworks_ai/accounts/fireworks/models/qwen2-7b-instruct | 32.768K | $0.2 | $0.2 | — | |||
| OpenAI: GPT-3.5 Turbo (batch)openai/gpt-3.5-turbo:batch | 16.385K | $0.25 | $0.75 | — | |||
| @cf/meta/llama-3.1-8b-instruct-fp8cloudflare/@cf/meta/llama-3.1-8b-instruct-fp8 | 32K | $0.152 | $0.287 | — | |||
| MoonshotAI: Kimi K2.7 Code (batch)moonshotai/kimi-k2.7-code:batch | 262.144K | $0.95 | $4 | — | |||
| gpt-5.5azure_ai/gpt-5.5 | 1.05M | $5 | $30 | — | |||
| gemini-3.8-flashgemini/gemini-3.8-flash | 1.04858M | $0.75 | $3.75 | — | |||
| us-gov-west-1/claude-sonnet-4-5-20250929-v1:0bedrock/us-gov-west-1/claude-sonnet-4-5-20250929-v1:0 | 200K | $3.6 | $18 | — | |||
| us-gov-west-1/meta.llama3-70b-instruct-v1:0bedrock/us-gov-west-1/meta.llama3-70b-instruct-v1:0 | 8K | $2.65 | $3.5 | — | |||
| us-gov-west-1/meta.llama3-8b-instruct-v1:0bedrock/us-gov-west-1/meta.llama3-8b-instruct-v1:0 | 8K | $0.3 | $2.65 | — | |||
| gemini-3.8-flashvertex_ai-language-models/gemini-3.8-flash | 1.04858M | $0.75 | $3.75 | — | |||
| zai-org/glm-5-maasvertex_ai-zai_models/zai-org/glm-5-maas | 200K | $1 | $3.2 | — | |||
| deepseek-ai/ESFT-token-intent-litedeepseek-ai/ESFT-token-intent-lite | Not documented | — | — | — | |||
| gpt-5.5-2026-04-23azure_ai/gpt-5.5-2026-04-23 | 1.05M | $5 | $30 | — | |||
| us-west-1/meta.llama3-70b-instruct-v1:0bedrock/us-west-1/meta.llama3-70b-instruct-v1:0 | 8.192K | $2.65 | $3.5 | — | |||
| us-gov.anthropic.claude-opus-4-8bedrock_converse/us-gov.anthropic.claude-opus-4-8 | 1M | $6 | $30 | — | |||
| Z.ai: GLM 4.5Vz-ai/glm-4.5v | 65.536K | $0.6 | $1.8 | — | |||
| us-west-1/meta.llama3-8b-instruct-v1:0bedrock/us-west-1/meta.llama3-8b-instruct-v1:0 | 8.192K | $0.3 | $0.6 | — | |||
| us-west-2/1-month-commitment/anthropic.claude-instant-v1bedrock/us-west-2/1-month-commitment/anthropic.claude-instant-v1 | 100K | — | — | — | |||
| gpt-5.4azure_ai/gpt-5.4 | 1.05M | $2.5 | $15 | — | |||
| grok-build-latestxai/grok-build-latest | 500K | $2 | $6 | — | |||
| accounts/fireworks/models/glm-5p3-flashfireworks_ai/accounts/fireworks/models/glm-5p3-flash | 1.04858M | $0.15 | $0.5 | — | |||
| accounts/fireworks/models/inklingfireworks_ai/accounts/fireworks/models/inkling | 1.04858M | $1 | $4.05 | — | |||
| glm-5.2scaleway/glm-5.2 | 256K | $1.8 | $5.5 | — | |||
| deepseek-ai/DeepSeek-V3-0324gmi/deepseek-ai/deepseek-v3-0324 | 163.84K | $0.28 | $0.88 | — |