No provider description is available for this model yet.
Models
Every model in the catalog with source-linked pricing, context limits, provider availability, and published benchmark results.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
Inflection 3 Productivity is optimized for following instructions. It is better for tasks requiring JSON output or precise adherence to provided guidelines. It has access to recent news. For emotional...
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
Sakana Namazu is a Japanese-specialized reasoning model from Sakana AI, based on Kimi K2.6 with additional training for Japanese language and business contexts. It is suited for Japanese instruction following,...
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
Qwen3.8 2.4T A95B is an open-weight sparse mixture-of-experts model from Qwen and the open-weight variant of [Qwen3.8 Max](/qwen/qwen3.8-max), with 95 billion active parameters out of 2.4 trillion total. It is...
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
GPT-3.5 Turbo is OpenAI's fastest model. It can understand and generate natural language or code, and is optimized for chat and traditional completion tasks. Training data up to Sep 2021.
No provider description is available for this model yet.
MoonshotAI: Kimi K2.7 Code is a coding-focused model in Moonshot AI's Kimi K2 family, built to complete end-to-end programming tasks reliably over long contexts. It uses a native multimodal mixture-of-experts...
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
Grok 4.3 is a reasoning model from SpaceXAI. It accepts text and image inputs with text output, and is suited for agentic workflows, instruction-following tasks, and applications requiring high factual...
No provider description is available for this model yet.
GLM-4.5V is a vision-language foundation model for multimodal agent applications. Built on a Mixture-of-Experts (MoE) architecture with 106B parameters and 12B activated parameters, it achieves state-of-the-art results in video understanding,...
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
| Model | Creator | Inputs | Context | Input | Output | Released | Compare |
|---|---|---|---|---|---|---|---|
| openai/gpt-5.6-terra-proopenrouter/openai/gpt-5.6-terra-pro | 1.05M | $2 | $12 | — | |||
| claude-opus-4-6azure_ai/claude-opus-4-6 | 1M | $5 | $25 | — | |||
| openai/gpt-oss-120b-Ultradeepinfra/openai/gpt-oss-120b-ultra | 131.072K | $0.2 | $0.95 | — | |||
| claude-opus-4-7azure_ai/claude-opus-4-7 | 1M | $5 | $25 | — | |||
| openai/gpt-6-astra-proopenrouter/openai/gpt-6-astra-pro | 1.05M | $10 | $50 | — | |||
| claude-fable-5azure_ai/claude-fable-5 | 1M | $10 | $50 | — | |||
| claude-opus-5azure_ai/claude-opus-5 | 1M | $5 | $25 | — | |||
| qwen/qwen3.8-max-0902openrouter/qwen/qwen3.8-max-0902 | 1M | $2 | $6 | — | |||
| claude-opus-4-8azure_ai/claude-opus-4-8 | 1M | $5 | $25 | — | |||
| Inflection: Inflection 3 Productivityinflection/inflection-3-productivity | 8K | $2.5 | $10 | — | |||
| claude-opus-4-1azure_ai/claude-opus-4-1 | 200K | $15 | $75 | — | |||
| openai/gpt-chat-latestopenrouter/openai/gpt-chat-latest | 400K | $5 | $30 | — | |||
| @cf/moonshotai/kimi-k2.7-codecloudflare/@cf/moonshotai/kimi-k2.7-code | 262.144K | $0.95 | $4 | — | |||
| Sakana: Namazusakana/namazu | 262.144K | $0.95 | $4 | — | |||
| claude-sonnet-4-5azure_ai/claude-sonnet-4-5 | 200K | $3 | $15 | — | |||
| snowflake-llama-3.1-405bsnowflake/snowflake-llama-3.1-405b | 8K | — | — | — | |||
| @cf/deepseek-ai/deepseek-r1-distill-qwen-32bcloudflare/@cf/deepseek-ai/deepseek-r1-distill-qwen-32b | 80K | $0.497 | $4.881 | — | |||
| Qwen: Qwen3.8 2.4T A95B (batch)qwen/qwen3.8-2.4t-a95b:batch | 1.01M | $2 | $6 | — | |||
| claude-sonnet-5azure_ai/claude-sonnet-5 | 1M | $2 | $10 | — | |||
| claude-sonnet-4-6azure_ai/claude-sonnet-4-6 | 1M | $3 | $15 | — | |||
| computer-use-previewazure/computer-use-preview | 8.192K | $3 | $12 | — | |||
| containerazure/container | Not documented | — | — | — | |||
| gemini-3.8-flashvertex_ai/gemini-3.8-flash | 1.04858M | $0.75 | $3.75 | — | |||
| gpt-oss-120bazure_ai/gpt-oss-120b | 131.072K | $0.15 | $0.6 | — | |||
| OpenAI: GPT-3.5 Turbo (batch)openai/gpt-3.5-turbo:batch | 16.385K | $0.25 | $0.75 | — | |||
| @cf/meta/llama-3.1-8b-instruct-fp8cloudflare/@cf/meta/llama-3.1-8b-instruct-fp8 | 32K | $0.152 | $0.287 | — | |||
| MoonshotAI: Kimi K2.7 Code (batch)moonshotai/kimi-k2.7-code:batch | 262.144K | $0.95 | $4 | — | |||
| gpt-5.5azure_ai/gpt-5.5 | 1.05M | $5 | $30 | — | |||
| gemini-3.8-flashgemini/gemini-3.8-flash | 1.04858M | $0.75 | $3.75 | — | |||
| us-gov-west-1/claude-sonnet-4-5-20250929-v1:0bedrock/us-gov-west-1/claude-sonnet-4-5-20250929-v1:0 | 200K | $3.6 | $18 | — | |||
| us-gov-west-1/meta.llama3-70b-instruct-v1:0bedrock/us-gov-west-1/meta.llama3-70b-instruct-v1:0 | 8K | $2.65 | $3.5 | — | |||
| us-gov-west-1/meta.llama3-8b-instruct-v1:0bedrock/us-gov-west-1/meta.llama3-8b-instruct-v1:0 | 8K | $0.3 | $2.65 | — | |||
| gemini-3.8-flashvertex_ai-language-models/gemini-3.8-flash | 1.04858M | $0.75 | $3.75 | — | |||
| us-gov.anthropic.claude-sonnet-5bedrock_converse/us-gov.anthropic.claude-sonnet-5 | 1M | $2.4 | $12 | — | |||
| deepseek-ai/ESFT-token-intent-litedeepseek-ai/ESFT-token-intent-lite | Not documented | — | — | — | |||
| gpt-5.5-2026-04-23azure_ai/gpt-5.5-2026-04-23 | 1.05M | $5 | $30 | — | |||
| us-west-1/meta.llama3-70b-instruct-v1:0bedrock/us-west-1/meta.llama3-70b-instruct-v1:0 | 8.192K | $2.65 | $3.5 | — | |||
| SpaceXAI: Grok 4.3 (batch)x-ai/grok-4.3:batch | 1M | $1 | $2 | — | |||
| us-gov.anthropic.claude-opus-4-8bedrock_converse/us-gov.anthropic.claude-opus-4-8 | 1M | $6 | $30 | — | |||
| Z.ai: GLM 4.5Vz-ai/glm-4.5v | 65.536K | $0.6 | $1.8 | — | |||
| us-west-1/meta.llama3-8b-instruct-v1:0bedrock/us-west-1/meta.llama3-8b-instruct-v1:0 | 8.192K | $0.3 | $0.6 | — | |||
| us-west-2/1-month-commitment/anthropic.claude-instant-v1bedrock/us-west-2/1-month-commitment/anthropic.claude-instant-v1 | 100K | — | — | — | |||
| gpt-5.4azure_ai/gpt-5.4 | 1.05M | $2.5 | $15 | — | |||
| grok-build-latestxai/grok-build-latest | 500K | $2 | $6 | — | |||
| accounts/fireworks/models/glm-5p3-flashfireworks_ai/accounts/fireworks/models/glm-5p3-flash | 1.04858M | $0.15 | $0.5 | — | |||
| accounts/fireworks/models/inklingfireworks_ai/accounts/fireworks/models/inkling | 1.04858M | $1 | $4.05 | — | |||
| glm-5.2scaleway/glm-5.2 | 256K | $1.8 | $5.5 | — | |||
| us-west-2/1-month-commitment/anthropic.claude-v1bedrock/us-west-2/1-month-commitment/anthropic.claude-v1 | 100K | — | — | — | |||
| us-west-2/1-month-commitment/anthropic.claude-v2:1bedrock/us-west-2/1-month-commitment/anthropic.claude-v2:1 | 100K | — | — | — | |||
| llama3.3-70b-instructgradient_ai/llama3.3-70b-instruct | 128K | $0.65 | $0.65 | — |