No provider description is available for this model yet.
Models
Every model in the catalog with source-linked pricing, context limits, provider availability, and published benchmark results.
No provider description is available for this model yet.
The Meta Llama 3.3 multilingual large language model (LLM) is a pretrained and instruction tuned generative model in 70B (text in/text out). The Llama 3.3 instruction tuned text only model...
GLM 5.2 is a large-scale reasoning model from Z.ai. It supports text input and output with a 1M-token context window, and is suited for long-horizon agent workflows, project-level software engineering,...
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
Grok 4.3 is a reasoning model from SpaceXAI. It accepts text and image inputs with text output, and is suited for agentic workflows, instruction-following tasks, and applications requiring high factual...
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
GPT-4o Search Previewis a specialized model for web search in Chat Completions. It is trained to understand and execute web search queries.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
As a 30B-class SOTA model, GLM-4.7-Flash offers a new option that balances performance and efficiency. It is further optimized for agentic coding use cases, strengthening coding capabilities, long-horizon task planning,...
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
| Model | Creator | Inputs | Context | Input | Output | Released | Compare |
|---|---|---|---|---|---|---|---|
| us-gov-east-1/openai.gpt-oss-20bbedrock_mantle/us-gov-east-1/openai.gpt-oss-20b | 131.072K | $0.084 | $0.36 | — | |||
| global.anthropic.claude-opus-4-8bedrock_converse/global.anthropic.claude-opus-4-8 | 1M | $5 | $25 | — | |||
| Meta: Llama 3.3 70B Instruct (free)meta-llama/llama-3.3-70b-instruct:free | 65.536K | Free | Free | — | |||
| Z.ai: GLM 5.2 (free)z-ai/glm-5.2:free | 256K | Free | Free | — | |||
| us.anthropic.claude-opus-4-8bedrock_converse/us.anthropic.claude-opus-4-8 | 1M | $5.5 | $27.5 | — | |||
| eu.anthropic.claude-opus-4-8bedrock_converse/eu.anthropic.claude-opus-4-8 | 1M | $5.5 | $27.5 | — | |||
| @cf/mistral/mistral-7b-instruct-v0.1cloudflare/@cf/mistral/mistral-7b-instruct-v0.1 | 8.192K | $1.923 | $1.923 | — | |||
| au.anthropic.claude-opus-4-8bedrock_converse/au.anthropic.claude-opus-4-8 | 1M | $5.5 | $27.5 | — | |||
| moonshotai/kimi-k2.6novita/moonshotai/kimi-k2.6 | 262.144K | $0.8 | $3.4 | — | |||
| jp.anthropic.claude-opus-4-8bedrock_converse/jp.anthropic.claude-opus-4-8 | 1M | $5.5 | $27.5 | — | |||
| glm-4.6zai/glm-4.6 | 200K | $0.6 | $2.2 | — | |||
| jp.anthropic.claude-opus-4-7bedrock_converse/jp.anthropic.claude-opus-4-7 | 1M | $5.5 | $27.5 | — | |||
| anthropic.claude-sonnet-5bedrock_converse/anthropic.claude-sonnet-5 | 1M | $2 | $10 | — | |||
| global.anthropic.claude-sonnet-5bedrock_converse/global.anthropic.claude-sonnet-5 | 1M | $2 | $10 | — | |||
| us-gov-east-1/xai.grok-4.6bedrock_mantle/us-gov-east-1/xai.grok-4.6 | 500K | $2.64 | $7.92 | — | |||
| SpaceXAI: Grok 4.3 (batch)x-ai/grok-4.3:batch | 1M | $1 | $2 | — | |||
| us.anthropic.claude-sonnet-5bedrock_converse/us.anthropic.claude-sonnet-5 | 1M | $2.2 | $11 | — | |||
| eu.anthropic.claude-sonnet-5bedrock_converse/eu.anthropic.claude-sonnet-5 | 1M | $2.2 | $11 | — | |||
| au.anthropic.claude-sonnet-5bedrock_converse/au.anthropic.claude-sonnet-5 | 1M | $2.2 | $11 | — | |||
| jp.anthropic.claude-sonnet-5bedrock_converse/jp.anthropic.claude-sonnet-5 | 1M | $2.2 | $11 | — | |||
| anthropic.claude-sonnet-4-6bedrock_converse/anthropic.claude-sonnet-4-6 | 1M | $3 | $15 | — | |||
| global.anthropic.claude-sonnet-4-6bedrock_converse/global.anthropic.claude-sonnet-4-6 | 1M | $3 | $15 | — | |||
| us.anthropic.claude-sonnet-4-6bedrock_converse/us.anthropic.claude-sonnet-4-6 | 1M | $3.3 | $16.5 | — | |||
| @hf/thebloke/codellama-7b-instruct-awqcloudflare/@hf/thebloke/codellama-7b-instruct-awq | 4.096K | $1.923 | $1.923 | — | |||
| us-gov-west-1/openai.gpt-oss-120bbedrock_mantle/us-gov-west-1/openai.gpt-oss-120b | 131.072K | $0.18 | $0.72 | — | |||
| OpenAI: GPT-4o Search Previewopenai/gpt-4o-search-preview | 128K | $2.5 | $10 | — | |||
| @cf/openai/gpt-oss-120bcloudflare/@cf/openai/gpt-oss-120b | 128K | $0.35 | $0.75 | — | |||
| eu.anthropic.claude-sonnet-4-6bedrock_converse/eu.anthropic.claude-sonnet-4-6 | 1M | $3.3 | $16.5 | — | |||
| au.anthropic.claude-sonnet-4-6bedrock_converse/au.anthropic.claude-sonnet-4-6 | 1M | $3.3 | $16.5 | — | |||
| qwen-3.8-27bcerebras/qwen-3.8-27b | 65.536K | $0.99 | $1.49 | — | |||
| jp.anthropic.claude-sonnet-4-6bedrock_converse/jp.anthropic.claude-sonnet-4-6 | 1M | $3.3 | $16.5 | — | |||
| computer-use-previewopenai/computer-use-preview | 8.192K | $3 | $12 | — | |||
| anthropic.claude-sonnet-4-20250514-v1:0bedrock_converse/anthropic.claude-sonnet-4-20250514-v1:0 | 1M | $3 | $15 | — | |||
| anthropic.claude-sonnet-4-5-20250929-v1:0bedrock_converse/anthropic.claude-sonnet-4-5-20250929-v1:0 | 200K | $3 | $15 | — | |||
| anthropic.claude-v1bedrock/anthropic.claude-v1 | 100K | $8 | $24 | — | |||
| anthropic.claude-v2:1bedrock/anthropic.claude-v2:1 | 100K | $8 | $24 | — | |||
| HuggingFaceH4/zephyr-7b-betaanyscale/huggingfaceh4/zephyr-7b-beta | 16.384K | $0.15 | $0.15 | — | |||
| deepseek/deepseek-v4.1-flashopenrouter/deepseek/deepseek-v4.1-flash | 1.04858M | $0.15 | $0.6 | — | |||
| openai/gpt-5.6-sol-proopenrouter/openai/gpt-5.6-sol-pro | 1.05M | $2 | $10 | — | |||
| deepseek-flashdeepseek/deepseek-flash | 1M | $0.3 | $1.2 | — | |||
| codellama/CodeLlama-34b-Instruct-hfanyscale/codellama/codellama-34b-instruct-hf | 4.096K | $1 | $1 | — | |||
| codellama/CodeLlama-70b-Instruct-hfanyscale/codellama/codellama-70b-instruct-hf | 4.096K | $1 | $1 | — | |||
| google/gemma-7b-itanyscale/google/gemma-7b-it | 8.192K | $0.15 | $0.15 | — | |||
| google/medgemma-1.5-4b-itgoogle/medgemma-1.5-4b-it | Not documented | — | — | — | |||
| meta-llama/Llama-2-13b-chat-hfanyscale/meta-llama/llama-2-13b-chat-hf | 4.096K | $0.25 | $0.25 | — | |||
| Z.ai: GLM 4.7 Flashz-ai/glm-4.7-flash | 131.072K | $0.061 | $0.4 | — | |||
| accounts/fireworks/models/deepseek-v4p1-flashfireworks_ai/accounts/fireworks/models/deepseek-v4p1-flash | 1.04858M | $0.22 | $0.66 | — | |||
| qwen/qwen3.6-27bnovita/qwen/qwen3.6-27b | 262.144K | $0.6 | $3.6 | — | |||
| meta-llama/Llama-2-70b-chat-hfanyscale/meta-llama/llama-2-70b-chat-hf | 4.096K | $1 | $1 | — | |||
| us-gov-west-1/openai.gpt-oss-20bbedrock_mantle/us-gov-west-1/openai.gpt-oss-20b | 131.072K | $0.084 | $0.36 | — |