Claude Opus 4.8 is Anthropic's most capable generally available model in the Opus family. It supports text, image, and file inputs with text output, with reasoning support and a 1M-token...
Models
Every model in the catalog with source-linked pricing, context limits, provider availability, and published benchmark results.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
A cost-efficient version of GPT Audio. The new snapshot features an upgraded decoder for more natural sounding voices and maintains better voice consistency. Input is priced at $0.60 per million...
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
Ling 3.0 Flash Fin is a finance-focused mixture-of-experts model from InclusionAI, built on Ling 3.0 Flash with 5.1B active parameters out of 124B total. It is designed for real-world investment...
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
GLM-4.5-Air is the lightweight variant of our latest flagship model family, also purpose-built for agent-centric applications. Like GLM-4.5, it adopts the Mixture-of-Experts (MoE) architecture but with a more compact parameter...
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
Qwen3-Coder-480B-A35B-Instruct is a Mixture-of-Experts (MoE) code generation model developed by the Qwen team. It is optimized for agentic coding tasks such as function calling, tool use, and long-context reasoning over...
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
GPT-6 Astra Pro is the same underlying model as [GPT-6 Astra](https://openrouter.ai/openai/gpt-6-astra), served with `reasoning.mode` set to `pro` for higher-quality responses on complex tasks. Learn more in OpenAI's docs: https://developers.openai.com/api/docs/guides/reasoning#reasoning-mode
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
Qwen3.8 Max 0902 is an updated snapshot of Qwen3.8 Max from Alibaba's Qwen team. It is a 2.4-trillion-parameter mixture-of-experts model that accepts text, image, and video input and returns text,...
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
| Model | Creator | Inputs | Context | Input | Output | Released | Compare |
|---|---|---|---|---|---|---|---|
| Anthropic: Claude Opus 4.8 (batch)anthropic/claude-opus-4.8:batch | 1M | $2.5 | $12.5 | — | |||
| eu.anthropic.claude-opus-4-7bedrock_converse/eu.anthropic.claude-opus-4-7 | 1M | $5.5 | $27.5 | — | |||
| us-gov-east-1/meta.llama3-70b-instruct-v1:0bedrock/us-gov-east-1/meta.llama3-70b-instruct-v1:0 | 8K | $2.65 | $3.5 | — | |||
| microsoft/Dayhoff-170M-UR90-HL-16000microsoft/Dayhoff-170M-UR90-HL-16000 | Not documented | — | — | — | |||
| microsoft/Dayhoff-170M-UR90-HL-8000microsoft/Dayhoff-170M-UR90-HL-8000 | Not documented | — | — | — | |||
| microsoft/Dayhoff-170M-GR-1000microsoft/Dayhoff-170M-GR-1000 | Not documented | — | — | — | |||
| microsoft/Dayhoff-170M-UR90-HL-4000microsoft/Dayhoff-170M-UR90-HL-4000 | Not documented | — | — | — | |||
| google/gemma-4-E2B-it-qat-w4a16-ctgoogle/gemma-4-E2B-it-qat-w4a16-ct | Not documented | — | — | — | |||
| deepseek-ai/DeepSeek-V3.1baseten/deepseek-ai/deepseek-v3.1 | Not documented | $0.5 | $1.5 | — | |||
| deepseek-ai/DeepSeek-V3-0324baseten/deepseek-ai/deepseek-v3-0324 | Not documented | $0.77 | $0.77 | — | |||
| us-east-1/1-month-commitment/anthropic.claude-instant-v1bedrock/us-east-1/1-month-commitment/anthropic.claude-instant-v1 | 100K | — | — | — | |||
| OpenAI: GPT Audio Miniopenai/gpt-audio-mini | 128K | $0.6 | $2.4 | — | |||
| us-gov-east-1/meta.llama3-8b-instruct-v1:0bedrock/us-gov-east-1/meta.llama3-8b-instruct-v1:0 | 8K | $0.3 | $2.65 | — | |||
| Qwen/Qwen3-VL-235B-A22B-Instruct-FP8gmi/qwen/qwen3-vl-235b-a22b-instruct-fp8 | 262.144K | $0.3 | $1.4 | — | |||
| zai-org/GLM-4.7-FP8gmi/zai-org/glm-4.7-fp8 | 202.752K | $0.4 | $2 | — | |||
| google.gemma-3-12b-itbedrock_converse/google.gemma-3-12b-it | 128K | $0.09 | $0.29 | — | |||
| inclusionAI: Ling 3.0 Flash Fin (free)inclusionai/ling-3.0-flash-fin:free | 262.144K | Free | Free | — | |||
| au.anthropic.claude-opus-4-7bedrock_converse/au.anthropic.claude-opus-4-7 | 1M | $5.5 | $27.5 | — | |||
| gpt-audio-miniazure/gpt-audio-mini | 128K | $0.6 | $2.4 | — | |||
| us-gov-west-1/amazon.nova-pro-v1:0bedrock/us-gov-west-1/amazon.nova-pro-v1:0 | 300K | $0.96 | $3.84 | — | |||
| google/gemma-4-12Bgoogle/gemma-4-12B | Not documented | — | — | — | |||
| microsoft/Mage-Flow-Basemicrosoft/Mage-Flow-Base | Not documented | — | — | — | |||
| GLM-5.2scx-ai/glm-5.2 | 1.04858M | $0.61 | $1.98 | — | |||
| anthropic.claude-fable-5bedrock_converse/anthropic.claude-fable-5 | 1M | $10 | $50 | — | |||
| Qwen3.8-Maxscx-ai/qwen3.8-max | 1M | $1.65 | $4.99 | — | |||
| us-gov-west-1/amazon.titan-text-express-v1bedrock/us-gov-west-1/amazon.titan-text-express-v1 | 42K | $1.3 | $1.7 | — | |||
| google/gemma-4-12B-itgoogle/gemma-4-12B-it | Not documented | — | — | — | |||
| Z.ai: GLM 4.5 Airz-ai/glm-4.5-air | 131.072K | $0.13 | $0.85 | — | |||
| together-ai-8.1b-21btogether_ai/together-ai-8.1b-21b | 1K | $0.3 | $0.3 | — | |||
| us.openai.gpt-5.6-solbedrock_converse/us.openai.gpt-5.6-sol | 1M | $5.5 | $33 | — | |||
| us-gov-west-1/amazon.titan-text-lite-v1bedrock/us-gov-west-1/amazon.titan-text-lite-v1 | 42K | $0.3 | $0.4 | — | |||
| together-ai-81.1b-110btogether_ai/together-ai-81.1b-110b | Not documented | $1.8 | $1.8 | — | |||
| together-ai-up-to-4btogether_ai/together-ai-up-to-4b | Not documented | $0.1 | $0.1 | — | |||
| Qwen/Qwen2.5-72B-Instruct-Turbotogether_ai/qwen/qwen2.5-72b-instruct-turbo | Not documented | — | — | — | |||
| Qwen: Qwen3 Coder 480B A35B (free)qwen/qwen3-coder:free | 262K | Free | Free | — | |||
| nvidia/Cosmos3-Super-Text2Image-4Stepnvidia/Cosmos3-Super-Text2Image-4Step | Not documented | — | — | — | |||
| us-gov-west-1/amazon.titan-text-premier-v1:0bedrock/us-gov-west-1/amazon.titan-text-premier-v1:0 | 42K | $0.5 | $1.5 | — | |||
| deepseek-ai/ESFT-gate-intent-litedeepseek-ai/ESFT-gate-intent-lite | Not documented | — | — | — | |||
| google/gemma-4-12B-it-assistantgoogle/gemma-4-12B-it-assistant | Not documented | — | — | — | |||
| global.openai.gpt-5.6-solbedrock_converse/global.openai.gpt-5.6-sol | 1M | $5 | $30 | — | |||
| OpenAI: GPT-6 Astra Pro (batch)openai/gpt-6-astra-pro:batch | 1.05M | $5 | $25 | — | |||
| @cf/nvidia/nemotron-3-120b-a12bcloudflare/@cf/nvidia/nemotron-3-120b-a12b | 256K | $0.5 | $1.5 | — | |||
| us.openai.gpt-5.6-terrabedrock_converse/us.openai.gpt-5.6-terra | 1M | $2.2 | $13.2 | — | |||
| deepseek-v4-flashdashscope/deepseek-v4-flash | 1M | $0.2 | $0.4 | — | |||
| deepseek-v4-flash-0731dashscope/deepseek-v4-flash-0731 | 1M | $0.2 | $0.4 | — | |||
| global.openai.gpt-5.6-terrabedrock_converse/global.openai.gpt-5.6-terra | 1M | $2 | $12 | — | |||
| Qwen: Qwen3.8 Max (0902)qwen/qwen3.8-max-0902 | 1M | $2 | $6 | — | |||
| google/gemma-4-31B-it-assistantgoogle/gemma-4-31B-it-assistant | Not documented | — | — | — | |||
| us-gov-west-1/anthropic.claude-3-7-sonnet-20250219-v1:0bedrock/us-gov-west-1/anthropic.claude-3-7-sonnet-20250219-v1:0 | 200K | $3.6 | $18 | — | |||
| google/gemma-4-31B-it-qat-w4a16-ctgoogle/gemma-4-31B-it-qat-w4a16-ct | Not documented | — | — | — |