No provider description is available for this model yet.
Models
Every model in the catalog with source-linked pricing, context limits, provider availability, and published benchmark results.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
Sakana Namazu is a Japanese-specialized reasoning model from Sakana AI, based on Kimi K2.6 with additional training for Japanese language and business contexts. It is suited for Japanese instruction following,...
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
Qwen3.8 2.4T A95B is an open-weight sparse mixture-of-experts model from Qwen and the open-weight variant of [Qwen3.8 Max](/qwen/qwen3.8-max), with 95 billion active parameters out of 2.4 trillion total. It is...
No provider description is available for this model yet.
No provider description is available for this model yet.
MoonshotAI: Kimi K2.7 Code is a coding-focused model in Moonshot AI's Kimi K2 family, built to complete end-to-end programming tasks reliably over long contexts. It uses a native multimodal mixture-of-experts...
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
DeepSeek V4 Flash Vision Exp is an experimental vision-enabled version of [DeepSeek V4 Flash 0731](https://openrouter.ai/deepseek/deepseek-v4-flash-0731) from DeepSeek, adding image understanding while matching the base model on text capabilities including agents,...
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
GPT-5.2 Pro is OpenAI’s most advanced model, offering major improvements in agentic coding and long context performance over GPT-5 Pro. It is optimized for complex tasks that require step-by-step reasoning,...
No provider description is available for this model yet.
No provider description is available for this model yet.
Gemini 2.5 Flash-Lite is a lightweight reasoning model in the Gemini 2.5 family, optimized for ultra-low latency and cost efficiency. It offers improved throughput, faster token generation, and better performance...
Llama 3.2 3B is a 3-billion-parameter multilingual large language model, optimized for advanced natural language processing tasks like dialogue generation, reasoning, and summarization. Designed with the latest transformer architecture, it...
No provider description is available for this model yet.
No provider description is available for this model yet.
Seed 2.1 Turbo is a multimodal model from ByteDance Seed for coding and long-horizon agent workflows. It is suited for end-to-end software delivery, multi-step task execution, and understanding visual and...
| Model | Creator | Inputs | Context | Input | Output | Released | Compare |
|---|---|---|---|---|---|---|---|
| claude-haiku-4-5azure_ai/claude-haiku-4-5 | 200K | $1 | $5 | — | |||
| claude-opus-4-5azure_ai/claude-opus-4-5 | 200K | $5 | $25 | — | |||
| openai/gpt-5.6-terra-proopenrouter/openai/gpt-5.6-terra-pro | 1.05M | $2 | $12 | — | |||
| claude-opus-4-6azure_ai/claude-opus-4-6 | 1M | $5 | $25 | — | |||
| claude-opus-4-7azure_ai/claude-opus-4-7 | 1M | $5 | $25 | — | |||
| openai/gpt-6-astra-proopenrouter/openai/gpt-6-astra-pro | 1.05M | $10 | $50 | — | |||
| claude-fable-5azure_ai/claude-fable-5 | 1M | $10 | $50 | — | |||
| claude-opus-5azure_ai/claude-opus-5 | 1M | $5 | $25 | — | |||
| qwen/qwen3.8-max-0902openrouter/qwen/qwen3.8-max-0902 | 1M | $2 | $6 | — | |||
| claude-opus-4-8azure_ai/claude-opus-4-8 | 1M | $5 | $25 | — | |||
| claude-opus-4-1azure_ai/claude-opus-4-1 | 200K | $15 | $75 | — | |||
| openai/gpt-chat-latestopenrouter/openai/gpt-chat-latest | 400K | $5 | $30 | — | |||
| @cf/moonshotai/kimi-k2.7-codecloudflare/@cf/moonshotai/kimi-k2.7-code | 262.144K | $0.95 | $4 | — | |||
| Sakana: Namazusakana/namazu | 262.144K | $0.95 | $4 | — | |||
| claude-sonnet-4-5azure_ai/claude-sonnet-4-5 | 200K | $3 | $15 | — | |||
| claude-sonnet-5azure_ai/claude-sonnet-5 | 1M | $2 | $10 | — | |||
| claude-sonnet-4-6azure_ai/claude-sonnet-4-6 | 1M | $3 | $15 | — | |||
| Qwen: Qwen3.8 2.4T A95B (batch)qwen/qwen3.8-2.4t-a95b:batch | 1.01M | $2 | $6 | — | |||
| gemini-3.8-flashvertex_ai/gemini-3.8-flash | 1.04858M | $0.75 | $3.75 | — | |||
| gpt-oss-120bazure_ai/gpt-oss-120b | 131.072K | $0.15 | $0.6 | — | |||
| MoonshotAI: Kimi K2.7 Code (batch)moonshotai/kimi-k2.7-code:batch | 262.144K | $0.95 | $4 | — | |||
| gpt-5.5azure_ai/gpt-5.5 | 1.05M | $5 | $30 | — | |||
| gemini-3.8-flashgemini/gemini-3.8-flash | 1.04858M | $0.75 | $3.75 | — | |||
| us-gov-west-1/claude-sonnet-4-5-20250929-v1:0bedrock/us-gov-west-1/claude-sonnet-4-5-20250929-v1:0 | 200K | $3.6 | $18 | — | |||
| DeepSeek: DeepSeek V4 Flash Vision Exp (batch)deepseek/deepseek-v4-flash-vision-exp:batch | 1.04858M | $0.11 | $0.33 | — | |||
| gemini-3.8-flashvertex_ai-language-models/gemini-3.8-flash | 1.04858M | $0.75 | $3.75 | — | |||
| us-gov.anthropic.claude-sonnet-5bedrock_converse/us-gov.anthropic.claude-sonnet-5 | 1M | $2.4 | $12 | — | |||
| gpt-5.5-2026-04-23azure_ai/gpt-5.5-2026-04-23 | 1.05M | $5 | $30 | — | |||
| us-gov.anthropic.claude-opus-4-8bedrock_converse/us-gov.anthropic.claude-opus-4-8 | 1M | $6 | $30 | — | |||
| gpt-5.4azure_ai/gpt-5.4 | 1.05M | $2.5 | $15 | — | |||
| grok-build-latestxai/grok-build-latest | 500K | $2 | $6 | — | |||
| accounts/fireworks/models/glm-5p3-flashfireworks_ai/accounts/fireworks/models/glm-5p3-flash | 1.04858M | $0.15 | $0.5 | — | |||
| accounts/fireworks/models/inklingfireworks_ai/accounts/fireworks/models/inkling | 1.04858M | $1 | $4.05 | — | |||
| glm-5.2scaleway/glm-5.2 | 256K | $1.8 | $5.5 | — | |||
| deepseek-v4-flash-0731scaleway/deepseek-v4-flash-0731 | 256K | $0.4 | $0.8 | — | |||
| kimi-k2.7-codeazure_ai/kimi-k2.7-code | 262.144K | $0.95 | $4 | — | |||
| us-gov-west-1/nvidia.nemotron-nano-3-30bbedrock/us-gov-west-1/nvidia.nemotron-nano-3-30b | 262.144K | $0.072 | $0.288 | — | |||
| us-gov-west-1/nvidia.nemotron-nano-12b-v2bedrock/us-gov-west-1/nvidia.nemotron-nano-12b-v2 | 128K | $0.24 | $0.72 | — | |||
| us-gov-west-1/nvidia.nemotron-super-3-120bbedrock/us-gov-west-1/nvidia.nemotron-super-3-120b | 256K | $0.18 | $0.78 | — | |||
| us-gov-west-1/openai.gpt-oss-20b-1:0bedrock/us-gov-west-1/openai.gpt-oss-20b-1:0 | 128K | $0.084 | $0.36 | — | |||
| us-gov-west-1/openai.gpt-oss-120b-1:0bedrock/us-gov-west-1/openai.gpt-oss-120b-1:0 | 128K | $0.18 | $0.72 | — | |||
| us-gov-west-1/anthropic.claude-sonnet-5bedrock/us-gov-west-1/anthropic.claude-sonnet-5 | 1M | $2.4 | $12 | — | |||
| OpenAI: GPT-5.2 Pro (batch)openai/gpt-5.2-pro:batch | 400K | $10.5 | $84 | — | |||
| us-gov-west-1/anthropic.claude-opus-4-8bedrock/us-gov-west-1/anthropic.claude-opus-4-8 | 1M | $6 | $30 | — | |||
| us-gov-east-1/nvidia.nemotron-nano-3-30bbedrock/us-gov-east-1/nvidia.nemotron-nano-3-30b | 262.144K | $0.072 | $0.288 | — | |||
| Google: Gemini 2.5 Flash Lite (batch)google/gemini-2.5-flash-lite:batch | 1.04858M | $0.05 | $0.2 | — | |||
| Meta: Llama 3.2 3B Instruct (free)meta-llama/llama-3.2-3b-instruct:free | 131.072K | Free | Free | — | |||
| gpt-5.4-2026-03-05azure_ai/gpt-5.4-2026-03-05 | 1.05M | $2.5 | $15 | — | |||
| us-gov-east-1/nvidia.nemotron-nano-12b-v2bedrock/us-gov-east-1/nvidia.nemotron-nano-12b-v2 | 128K | $0.24 | $0.72 | — | |||
| ByteDance Seed: Seed 2.1 Turbobytedance-seed/seed-2-1-turbo | 262.144K | $0.5 | $2.5 | — |