No provider description is available for this model yet.
Models
Every model in the catalog with source-linked pricing, context limits, provider availability, and published benchmark results.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
This model always redirects to the latest model in the GLM Flash family.
No provider description is available for this model yet.
No provider description is available for this model yet.
The o-series of models are trained with reinforcement learning to think before they answer and perform complex reasoning. The o3-pro model uses more compute to think harder and provide consistently...
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
Gemini 2.5 Flash-Lite is a lightweight reasoning model in the Gemini 2.5 family, optimized for ultra-low latency and cost efficiency. It offers improved throughput, faster token generation, and better performance...
No provider description is available for this model yet.
Llama 3.2 3B is a 3-billion-parameter multilingual large language model, optimized for advanced natural language processing tasks like dialogue generation, reasoning, and summarization. Designed with the latest transformer architecture, it...
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
Seed 2.1 Turbo is a multimodal model from ByteDance Seed for coding and long-horizon agent workflows. It is suited for end-to-end software delivery, multi-step task execution, and understanding visual and...
No provider description is available for this model yet.
No provider description is available for this model yet.
GPT-5.4 is OpenAI’s latest frontier model, unifying the Codex and GPT lines into a single system. It features a 1M+ token context window (922K input, 128K output) with support for...
No provider description is available for this model yet.
No provider description is available for this model yet.
Kimi K3 is a 2.8T parameter open-weight multimodal reasoning model from Moonshot AI. It is suited for complex coding, knowledge work, and long-horizon agentic workflows, and is particularly strong at...
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
Coder‑Large is a 32 B‑parameter offspring of Qwen 2.5‑Instruct that has been further trained on permissively‑licensed GitHub, CodeSearchNet and synthetic bug‑fix corpora. It supports a 32k context window, enabling multi‑file...
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
| Model | Creator | Inputs | Context | Input | Output | Released | Compare |
|---|---|---|---|---|---|---|---|
| grok-build-latestxai/grok-build-latest | 500K | $2 | $6 | — | |||
| accounts/fireworks/models/glm-5p3-flashfireworks_ai/accounts/fireworks/models/glm-5p3-flash | 1.04858M | $0.15 | $0.5 | — | |||
| accounts/fireworks/models/inklingfireworks_ai/accounts/fireworks/models/inkling | 1.04858M | $1 | $4.05 | — | |||
| glm-5.2scaleway/glm-5.2 | 256K | $1.8 | $5.5 | — | |||
| Z.ai: GLM Flash Latest~z-ai/glm-flash-latest | 1.04858M | $0.075 | $0.25 | — | |||
| us-west-2/1-month-commitment/anthropic.claude-v2:1bedrock/us-west-2/1-month-commitment/anthropic.claude-v2:1 | 100K | — | — | — | |||
| deepseek-v4-flash-0731scaleway/deepseek-v4-flash-0731 | 256K | $0.4 | $0.8 | — | |||
| OpenAI: o3 Pro (batch)openai/o3-pro:batch | 200K | $10 | $40 | — | |||
| us-west-2/6-month-commitment/anthropic.claude-instant-v1bedrock/us-west-2/6-month-commitment/anthropic.claude-instant-v1 | 100K | — | — | — | |||
| kimi-k2.7-codeazure_ai/kimi-k2.7-code | 262.144K | $0.95 | $4 | — | |||
| accounts/fireworks/models/qwen2p5-coder-32b-instruct-128kfireworks_ai/accounts/fireworks/models/qwen2p5-coder-32b-instruct-128k | 131.072K | $0.9 | $0.9 | — | |||
| us-west-2/6-month-commitment/anthropic.claude-v1bedrock/us-west-2/6-month-commitment/anthropic.claude-v1 | 100K | — | — | — | |||
| claude-3-5-haiku@20241022vertex_ai-anthropic_models/claude-3-5-haiku@20241022 | 200K | $1 | $5 | — | |||
| us-gov-west-1/nvidia.nemotron-nano-3-30bbedrock/us-gov-west-1/nvidia.nemotron-nano-3-30b | 262.144K | $0.072 | $0.288 | — | |||
| us-gov-east-1/anthropic.claude-sonnet-4-5-20250929-v1:0bedrock/us-gov-east-1/anthropic.claude-sonnet-4-5-20250929-v1:0 | 200K | $3.6 | $18 | — | |||
| us-gov-west-1/nvidia.nemotron-super-3-120bbedrock/us-gov-west-1/nvidia.nemotron-super-3-120b | 256K | $0.18 | $0.78 | — | |||
| us.anthropic.claude-opus-4-7bedrock_converse/us.anthropic.claude-opus-4-7 | 1M | $5.5 | $27.5 | — | |||
| us-gov-west-1/openai.gpt-oss-20b-1:0bedrock/us-gov-west-1/openai.gpt-oss-20b-1:0 | 128K | $0.084 | $0.36 | — | |||
| us-gov-west-1/openai.gpt-oss-120b-1:0bedrock/us-gov-west-1/openai.gpt-oss-120b-1:0 | 128K | $0.18 | $0.72 | — | |||
| us-gov-west-1/anthropic.claude-sonnet-5bedrock/us-gov-west-1/anthropic.claude-sonnet-5 | 1M | $2.4 | $12 | — | |||
| global.anthropic.claude-opus-4-7bedrock_converse/global.anthropic.claude-opus-4-7 | 1M | $5 | $25 | — | |||
| us-west-2/anthropic.claude-v1bedrock/us-west-2/anthropic.claude-v1 | 100K | $8 | $24 | — | |||
| chatgpt-4o-latestopenai/chatgpt-4o-latest | 128K | $5 | $15 | — | |||
| Google: Gemini 2.5 Flash Lite (batch)google/gemini-2.5-flash-lite:batch | 1.04858M | $0.05 | $0.2 | — | |||
| us-gov-east-1/nvidia.nemotron-nano-3-30bbedrock/us-gov-east-1/nvidia.nemotron-nano-3-30b | 262.144K | $0.072 | $0.288 | — | |||
| Meta: Llama 3.2 3B Instruct (free)meta-llama/llama-3.2-3b-instruct:free | 131.072K | Free | Free | — | |||
| us-west-2/mistral.mistral-7b-instruct-v0:2bedrock/us-west-2/mistral.mistral-7b-instruct-v0:2 | 32K | $0.15 | $0.2 | — | |||
| gpt-5.4-2026-03-05azure_ai/gpt-5.4-2026-03-05 | 1.05M | $2.5 | $15 | — | |||
| us-gov-east-1/nvidia.nemotron-nano-12b-v2bedrock/us-gov-east-1/nvidia.nemotron-nano-12b-v2 | 128K | $0.24 | $0.72 | — | |||
| us-west-2/mistral.mistral-large-2402-v1:0bedrock/us-west-2/mistral.mistral-large-2402-v1:0 | 32K | $8 | $24 | — | |||
| qwen3-coder-plusqwencloud/qwen3-coder-plus | 997.952K | — | — | — | |||
| us-west-2/mistral.mixtral-8x7b-instruct-v0:1bedrock/us-west-2/mistral.mixtral-8x7b-instruct-v0:1 | 32K | $0.45 | $0.7 | — | |||
| ByteDance Seed: Seed 2.1 Turbobytedance-seed/seed-2-1-turbo | 262.144K | $0.5 | $2.5 | — | |||
| gpt-5.4-miniazure_ai/gpt-5.4-mini | 272K | $0.75 | $4.5 | — | |||
| accounts/fireworks/models/qwen2p5-coder-32bfireworks_ai/accounts/fireworks/models/qwen2p5-coder-32b | 32.768K | $0.9 | $0.9 | — | |||
| OpenAI: GPT-5.4 (batch)openai/gpt-5.4:batch | 1.05M | $1.25 | $7.5 | — | |||
| gpt-5.4-mini-2026-03-17azure_ai/gpt-5.4-mini-2026-03-17 | 272K | $0.75 | $4.5 | — | |||
| us-gov-east-1/openai.gpt-oss-120b-1:0bedrock/us-gov-east-1/openai.gpt-oss-120b-1:0 | 128K | $0.18 | $0.72 | — | |||
| MoonshotAI: Kimi K3 (batch)moonshotai/kimi-k3:batch | 1.04858M | $3 | $15 | — | |||
| us-gov-east-1/anthropic.claude-sonnet-5bedrock/us-gov-east-1/anthropic.claude-sonnet-5 | 1M | $2.4 | $12 | — | |||
| us-gov-east-1/anthropic.claude-3-haiku-20240307-v1:0bedrock/us-gov-east-1/anthropic.claude-3-haiku-20240307-v1:0 | 200K | $0.3 | $1.5 | — | |||
| us-gov-east-1/anthropic.claude-opus-4-8bedrock/us-gov-east-1/anthropic.claude-opus-4-8 | 1M | $6 | $30 | — | |||
| Arcee AI: Coder Largearcee-ai/coder-large | 32.768K | $0.5 | $0.8 | — | |||
| eu/gpt-4o-2024-08-06azure/eu/gpt-4o-2024-08-06 | 128K | $2.75 | $11 | — | |||
| eu/gpt-4o-2024-11-20azure/eu/gpt-4o-2024-11-20 | 128K | $2.75 | $11 | — | |||
| @cf/meta/llama-3.2-1b-instructcloudflare/@cf/meta/llama-3.2-1b-instruct | 60K | $0.027 | $0.201 | — | |||
| qwen3-coder-flash-2025-07-28qwencloud/qwen3-coder-flash-2025-07-28 | 997.952K | — | — | — | |||
| eu/gpt-5-2025-08-07azure/eu/gpt-5-2025-08-07 | 272K | $1.375 | $11 | — | |||
| eu/gpt-5-mini-2025-08-07azure/eu/gpt-5-mini-2025-08-07 | 272K | $0.275 | $2.2 | — | |||
| accounts/fireworks/models/qwen2p5-coder-1p5b-instructfireworks_ai/accounts/fireworks/models/qwen2p5-coder-1p5b-instruct | 32.768K | $0.1 | $0.1 | — |