No provider description is available for this model yet.
Models
Every model in the catalog with source-linked pricing, context limits, provider availability, and published benchmark results.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
Mistral's cutting-edge language model for coding released end of July 2025. Codestral specializes in low-latency, high-frequency tasks such as fill-in-the-middle (FIM), code correction and test generation. [Blog Post](https://mistral.ai/news/codestral-25-08)
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
Claude Opus 4.8 is Anthropic's most capable generally available model in the Opus family. It supports text, image, and file inputs with text output, with reasoning support and a 1M-token...
Uncensored and creative writing model based on Mistral Small 3.2 24B with good recall, prompt adherence, and intelligence.
No provider description is available for this model yet.
Qwen3-VL-235B-A22B Thinking is a multimodal model that unifies strong text generation with visual understanding across images and video. The Thinking model is optimized for multimodal reasoning in STEM and math....
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
GPT-5.6 Sol is the flagship model in OpenAI's GPT-5.6 series. It is suited for complex reasoning, coding, and agentic workflows, and is particularly strong at command-line and multi-step coding tasks...
No provider description is available for this model yet.
No provider description is available for this model yet.
The o1 series of models are trained with reinforcement learning to think before they answer and perform complex reasoning. The o1-pro model uses more compute to think harder and provide...
GPT-5.4 Pro is OpenAI's most advanced model, building on GPT-5.4's unified architecture with enhanced reasoning capabilities for complex, high-stakes tasks. It features a 1M+ token context window (922K input, 128K...
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
Ling 3.0 Flash VL builds on Ling 3.0 Flash (124B total / 5.5B active MoE from InclusionAI), further strengthening its language capabilities while adding native visual perception and advanced visual...
No provider description is available for this model yet.
No provider description is available for this model yet.
Muse Spark 1.2 contributor tier is a reasoning model from Meta designed for developers who want to start building at an even lower cost. It’s meaningfully cheaper than Muse Spark...
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
DeepSeek V4 Flash 0731 is a sparse mixture-of-experts model from DeepSeek, with 13B active parameters out of 284B total. This re-post-trained revision is suited for coding, reasoning, and agent workflows....
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
gpt-oss-20b is an open-weight 21B parameter model released by OpenAI under the Apache 2.0 license. It uses a Mixture-of-Experts (MoE) architecture with 3.6B active parameters per forward pass, optimized for...
Claude Opus 4.1 is an updated version of Anthropic’s flagship model, offering improved performance in coding, reasoning, and agentic tasks. It achieves 74.5% on SWE-bench Verified and shows notable gains...
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
*Ling-3.0-flash* is a *124B-parameter Mixture-of-Experts (MoE) model*, with approximately *5.1B parameters activated per token*. The model is designed with *token efficiency and production-scale agentic inference* as key priorities, enabling developers...
No provider description is available for this model yet.
Llama 3.2 11B Vision is a multimodal model with 11 billion parameters, designed to handle tasks combining visual and textual data. It excels in tasks such as image captioning and...
No provider description is available for this model yet.
No provider description is available for this model yet.
| Model | Creator | Inputs | Context | Input | Output | Released | Compare |
|---|---|---|---|---|---|---|---|
| gemma-4-31b-itgemini/gemma-4-31b-it | 262.144K | — | — | — | |||
| grok-3xai/grok-3 | 131.072K | $3 | $15 | — | |||
| openai/gpt-4o-minigmi/openai/gpt-4o-mini | 131.072K | $0.15 | $0.6 | — | |||
| Qwen/Qwen3-Coder-480B-A35B-Instruct-FP8together_ai/qwen/qwen3-coder-480b-a35b-instruct-fp8 | 256K | $2 | $2 | — | |||
| Mistral: Codestral 2508 (batch)mistralai/codestral-2508:batch | 256K | $0.15 | $0.45 | — | |||
| amazon.titan-text-express-v1bedrock/amazon.titan-text-express-v1 | 42K | $1.3 | $1.7 | — | |||
| amazon.titan-text-lite-v1bedrock/amazon.titan-text-lite-v1 | 42K | $0.3 | $0.4 | — | |||
| amazon.titan-text-premier-v1:0bedrock/amazon.titan-text-premier-v1:0 | 42K | $0.5 | $1.5 | — | |||
| deepseek-ai/DeepSeek-V3.2gmi/deepseek-ai/deepseek-v3.2 | 163.84K | $0.28 | $0.4 | — | |||
| Anthropic: Claude Opus 4.8 (batch)anthropic/claude-opus-4.8:batch | 1M | $2.5 | $12.5 | — | |||
| TheDrummer: Cydonia 24B V4.1thedrummer/cydonia-24b-v4.1 | 131.072K | $0.3 | $0.5 | — | |||
| anthropic.claude-3-5-haiku-20241022-v1:0bedrock/anthropic.claude-3-5-haiku-20241022-v1:0 | 200K | $0.8 | $4 | — | |||
| Qwen: Qwen3 VL 235B A22B Thinkingqwen/qwen3-vl-235b-a22b-thinking | 131.072K | $0.4 | $4 | — | |||
| us-gov-east-1/amazon.nova-pro-v1:0bedrock/us-gov-east-1/amazon.nova-pro-v1:0 | 300K | $0.96 | $3.84 | — | |||
| labs-leanstral-1-5mistral/labs-leanstral-1-5 | 262.144K | — | — | — | |||
| gemma-4-26b-a4b-itgemini/gemma-4-26b-a4b-it | 262.144K | — | — | — | |||
| OpenAI: GPT-5.6 Sol (batch)openai/gpt-5.6-sol:batch | 1.05M | $1 | $5 | — | |||
| anthropic.claude-haiku-4-5-20251001-v1:0bedrock_converse/anthropic.claude-haiku-4-5-20251001-v1:0 | 200K | $1 | $5 | — | |||
| claude-mythos-previewanthropic/claude-mythos-preview | 1M | $10 | $50 | — | |||
| OpenAI: o1-pro (batch)openai/o1-pro:batch | 200K | $75 | $300 | — | |||
| OpenAI: GPT-5.4 Pro (batch)openai/gpt-5.4-pro:batch | 1.05M | $15 | $90 | — | |||
| us-gov-east-1/amazon.titan-text-express-v1bedrock/us-gov-east-1/amazon.titan-text-express-v1 | 42K | $1.3 | $1.7 | — | |||
| databricks-glm-5-3-flashdatabricks/databricks-glm-5-3-flash | 1.04858M | — | — | — | |||
| anthropic.claude-haiku-4-5@20251001bedrock_converse/anthropic.claude-haiku-4-5@20251001 | 200K | $1 | $5 | — | |||
| inclusionAI: Ling 3.0 Flash VLinclusionai/ling-3.0-flash-vl | 131.072K | $0.06 | $0.18 | — | |||
| anthropic.claude-3-5-sonnet-20240620-v1:0bedrock/anthropic.claude-3-5-sonnet-20240620-v1:0 | 1M | $3 | $15 | — | |||
| grok-2-vision-latestxai/grok-2-vision-latest | 32.768K | $2 | $10 | — | |||
| Meta: Muse Spark 1.2 Contributormeta/muse-spark-1.2-contributor | 1.04858M | $0.1 | $0.2 | — | |||
| us-gov-east-1/amazon.titan-text-lite-v1bedrock/us-gov-east-1/amazon.titan-text-lite-v1 | 42K | $0.3 | $0.4 | — | |||
| grok-4.20-multi-agent-0309xai/grok-4.20-multi-agent-0309 | 1M | $1.25 | $2.5 | — | |||
| us-gov-east-1/amazon.titan-text-premier-v1:0bedrock/us-gov-east-1/amazon.titan-text-premier-v1:0 | 42K | $0.5 | $1.5 | — | |||
| DeepSeek: DeepSeek V4 Flash 0731 (batch)deepseek/deepseek-v4-flash-0731:batch | 1.04858M | $0.11 | $0.33 | — | |||
| anthropic.claude-3-5-sonnet-20241022-v2:0bedrock/anthropic.claude-3-5-sonnet-20241022-v2:0 | 1M | $3 | $15 | — | |||
| anthropic.claude-3-7-sonnet-20240620-v1:0bedrock/anthropic.claude-3-7-sonnet-20240620-v1:0 | 200K | $3.6 | $18 | — | |||
| anthropic.claude-3-7-sonnet-20250219-v1:0bedrock_converse/anthropic.claude-3-7-sonnet-20250219-v1:0 | 200K | $3 | $15 | — | |||
| anthropic.claude-3-haiku-20240307-v1:0bedrock/anthropic.claude-3-haiku-20240307-v1:0 | 200K | $0.25 | $1.25 | — | |||
| anthropic.claude-3-opus-20240229-v1:0bedrock/anthropic.claude-3-opus-20240229-v1:0 | 200K | $15 | $75 | — | |||
| OpenAI: gpt-oss-20b (batch)openai/gpt-oss-20b:batch | 131.072K | $0.05 | $0.2 | — | |||
| Anthropic: Claude Opus 4.1 (batch)anthropic/claude-opus-4.1:batch | 200K | $7.5 | $37.5 | — | |||
| anthropic.claude-3-sonnet-20240229-v1:0bedrock/anthropic.claude-3-sonnet-20240229-v1:0 | 200K | $3 | $15 | — | |||
| us-east-1/qwen.qwen3-coder-nextbedrock/us-east-1/qwen.qwen3-coder-next | 262.144K | $0.5 | $1.2 | — | |||
| anthropic.claude-instant-v1bedrock/anthropic.claude-instant-v1 | 100K | $0.8 | $2.4 | — | |||
| anthropic.claude-opus-4-1-20250805-v1:0bedrock_converse/anthropic.claude-opus-4-1-20250805-v1:0 | 200K | $15 | $75 | — | |||
| anthropic.claude-opus-4-20250514-v1:0bedrock_converse/anthropic.claude-opus-4-20250514-v1:0 | 200K | $15 | $75 | — | |||
| zai-glm-4.7cerebras/zai-glm-4.7 | 128K | $2.25 | $2.75 | — | |||
| inclusionAI: Ling 3.0 Flashinclusionai/ling-3.0-flash | 262.144K | $0.021 | $0.063 | — | |||
| anthropic.claude-opus-4-5-20251101-v1:0bedrock_converse/anthropic.claude-opus-4-5-20251101-v1:0 | 200K | $5 | $25 | — | |||
| Meta: Llama 3.2 11B Vision Instructmeta-llama/llama-3.2-11b-vision-instruct | 131.072K | $0.345 | $0.345 | — | |||
| anthropic.claude-opus-4-6-v1bedrock_converse/anthropic.claude-opus-4-6-v1 | 1M | $5 | $25 | — | |||
| google/gemma-4-31B-it-turbodeepinfra/google/gemma-4-31b-it-turbo | 262.144K | $0.09 | $0.34 | — |