Ling 3.0 Tiny is a mixture-of-experts model from InclusionAI, with 1.3B active parameters out of 7.9B total. It is designed for responsive agents, instruction following, and multi-turn conversations, with switchable...
Models
Every model in the catalog with source-linked pricing, context limits, provider availability, and published benchmark results.
Claude Fable 5.1 improves on Claude Fable 5 across the board, with the biggest gains in agentic coding, long-running agentic workflows, and knowledge work: long code refactors, front-end and visual...
GPT-5.1 is the latest frontier-grade model in the GPT-5 series, offering stronger general-purpose reasoning, improved instruction adherence, and a more natural conversational style compared to GPT-5. It uses adaptive reasoning...
gpt-oss-120b is an open-weight, 117B-parameter Mixture-of-Experts (MoE) language model from OpenAI designed for high-reasoning, agentic, and general-purpose production use cases. It activates 5.1B parameters per forward pass and is optimized...
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
Kimi K3 is a 2.8T parameter open-weight multimodal reasoning model from Moonshot AI. It is suited for complex coding, knowledge work, and long-horizon agentic workflows, and is particularly strong at...
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
Qwen 3.6 Plus builds on a hybrid architecture that combines efficient linear attention with sparse mixture-of-experts routing, enabling strong scalability and high-performance inference. Compared to the 3.5 series, it delivers...
No provider description is available for this model yet.
Gemma 4 31B Instruct is Google DeepMind's 30.7B dense multimodal model supporting text and image input with text output. Features a 256K token context window, configurable thinking/reasoning mode, native function...
No provider description is available for this model yet.
No provider description is available for this model yet.
The o1 series of models are trained with reinforcement learning to think before they answer and perform complex reasoning. The o1-pro model uses more compute to think harder and provide...
Gemini 3.8 Flash is Google's most intelligent Flash model with significant gains from 3.7 Flash across software engineering, agentic tasks, and multi-step reasoning.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
This model always redirects to the latest model in the DeepSeek Pro family.
No provider description is available for this model yet.
No provider description is available for this model yet.
Mercury 2.5 is the fastest reasoning LLM, and the latest diffusion LLM (dLLM) from Inception. Instead of generating tokens sequentially, Mercury 2.5 produces and refines multiple tokens in parallel, achieving...
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
GPT-4o mini is OpenAI's newest model after [GPT-4 Omni](/models/openai/gpt-4o), supporting both text and image inputs with text outputs. As their most advanced small model, it is many multiples more affordable...
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
Mistral's cutting-edge language model for coding released end of July 2025. Codestral specializes in low-latency, high-frequency tasks such as fill-in-the-middle (FIM), code correction and test generation. [Blog Post](https://mistral.ai/news/codestral-25-08)
No provider description is available for this model yet.
Llama 3.2 11B Vision is a multimodal model with 11 billion parameters, designed to handle tasks combining visual and textual data. It excels in tasks such as image captioning and...
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
Hermes 3 is a generalist language model with many improvements over Hermes 2, including advanced agentic capabilities, much better roleplaying, reasoning, multi-turn conversation, long context coherence, and improvements across the...
No provider description is available for this model yet.
| Model | Creator | Inputs | Context | Input | Output | Released | Compare |
|---|---|---|---|---|---|---|---|
| inclusionAI: Ling 3.0 Tiny (free)inclusionai/ling-3.0-tiny:free | 262.144K | Free | Free | — | |||
| Anthropic: Claude Fable 5.1 (batch)anthropic/claude-fable-5.1:batch | 1M | $5 | $25 | — | |||
| OpenAI: GPT-5.1 (batch)openai/gpt-5.1:batch | 400K | $0.625 | $5 | — | |||
| OpenAI: gpt-oss-120b (batch)openai/gpt-oss-120b:batch | 131.072K | $0.15 | $0.6 | — | |||
| eu-central-1/1-month-commitment/anthropic.claude-v1bedrock/eu-central-1/1-month-commitment/anthropic.claude-v1 | 100K | — | — | — | |||
| eu-central-1/1-month-commitment/anthropic.claude-instant-v1bedrock/eu-central-1/1-month-commitment/anthropic.claude-instant-v1 | 100K | — | — | — | |||
| amazon.titan-text-express-v1bedrock/amazon.titan-text-express-v1 | 42K | $1.3 | $1.7 | — | |||
| amazon.titan-text-lite-v1bedrock/amazon.titan-text-lite-v1 | 42K | $0.3 | $0.4 | — | |||
| MoonshotAI: Kimi K3 (batch)moonshotai/kimi-k3:batch | 1.04858M | $3 | $15 | — | |||
| deepseek-ai/DeepSeek-V3.2gmi/deepseek-ai/deepseek-v3.2 | 163.84K | $0.28 | $0.4 | — | |||
| openai/gpt-4-turbovercel_ai_gateway/openai/gpt-4-turbo | 128K | $10 | $30 | — | |||
| anthropic.claude-3-5-haiku-20241022-v1:0bedrock/anthropic.claude-3-5-haiku-20241022-v1:0 | 200K | $0.8 | $4 | — | |||
| Qwen: Qwen3.6 Plusqwen/qwen3.6-plus | 1M | $0.325 | $1.95 | — | |||
| us-gov-east-1/amazon.nova-pro-v1:0bedrock/us-gov-east-1/amazon.nova-pro-v1:0 | 300K | $0.96 | $3.84 | — | |||
| Google: Gemma 4 31B (batch)google/gemma-4-31b-it:batch | 262.144K | $0.39 | $0.97 | — | |||
| anthropic.claude-haiku-4-5-20251001-v1:0bedrock_converse/anthropic.claude-haiku-4-5-20251001-v1:0 | 200K | $1 | $5 | — | |||
| eu-north-1/moonshotai.kimi-k2.5bedrock/eu-north-1/moonshotai.kimi-k2.5 | 262.144K | $0.72 | $3.6 | — | |||
| OpenAI: o1-pro (batch)openai/o1-pro:batch | 200K | $75 | $300 | — | |||
| Google: Gemini 3.8 Flash (batch)google/gemini-3.8-flash:batch | 1.04858M | $0.375 | $1.875 | — | |||
| us-gov-east-1/amazon.titan-text-express-v1bedrock/us-gov-east-1/amazon.titan-text-express-v1 | 42K | $1.3 | $1.7 | — | |||
| eu-north-1/minimax.minimax-m2.5bedrock/eu-north-1/minimax.minimax-m2.5 | 1M | $0.36 | $1.44 | — | |||
| gemini-3.7-flashaihubmix/gemini-3.7-flash | 1.04858M | $0.75 | $3.75 | — | |||
| anthropic.claude-3-5-sonnet-20240620-v1:0bedrock/anthropic.claude-3-5-sonnet-20240620-v1:0 | 1M | $3 | $15 | — | |||
| DeepSeek: DeepSeek Pro Latest~deepseek/deepseek-pro-latest | 1.04858M | $0.66 | $1.98 | — | |||
| us-gov-east-1/amazon.titan-text-lite-v1bedrock/us-gov-east-1/amazon.titan-text-lite-v1 | 42K | $0.3 | $0.4 | — | |||
| eu-north-1/minimax.minimax-m2.1bedrock/eu-north-1/minimax.minimax-m2.1 | 196K | $0.36 | $1.44 | — | |||
| Inception: Mercury 2.5inception/mercury-2.5 | 260K | $0.04 | $0.15 | — | |||
| us-gov-east-1/amazon.titan-text-premier-v1:0bedrock/us-gov-east-1/amazon.titan-text-premier-v1:0 | 42K | $0.5 | $1.5 | — | |||
| anthropic.claude-3-5-sonnet-20241022-v2:0bedrock/anthropic.claude-3-5-sonnet-20241022-v2:0 | 1M | $3 | $15 | — | |||
| anthropic.claude-3-7-sonnet-20240620-v1:0bedrock/anthropic.claude-3-7-sonnet-20240620-v1:0 | 200K | $3.6 | $18 | — | |||
| anthropic.claude-3-7-sonnet-20250219-v1:0bedrock_converse/anthropic.claude-3-7-sonnet-20250219-v1:0 | 200K | $3 | $15 | — | |||
| anthropic.claude-3-haiku-20240307-v1:0bedrock/anthropic.claude-3-haiku-20240307-v1:0 | 200K | $0.25 | $1.25 | — | |||
| eu-north-1/deepseek.v3.2bedrock/eu-north-1/deepseek.v3.2 | 163.84K | $0.74 | $2.22 | — | |||
| anthropic.claude-3-opus-20240229-v1:0bedrock/anthropic.claude-3-opus-20240229-v1:0 | 200K | $15 | $75 | — | |||
| OpenAI: GPT-4o-mini (batch)openai/gpt-4o-mini:batch | 128K | $0.075 | $0.3 | — | |||
| morph/morph-v3-largevercel_ai_gateway/morph/morph-v3-large | 32.768K | $0.9 | $1.9 | — | |||
| anthropic.claude-3-sonnet-20240229-v1:0bedrock/anthropic.claude-3-sonnet-20240229-v1:0 | 200K | $3 | $15 | — | |||
| anthropic.claude-instant-v1bedrock/anthropic.claude-instant-v1 | 100K | $0.8 | $2.4 | — | |||
| anthropic.claude-opus-4-1-20250805-v1:0bedrock_converse/anthropic.claude-opus-4-1-20250805-v1:0 | 200K | $15 | $75 | — | |||
| anthropic.claude-opus-4-20250514-v1:0bedrock_converse/anthropic.claude-opus-4-20250514-v1:0 | 200K | $15 | $75 | — | |||
| zai-glm-4.7cerebras/zai-glm-4.7 | 128K | $2.25 | $2.75 | — | |||
| Mistral: Codestral 2508 (batch)mistralai/codestral-2508:batch | 256K | $0.15 | $0.45 | — | |||
| anthropic.claude-opus-4-5-20251101-v1:0bedrock_converse/anthropic.claude-opus-4-5-20251101-v1:0 | 200K | $5 | $25 | — | |||
| Meta: Llama 3.2 11B Vision Instructmeta-llama/llama-3.2-11b-vision-instruct | 131.072K | $0.345 | $0.345 | — | |||
| anthropic.claude-opus-4-6-v1bedrock_converse/anthropic.claude-opus-4-6-v1 | 1M | $5 | $25 | — | |||
| global.anthropic.claude-opus-4-6-v1bedrock_converse/global.anthropic.claude-opus-4-6-v1 | 1M | $5 | $25 | — | |||
| us.anthropic.claude-opus-4-6-v1bedrock_converse/us.anthropic.claude-opus-4-6-v1 | 1M | $5.5 | $27.5 | — | |||
| eu.anthropic.claude-opus-4-6-v1bedrock_converse/eu.anthropic.claude-opus-4-6-v1 | 1M | $5.5 | $27.5 | — | |||
| Nous: Hermes 3 405B Instruct (free)nousresearch/hermes-3-llama-3.1-405b:free | 131.072K | Free | Free | — | |||
| openai.gpt-oss-120b-1:0bedrock_converse/openai.gpt-oss-120b-1:0 | 128K | $0.15 | $0.6 | — |