The preview GPT-4 model with improved instruction following, JSON mode, reproducible outputs, parallel function calling, and more. Training data: up to Dec 2023. **Note:** heavily rate limited by OpenAI while...
Models
Every model in the catalog with source-linked pricing, context limits, provider availability, and published benchmark results.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
Schematron V2 Small is a 3B-parameter HTML-to-JSON extraction model from Inference.net. It prioritizes extraction quality for complex schemas and long pages. Extraction instructions must be supplied through a JSON schema...
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
gpt-oss-20b is an open-weight 21B parameter model released by OpenAI under the Apache 2.0 license. It uses a Mixture-of-Experts (MoE) architecture with 3.6B active parameters per forward pass, optimized for...
GPT-4o Search Previewis a specialized model for web search in Chat Completions. It is trained to understand and execute web search queries.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
Kimi K2 Instruct is a large-scale Mixture-of-Experts (MoE) language model developed by Moonshot AI, featuring 1 trillion total parameters with 32 billion active per forward pass. It is optimized for...
No provider description is available for this model yet.
Claude Opus 4 is benchmarked as the world’s best coding model, at time of release, bringing sustained performance on complex, long-running tasks and agent workflows. It sets new benchmarks in...
Qwen3-VL-30B-A3B-Thinking is a multimodal model that unifies strong text generation with visual understanding for images and videos. Its Thinking variant enhances reasoning in STEM, math, and complex tasks. It excels...
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
Qwen3.7-Plus is a cost-effective model in Alibaba's Qwen3.7 series. It supports text and image input with text output, building on the series' text capabilities with a comprehensive upgrade to its...
No provider description is available for this model yet.
Qwen3-14B is a dense 14.8B parameter causal language model from the Qwen3 series, designed for both complex reasoning and efficient dialogue. It supports seamless switching between a "thinking" mode for...
No provider description is available for this model yet.
No provider description is available for this model yet.
GPT-4o mini Search Preview is a specialized model for web search in Chat Completions. It is trained to understand and execute web search queries.
Qwen 3.6 Plus builds on a hybrid architecture that combines efficient linear attention with sparse mixture-of-experts routing, enabling strong scalability and high-performance inference. Compared to the 3.5 series, it delivers...
No provider description is available for this model yet.
No provider description is available for this model yet.
GPT-5 Chat is designed for advanced, natural, multimodal, and context-aware conversations for enterprise applications.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
Inkling Small is an open-weight multimodal mixture-of-experts model from Thinking Machines Lab, with 12B active parameters out of 276B total. It is positioned as the smaller, more efficient member of...
No provider description is available for this model yet.
No provider description is available for this model yet.
Claude Fable 5.1 improves on Claude Fable 5 across the board, with the biggest gains in agentic coding, long-running agentic workflows, and knowledge work: long code refactors, front-end and visual...
No provider description is available for this model yet.
No provider description is available for this model yet.
| Model | Creator | Inputs | Context | Input | Output | Released | Compare |
|---|---|---|---|---|---|---|---|
| OpenAI: GPT-4 Turbo Previewopenai/gpt-4-turbo-preview | 128K | $10 | $30 | — | |||
| jp.anthropic.claude-opus-4-7bedrock_converse/jp.anthropic.claude-opus-4-7 | 1M | $5.5 | $27.5 | — | |||
| anthropic.claude-sonnet-5bedrock_converse/anthropic.claude-sonnet-5 | 1M | $2 | $10 | — | |||
| global.anthropic.claude-sonnet-5bedrock_converse/global.anthropic.claude-sonnet-5 | 1M | $2 | $10 | — | |||
| Inference.net: Schematron V2 Smallinference-net/schematron-v2-small | 128K | $0.05 | $0.23 | — | |||
| us.anthropic.claude-sonnet-5bedrock_converse/us.anthropic.claude-sonnet-5 | 1M | $2.2 | $11 | — | |||
| eu.anthropic.claude-sonnet-5bedrock_converse/eu.anthropic.claude-sonnet-5 | 1M | $2.2 | $11 | — | |||
| au.anthropic.claude-sonnet-5bedrock_converse/au.anthropic.claude-sonnet-5 | 1M | $2.2 | $11 | — | |||
| jp.anthropic.claude-sonnet-5bedrock_converse/jp.anthropic.claude-sonnet-5 | 1M | $2.2 | $11 | — | |||
| anthropic.claude-sonnet-4-6bedrock_converse/anthropic.claude-sonnet-4-6 | 1M | $3 | $15 | — | |||
| global.anthropic.claude-sonnet-4-6bedrock_converse/global.anthropic.claude-sonnet-4-6 | 1M | $3 | $15 | — | |||
| us.anthropic.claude-sonnet-4-6bedrock_converse/us.anthropic.claude-sonnet-4-6 | 1M | $3.3 | $16.5 | — | |||
| OpenAI: gpt-oss-20b (batch)openai/gpt-oss-20b:batch | 131.072K | $0.05 | $0.2 | — | |||
| OpenAI: GPT-4o Search Previewopenai/gpt-4o-search-preview | 128K | $2.5 | $10 | — | |||
| @cf/openai/gpt-oss-120bcloudflare/@cf/openai/gpt-oss-120b | 128K | $0.35 | $0.75 | — | |||
| eu.anthropic.claude-sonnet-4-6bedrock_converse/eu.anthropic.claude-sonnet-4-6 | 1M | $3.3 | $16.5 | — | |||
| au.anthropic.claude-sonnet-4-6bedrock_converse/au.anthropic.claude-sonnet-4-6 | 1M | $3.3 | $16.5 | — | |||
| MoonshotAI: Kimi K2 0711moonshotai/kimi-k2 | 131.072K | $0.57 | $2.3 | — | |||
| jp.anthropic.claude-sonnet-4-6bedrock_converse/jp.anthropic.claude-sonnet-4-6 | 1M | $3.3 | $16.5 | — | |||
| Anthropic: Claude Opus 4anthropic/claude-opus-4 | 200K | $15 | $75 | — | |||
| Qwen: Qwen3 VL 30B A3B Thinkingqwen/qwen3-vl-30b-a3b-thinking | 131.072K | $0.2 | $2.4 | — | |||
| anthropic.claude-sonnet-4-20250514-v1:0bedrock_converse/anthropic.claude-sonnet-4-20250514-v1:0 | 1M | $3 | $15 | — | |||
| anthropic.claude-sonnet-4-5-20250929-v1:0bedrock_converse/anthropic.claude-sonnet-4-5-20250929-v1:0 | 200K | $3 | $15 | — | |||
| anthropic.claude-v1bedrock/anthropic.claude-v1 | 100K | $8 | $24 | — | |||
| anthropic.claude-v2:1bedrock/anthropic.claude-v2:1 | 100K | $8 | $24 | — | |||
| HuggingFaceH4/zephyr-7b-betaanyscale/huggingfaceh4/zephyr-7b-beta | 16.384K | $0.15 | $0.15 | — | |||
| google/gemma-7b-itanyscale/google/gemma-7b-it | 8.192K | $0.15 | $0.15 | — | |||
| Qwen: Qwen3.7 Plusqwen/qwen3.7-plus | 1M | $0.32 | $1.28 | — | |||
| gemini-3.7-flashvertex_ai-language-models/gemini-3.7-flash | 1.04858M | $0.75 | $3.75 | — | |||
| Qwen: Qwen3 14Bqwen/qwen3-14b | 131.072K | $0.227 | $0.91 | — | |||
| meta-llama/Meta-Llama-3-70B-Instructanyscale/meta-llama/meta-llama-3-70b-instruct | 8.192K | $1 | $1 | — | |||
| meta-llama/Meta-Llama-3-8B-Instructanyscale/meta-llama/meta-llama-3-8b-instruct | 8.192K | $0.15 | $0.15 | — | |||
| OpenAI: GPT-4o-mini Search Previewopenai/gpt-4o-mini-search-preview | 128K | $0.15 | $0.6 | — | |||
| Qwen: Qwen3.6 Plusqwen/qwen3.6-plus | 1M | $0.325 | $1.95 | — | |||
| mistralai/Mistral-7B-Instruct-v0.1anyscale/mistralai/mistral-7b-instruct-v0.1 | 16.384K | $0.15 | $0.15 | — | |||
| @cf/google/gemma-2b-it-loracloudflare/@cf/google/gemma-2b-it-lora | 8.192K | — | — | — | |||
| OpenAI: GPT-5 Chatopenai/gpt-5-chat | 128K | $1.25 | $10 | — | |||
| mistralai/Mixtral-8x22B-Instruct-v0.1anyscale/mistralai/mixtral-8x22b-instruct-v0.1 | 65.536K | $0.9 | $0.9 | — | |||
| mistralai/Mixtral-8x7B-Instruct-v0.1anyscale/mistralai/mixtral-8x7b-instruct-v0.1 | 16.384K | $0.15 | $0.15 | — | |||
| apac.amazon.nova-lite-v1:0bedrock_converse/apac.amazon.nova-lite-v1:0 | 300K | $0.063 | $0.252 | — | |||
| apac.amazon.nova-micro-v1:0bedrock_converse/apac.amazon.nova-micro-v1:0 | 128K | $0.037 | $0.148 | — | |||
| @cf/meta/llama-3.2-3b-instructcloudflare/@cf/meta/llama-3.2-3b-instruct | 80K | $0.051 | $0.335 | — | |||
| apac.amazon.nova-pro-v1:0bedrock_converse/apac.amazon.nova-pro-v1:0 | 300K | $0.84 | $3.36 | — | |||
| apac.anthropic.claude-3-5-sonnet-20240620-v1:0bedrock/apac.anthropic.claude-3-5-sonnet-20240620-v1:0 | 200K | $3 | $15 | — | |||
| Thinking Machines: Inkling Small (free)thinkingmachines/inkling-small:free | 1.04858M | Free | Free | — | |||
| apac.anthropic.claude-3-5-sonnet-20241022-v2:0bedrock/apac.anthropic.claude-3-5-sonnet-20241022-v2:0 | 200K | $3 | $15 | — | |||
| anthropic.claude-opus-4-8bedrock_converse/anthropic.claude-opus-4-8 | 1M | $5 | $25 | — | |||
| Anthropic: Claude Fable 5.1anthropic/claude-fable-5.1 | 1M | $10 | $50 | — | |||
| apac.anthropic.claude-3-haiku-20240307-v1:0bedrock/apac.anthropic.claude-3-haiku-20240307-v1:0 | 200K | $0.25 | $1.25 | — | |||
| gpt-5.6-cyberopenai/gpt-5.6-cyber | 400K | $12.5 | $75 | — |