No provider description is available for this model yet.
Models
Every model in the catalog with source-linked pricing, context limits, provider availability, and published benchmark results.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
Schematron V2 Small is a 3B-parameter HTML-to-JSON extraction model from Inference.net. It prioritizes extraction quality for complex schemas and long pages. Extraction instructions must be supplied through a JSON schema...
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
Ling 3.0 Flash Fin is a finance-focused mixture-of-experts model from InclusionAI, built on Ling 3.0 Flash with 5.1B active parameters out of 124B total. It is designed for real-world investment...
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
Gemini 2.5 Flash-Lite is a lightweight reasoning model in the Gemini 2.5 family, optimized for ultra-low latency and cost efficiency. It offers improved throughput, faster token generation, and better performance...
| Model | Creator | Inputs | Context | Input | Output | Released | Compare |
|---|---|---|---|---|---|---|---|
| us-gov-east-1/anthropic.claude-sonnet-5bedrock/us-gov-east-1/anthropic.claude-sonnet-5 | 1M | $2.4 | $12 | — | |||
| gpt-5.4-nano-2026-03-17azure_ai/gpt-5.4-nano-2026-03-17 | 272K | $0.2 | $1.25 | — | |||
| us-gov-east-1/anthropic.claude-opus-4-8bedrock/us-gov-east-1/anthropic.claude-opus-4-8 | 1M | $6 | $30 | — | |||
| eu/gpt-4o-2024-08-06azure/eu/gpt-4o-2024-08-06 | 128K | $2.75 | $11 | — | |||
| eu/gpt-4o-2024-11-20azure/eu/gpt-4o-2024-11-20 | 128K | $2.75 | $11 | — | |||
| @cf/meta/llama-3.2-1b-instructcloudflare/@cf/meta/llama-3.2-1b-instruct | 60K | $0.027 | $0.201 | — | |||
| eu/gpt-4o-mini-2024-07-18azure/eu/gpt-4o-mini-2024-07-18 | 128K | $0.165 | $0.66 | — | |||
| eu/gpt-5-2025-08-07azure/eu/gpt-5-2025-08-07 | 272K | $1.375 | $11 | — | |||
| eu/gpt-5-mini-2025-08-07azure/eu/gpt-5-mini-2025-08-07 | 272K | $0.275 | $2.2 | — | |||
| eu/gpt-5.1azure/eu/gpt-5.1 | 272K | $1.38 | $11 | — | |||
| eu/gpt-5.1-chatazure/eu/gpt-5.1-chat | 128K | $1.38 | $11 | — | |||
| eu/gpt-5-nano-2025-08-07azure/eu/gpt-5-nano-2025-08-07 | 272K | $0.055 | $0.44 | — | |||
| eu/o1-2024-12-17azure/eu/o1-2024-12-17 | 200K | $16.5 | $66 | — | |||
| eu/o1-mini-2024-09-12azure/eu/o1-mini-2024-09-12 | 128K | $1.21 | $4.84 | — | |||
| us-gov-west-1/xai.grok-4.3bedrock_mantle/us-gov-west-1/xai.grok-4.3 | 131.072K | $1.5 | $3 | — | |||
| eu/o1-preview-2024-09-12azure/eu/o1-preview-2024-09-12 | 128K | $16.5 | $66 | — | |||
| us-gov/gpt-5.1azure/us-gov/gpt-5.1 | 272K | $1.719 | $13.75 | — | |||
| Inference.net: Schematron V2 Smallinference-net/schematron-v2-small | 128K | $0.05 | $0.23 | — | |||
| us-gov/o3-miniazure/us-gov/o3-mini | 200K | $1.513 | $6.05 | — | |||
| eu/o3-mini-2025-01-31azure/eu/o3-mini-2025-01-31 | 200K | $1.21 | $4.84 | — | |||
| global-standard/gpt-4o-2024-08-06azure/global-standard/gpt-4o-2024-08-06 | 128K | $2.5 | $10 | — | |||
| global-standard/gpt-4o-2024-11-20azure/global-standard/gpt-4o-2024-11-20 | 128K | $2.5 | $10 | — | |||
| inclusionAI: Ling 3.0 Flash Fin (free)inclusionai/ling-3.0-flash-fin:free | 262.144K | Free | Free | — | |||
| global-standard/gpt-4o-miniazure/global-standard/gpt-4o-mini | 128K | $0.15 | $0.6 | — | |||
| us-gov-west-1/openai.gpt-oss-120b-1:0bedrock/us-gov-west-1/openai.gpt-oss-120b-1:0 | 128K | $0.18 | $0.72 | — | |||
| global/gpt-4o-2024-11-20azure/global/gpt-4o-2024-11-20 | 128K | $2.5 | $10 | — | |||
| global/gpt-5.1azure/global/gpt-5.1 | 272K | $1.25 | $10 | — | |||
| global/gpt-5.1-chatazure/global/gpt-5.1-chat | 128K | $1.25 | $10 | — | |||
| databricks-claude-fable-5databricks/databricks-claude-fable-5 | 1M | $10 | $50 | — | |||
| global/gpt-4o-2024-08-06azure/global/gpt-4o-2024-08-06 | 128K | $2.5 | $10 | — | |||
| us-gov-west-1/anthropic.claude-sonnet-5bedrock/us-gov-west-1/anthropic.claude-sonnet-5 | 1M | $2.4 | $12 | — | |||
| us-gov-west-1/anthropic.claude-opus-4-8bedrock/us-gov-west-1/anthropic.claude-opus-4-8 | 1M | $6 | $30 | — | |||
| us-west-2/anthropic.claude-v1bedrock/us-west-2/anthropic.claude-v1 | 100K | $8 | $24 | — | |||
| us-west-2/anthropic.claude-v2:1bedrock/us-west-2/anthropic.claude-v2:1 | 100K | $8 | $24 | — | |||
| us-gov-east-1/nvidia.nemotron-nano-3-30bbedrock/us-gov-east-1/nvidia.nemotron-nano-3-30b | 262.144K | $0.072 | $0.288 | — | |||
| Google: Gemini 2.5 Flash Lite (batch)google/gemini-2.5-flash-lite:batch | 1.04858M | $0.05 | $0.2 | — |