No provider description is available for this model yet.
Models
Every model in the catalog with source-linked pricing, context limits, provider availability, and published benchmark results.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
Fugu Max is the cost-performance model in Sakana AI's Fugu family. Rather than a single monolithic model, Fugu is a learned multi-agent orchestration system: a language model trained to route...
No provider description is available for this model yet.
No provider description is available for this model yet.
Gemini 2.5 Flash-Lite is a lightweight reasoning model in the Gemini 2.5 family, optimized for ultra-low latency and cost efficiency. It offers improved throughput, faster token generation, and better performance...
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
Hermes 4 is a large-scale reasoning model built on Meta-Llama-3.1-405B and released by Nous Research. It introduces a hybrid reasoning mode, where the model can choose to deliberate internally with...
No provider description is available for this model yet.
No provider description is available for this model yet.
GPT-5 Pro is OpenAI’s most advanced model, offering major improvements in reasoning, code quality, and user experience. It is optimized for complex tasks that require step-by-step reasoning, instruction following, and...
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
Mistral Medium 3.1 is an updated version of Mistral Medium 3, which is a high-performance enterprise-grade language model designed to deliver frontier-level capabilities at significantly reduced operational cost. It balances...
No provider description is available for this model yet.
No provider description is available for this model yet.
Ministral 8B is an 8B parameter model featuring a unique interleaved sliding-window attention pattern for faster, memory-efficient inference. Designed for edge use cases, it supports up to 128k context length...
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
GPT-5.5 is OpenAI’s frontier model designed for complex professional workloads, building on GPT-5.4 with stronger reasoning, higher reliability, and improved token efficiency on hard tasks. It features a 1M+ token...
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
Muse Spark 1.3 Contributor is the cost-efficient contributor tier of Meta’s multimodal reasoning model for experimentation, learning, and early-stage agentic, multi-agent, and coding workflows. It is designed to track information...
GLM-5 is Z.ai’s flagship open-source foundation model engineered for complex systems design and long-horizon agent workflows. Built for expert developers, it delivers production-grade performance on large-scale programming tasks, rivaling leading...
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
| Model | Creator | Inputs | Context | Input | Output | Released | Compare |
|---|---|---|---|---|---|---|---|
| gemini-2.5-flash-native-audio-latestgemini/gemini-2.5-flash-native-audio-latest | 1.04858M | $0.3 | $2.5 | — | |||
| gemma-4-31b-it-thinkinglibertai/gemma-4-31b-it-thinking | 262.144K | $0.15 | $0.4 | — | |||
| deepseek-r1-7b-qwenllamagate/deepseek-r1-7b-qwen | 131.072K | $0.08 | $0.15 | — | |||
| qwen3-8bllamagate/qwen3-8b | 32.768K | $0.04 | $0.14 | — | |||
| databricks-gpt-5-4databricks/databricks-gpt-5-4 | 272K | $2.5 | $15 | — | |||
| databricks-gpt-5-3-codexdatabricks/databricks-gpt-5-3-codex | 272K | $1.75 | $14 | — | |||
| databricks-gpt-5-2-codexdatabricks/databricks-gpt-5-2-codex | 272K | $1.75 | $14 | — | |||
| databricks-gpt-5-2databricks/databricks-gpt-5-2 | 272K | $1.75 | $14 | — | |||
| openai/gpt-5.6-solopenrouter/openai/gpt-5.6-sol | 1.05M | $2 | $10 | — | |||
| us-gov-west-1/anthropic.claude-haiku-4-5-20251001-v1:0bedrock/us-gov-west-1/anthropic.claude-haiku-4-5-20251001-v1:0 | 200K | $1.2 | $6 | — | |||
| databricks-gpt-5-1-codex-minidatabricks/databricks-gpt-5-1-codex-mini | 272K | $0.25 | $2 | — | |||
| databricks-gpt-5-1-codex-maxdatabricks/databricks-gpt-5-1-codex-max | 272K | $1.25 | $10 | — | |||
| Sakana: Fugu Maxsakana/fugu-max | 1M | $2 | $6 | — | |||
| databricks-gemini-3-prodatabricks/databricks-gemini-3-pro | 1.04858M | $2.5 | $15 | — | |||
| deepseek/deepseek-r1novita/deepseek/deepseek-r1 | 64K | $4 | $4 | — | |||
| Google: Gemini 2.5 Flash Lite (batch)google/gemini-2.5-flash-lite:batch | 1.04858M | $0.05 | $0.2 | — | |||
| us-gov-east-1/anthropic.claude-haiku-4-5-20251001-v1:0bedrock/us-gov-east-1/anthropic.claude-haiku-4-5-20251001-v1:0 | 200K | $1.2 | $6 | — | |||
| gemini-2.0-flash-lite-001gemini/gemini-2.0-flash-lite-001 | 1.04858M | $0.075 | $0.3 | — | |||
| deepseek/deepseek_v3novita/deepseek/deepseek_v3 | 64K | $0.89 | $0.89 | — | |||
| Nous: Hermes 4 405Bnousresearch/hermes-4-405b | 131.072K | $1 | $3 | — | |||
| qwen/qwen3.6-35b-a3bnovita/qwen/qwen3.6-35b-a3b | 262.144K | $0.248 | $1.485 | — | |||
| zai-org/glm-4.7-flashnovita/zai-org/glm-4.7-flash | 200K | $0.07 | $0.4 | — | |||
| OpenAI: GPT-5 Pro (batch)openai/gpt-5-pro:batch | 400K | $7.5 | $60 | — | |||
| us-west-2/zai.glm-5bedrock/us-west-2/zai.glm-5 | 200K | $1 | $3.2 | — | |||
| zai-org/glm-4.7-hnovita/zai-org/glm-4.7-h | 204.8K | $0.6 | $2.2 | — | |||
| moonshotai/kimi-k2.5novita/moonshotai/kimi-k2.5 | 262.144K | $0.6 | $3 | — | |||
| Mistral: Mistral Medium 3.1 (batch)mistralai/mistral-medium-3.1:batch | 131.072K | $0.2 | $1 | — | |||
| databricks-gemini-3-flashdatabricks/databricks-gemini-3-flash | 1.04858M | $0.625 | $3.75 | — | |||
| databricks-gemini-3-1-prodatabricks/databricks-gemini-3-1-pro | 1.04858M | $2.5 | $15 | — | |||
| Mistral: Ministral 8Bmistralai/ministral-8b | 128K | $0.11 | $0.11 | — | |||
| us-east-1/zai.glm-5bedrock/us-east-1/zai.glm-5 | 200K | $1 | $3.2 | — | |||
| gpt-5-search-api-2025-10-14openai/gpt-5-search-api-2025-10-14 | 272K | $1.25 | $10 | — | |||
| gemma-4-31b-itlibertai/gemma-4-31b-it | 262.144K | $0.15 | $0.4 | — | |||
| qwen/qwen3-coder-nextnovita/qwen/qwen3-coder-next | 262.144K | $0.2 | $1.5 | — | |||
| qwen-flash-2025-07-28qwencloud/qwen-flash-2025-07-28 | 997.952K | — | — | — | |||
| zai-org/glm-5novita/zai-org/glm-5 | 202.8K | $1 | $3.2 | — | |||
| minimax/minimax-m2.5novita/minimax/minimax-m2.5 | 204.8K | $0.3 | $1.2 | — | |||
| qwen-flashqwencloud/qwen-flash | 997.952K | — | — | — | |||
| qwen-coderqwencloud/qwen-coder | 1M | $0.3 | $1.5 | — | |||
| qwen/qwen3.5-397b-a17bnovita/qwen/qwen3.5-397b-a17b | 262.144K | $0.6 | $3.6 | — | |||
| OpenAI: GPT-5.5 (batch)openai/gpt-5.5:batch | 1.05M | $2.5 | $15 | — | |||
| doubao-seed-2-0-code-preview-260215volcengine/doubao-seed-2-0-code-preview-260215 | 256K | — | — | — | |||
| kimi-k2.7-codeqwencloud/kimi-k2.7-code | 229.376K | $0.95 | $4 | — | |||
| databricks-gemini-3-1-flash-litedatabricks/databricks-gemini-3-1-flash-lite | 1.04858M | $0.312 | $1.875 | — | |||
| Meta: Muse Spark 1.3 Contributormeta/muse-spark-1.3-contributor | 1.04858M | $0.1 | $0.2 | — | |||
| Z.ai: GLM 5z-ai/glm-5 | 198K | $0.6 | $1.92 | — | |||
| glm-5.2qwencloud/glm-5.2 | 1.04858M | $1.4 | $4.4 | — | |||
| qwen-plusqwencloud/qwen-plus | 129.024K | $0.4 | $1.2 | — | |||
| qwen-plus-2025-01-25qwencloud/qwen-plus-2025-01-25 | 129.024K | $0.4 | $1.2 | — | |||
| glm-5.1qwencloud/glm-5.1 | 202.745K | $1.4 | $4.4 | — |