No provider description is available for this model yet.
Models
Every model in the catalog with source-linked pricing, context limits, provider availability, and published benchmark results.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
Mercury 2 is an extremely fast reasoning LLM, and the first reasoning diffusion LLM (dLLM). Instead of generating tokens sequentially, Mercury 2 produces and refines multiple tokens in parallel, achieving...
A cost-efficient version of GPT Audio. The new snapshot features an upgraded decoder for more natural sounding voices and maintains better voice consistency. Input is priced at $0.60 per million...
LFM2.5-2.6B is a compact reasoning model from Liquid AI. It is suited for agent workflows, data extraction, RAG, and long-context processing. Liquid advises against using it for agentic coding or...
Mistral Medium 3.1 is an updated version of Mistral Medium 3, which is a high-performance enterprise-grade language model designed to deliver frontier-level capabilities at significantly reduced operational cost. It balances...
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
Gemini 2.5 Flash is Google's state-of-the-art workhorse model, specifically designed for advanced reasoning, coding, mathematics, and scientific tasks. It includes built-in "thinking" capabilities, enabling it to provide responses with greater...
Ministral 8B is an 8B parameter model featuring a unique interleaved sliding-window attention pattern for faster, memory-efficient inference. Designed for edge use cases, it supports up to 128k context length...
GLM-5 is Z.ai’s flagship open-source foundation model engineered for complex systems design and long-horizon agent workflows. Built for expert developers, it delivers production-grade performance on large-scale programming tasks, rivaling leading...
| Model | Creator | Inputs | Context | Input | Output | Released | Compare |
|---|---|---|---|---|---|---|---|
| doubao-seed-2-0-mini-260215volcengine/doubao-seed-2-0-mini-260215 | 256K | — | — | — | |||
| doubao-seed-2-0-code-preview-260215volcengine/doubao-seed-2-0-code-preview-260215 | 256K | — | — | — | |||
| us-east-1/zai.glm-5bedrock/us-east-1/zai.glm-5 | 200K | $1 | $3.2 | — | |||
| us-west-2/zai.glm-5bedrock/us-west-2/zai.glm-5 | 200K | $1 | $3.2 | — | |||
| us-gov-east-1/anthropic.claude-haiku-4-5-20251001-v1:0bedrock/us-gov-east-1/anthropic.claude-haiku-4-5-20251001-v1:0 | 200K | $1.2 | $6 | — | |||
| us-gov-west-1/anthropic.claude-haiku-4-5-20251001-v1:0bedrock/us-gov-west-1/anthropic.claude-haiku-4-5-20251001-v1:0 | 200K | $1.2 | $6 | — | |||
| claude-sonnet-4-5snowflake/claude-sonnet-4-5 | 200K | $3 | $15 | — | |||
| claude-sonnet-4-6snowflake/claude-sonnet-4-6 | 200K | $3 | $15 | — | |||
| claude-4-opussnowflake/claude-4-opus | 200K | $5 | $25 | — | |||
| claude-haiku-4-5snowflake/claude-haiku-4-5 | 200K | $1 | $5 | — | |||
| claude-3-7-sonnetsnowflake/claude-3-7-sonnet | 200K | $3 | $15 | — | |||
| openai-gpt-4.1snowflake/openai-gpt-4.1 | 300K | $2 | $8 | — | |||
| openai-gpt-5snowflake/openai-gpt-5 | 300K | $1.25 | $10 | — | |||
| openai-gpt-5-minisnowflake/openai-gpt-5-mini | 1M | $0.3 | $1.2 | — | |||
| openai-gpt-5-nanosnowflake/openai-gpt-5-nano | 5M | $0.15 | $0.6 | — | |||
| llama4-mavericksnowflake/llama4-maverick | 128K | $0.24 | $0.97 | — | |||
| Qwen/Qwen3.5-397B-A17B-FP8tensormesh/qwen/qwen3.5-397b-a17b-fp8 | 262.144K | $0.6 | $3.6 | — | |||
| Qwen/Qwen3-Coder-480B-A35B-Instruct-FP8tensormesh/qwen/qwen3-coder-480b-a35b-instruct-fp8 | 262.144K | $0.45 | $1.8 | — | |||
| Qwen/Qwen3.6-27B-FP8tensormesh/qwen/qwen3.6-27b-fp8 | 262.144K | $0.32 | $3.2 | — | |||
| lukealonso/GLM-5.1-NVFP4-MTPtensormesh/lukealonso/glm-5.1-nvfp4-mtp | 202.752K | $1.4 | $4.4 | — | |||
| deepseek-ai/DeepSeek-V4-Flashtensormesh/deepseek-ai/deepseek-v4-flash | 32.768K | $0.14 | $0.28 | — | |||
| moonshotai/Kimi-K2.6tensormesh/moonshotai/kimi-k2.6 | 32.768K | $0.96 | $4 | — | |||
| MiniMaxAI/MiniMax-M2.5tensormesh/minimaxai/minimax-m2.5 | 196.608K | $0.3 | $1.2 | — | |||
| google/gemma-4-31B-ittensormesh/google/gemma-4-31b-it | 32.768K | $0.14 | $0.56 | — | |||
| openai/gpt-oss-120btensormesh/openai/gpt-oss-120b | 131.072K | $0.15 | $0.6 | — | |||
| openai/gpt-oss-20btensormesh/openai/gpt-oss-20b | 131.072K | $0.07 | $0.28 | — | |||
| deepseek-v4-protencent/deepseek-v4-pro | 1M | $0.435 | $0.87 | — | |||
| deepseek-v4-flashtencent/deepseek-v4-flash | 1M | $0.14 | $0.28 | — | |||
| ps/glm-4.5-airpinstripes/ps/glm-4.5-air | 128K | $0.125 | $0.45 | — | |||
| ps/qwen3.6-35b-a3bpinstripes/ps/qwen3.6-35b-a3b | 131.072K | $0.14 | $0.45 | — | |||
| ps/qwen3-30b-a3bpinstripes/ps/qwen3-30b-a3b | 131.072K | $0.09 | $0.2 | — | |||
| ps/qwen3-coder-30b-a3bpinstripes/ps/qwen3-coder-30b-a3b | 131.072K | $0.3 | $0.6 | — | |||
| ps/deepseek-v4-flashpinstripes/ps/deepseek-v4-flash | 163.84K | $0.1 | $0.2 | — | |||
| ps/minimax-m2.7pinstripes/ps/minimax-m2.7 | 1.00019M | $0.255 | $0.55 | — | |||
| gemma-4-26bdarkbloom/gemma-4-26b | 131.072K | $0.03 | $0.165 | — | |||
| gpt-oss-20bdarkbloom/gpt-oss-20b | 131.072K | $0.015 | $0.07 | — | |||
| fallback_generalizationsunknown/fallback_generalizations | Not documented | — | — | — | |||
| Inception: Mercury 2inception/mercury-2 | 128K | $0.25 | $0.75 | — | |||
| OpenAI: GPT Audio Miniopenai/gpt-audio-mini | 128K | $0.6 | $2.4 | — | |||
| LiquidAI: LFM2.5-2.6B (free)liquid/lfm-2.5-2.6b:free | 65.536K | Free | Free | — | |||
| Mistral: Mistral Medium 3.1mistralai/mistral-medium-3.1 | 131.072K | $0.4 | $2 | — | |||
| Qwen/Qwen3.8-2.4T-A95BQwen/Qwen3.8-2.4T-A95B | Not documented | — | — | — | |||
| nvidia/NVIDIA-Nemotron-Labs-Teacher-STEMnvidia/NVIDIA-Nemotron-Labs-Teacher-STEM | Not documented | — | — | — | |||
| nvidia/NVIDIA-Nemotron-Labs-Teacher-Instruction-Followingnvidia/NVIDIA-Nemotron-Labs-Teacher-Instruction-Following | Not documented | — | — | — | |||
| nvidia/NVIDIA-Nemotron-Labs-Teacher-Competition-Codingnvidia/NVIDIA-Nemotron-Labs-Teacher-Competition-Coding | Not documented | — | — | — | |||
| openai/gpt-oss-20breplicate/openai/gpt-oss-20b | Not documented | $0.09 | $0.36 | — | |||
| nvidia/NVIDIA-Nemotron-Labs-Teacher-Chatnvidia/NVIDIA-Nemotron-Labs-Teacher-Chat | Not documented | — | — | — | |||
| Google: Gemini 2.5 Flash (batch)google/gemini-2.5-flash:batch | 1.04858M | $0.15 | $1.25 | — | |||
| Mistral: Ministral 8Bmistralai/ministral-8b | 128K | $0.11 | $0.11 | — | |||
| Z.ai: GLM 5z-ai/glm-5 | 198K | $0.6 | $1.92 | — |