Mistral-Small-3.2-24B-Instruct-2506 is an updated 24B parameter model from Mistral optimized for instruction following, repetition reduction, and improved function calling. Compared to the 3.1 release, version 3.2 significantly improves accuracy on...
Models
Every model in the catalog with source-linked pricing, context limits, provider availability, and published benchmark results.
No provider description is available for this model yet.
Aion-RP-Llama-3.1-8B ranks the highest in the character evaluation portion of the RPBench-Auto benchmark, a roleplaying-specific variant of Arena-Hard-Auto, where LLMs evaluate each other’s responses. It is a fine-tuned base model...
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
GPT-4o ("o" for "omni") is OpenAI's latest AI model, supporting both text and image inputs with text outputs. It maintains the intelligence level of [GPT-4 Turbo](/models/openai/gpt-4-turbo) while being twice as...
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
MiniMax-M2 is a compact, high-efficiency large language model optimized for end-to-end coding and agentic workflows. With 10 billion activated parameters (230 billion total), it delivers near-frontier intelligence across general reasoning,...
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
| Model | Creator | Inputs | Context | Input | Output | Released | Compare |
|---|---|---|---|---|---|---|---|
| Mistral: Mistral Small 3.2 24Bmistralai/mistral-small-3.2-24b-instruct | 256K | $0.094 | $0.25 | — | |||
| qwen/qwen3.5-397b-a17bscaleway/qwen/qwen3.5-397b-a17b | 256K | $0.6 | $3.6 | — | |||
| AionLabs: Aion-RP 1.0 (8B)aion-labs/aion-rp-llama-3.1-8b | 32.768K | $0.8 | $1.6 | — | |||
| accounts/fireworks/routers/kimi-k2p7-code-fastfireworks_ai/accounts/fireworks/routers/kimi-k2p7-code-fast | 262.144K | $1.9 | $8 | — | |||
| zai-org/GLM-4.5deepinfra/zai-org/glm-4.5 | 131.072K | $0.4 | $1.6 | — | |||
| apac.anthropic.claude-haiku-4-5-20251001-v1:0bedrock_converse/apac.anthropic.claude-haiku-4-5-20251001-v1:0 | 200K | $1.1 | $5.5 | — | |||
| gpt-6-astraazure_ai/gpt-6-astra | 922K | $10 | $50 | — | |||
| OpenAI: GPT-4o (batch)openai/gpt-4o:batch | 128K | $1.25 | $5 | — | |||
| accounts/fireworks/routers/kimi-k2p6-fastfireworks_ai/accounts/fireworks/routers/kimi-k2p6-fast | 262.144K | $2 | $8 | — | |||
| gpt-chat-latestazure_ai/gpt-chat-latest | 272K | $5 | $30 | — | |||
| accounts/fireworks/routers/glm-5p1-fastfireworks_ai/accounts/fireworks/routers/glm-5p1-fast | 202.8K | $2.8 | $8.8 | — | |||
| model-routerazure_ai/model-router | 200K | $0.14 | — | — | |||
| cohere-command-aazure_ai/cohere-command-a | 131.072K | $2.5 | $10 | — | |||
| xai/grok-4.3vertex_ai/xai/grok-4.3 | 200K | $1.25 | $2.5 | — | |||
| openai/gpt-oss-20bdeepinfra/openai/gpt-oss-20b | 131.072K | $0.04 | $0.15 | — | |||
| grok-4-20-reasoningazure_ai/grok-4-20-reasoning | 262K | $1.25 | $2.5 | — | |||
| xai/grok-4.6vertex_ai/xai/grok-4.6 | 524.288K | $2 | $6 | — | |||
| grok-4-20-non-reasoningazure_ai/grok-4-20-non-reasoning | 262K | $1.25 | $2.5 | — | |||
| accounts/fireworks/models/zephyr-7b-betafireworks_ai/accounts/fireworks/models/zephyr-7b-beta | 32.768K | $0.2 | $0.2 | — | |||
| us.openai.gpt-6-astrabedrock_converse/us.openai.gpt-6-astra | 1.05M | $11 | $55 | — | |||
| accounts/fireworks/models/yi-34b-200k-capybarafireworks_ai/accounts/fireworks/models/yi-34b-200k-capybara | 200K | $0.9 | $0.9 | — | |||
| lyria-3.5gemini/lyria-3.5 | 1.04858M | — | — | — | |||
| gpt-6-astraazure/gpt-6-astra | 922K | $10 | $50 | — | |||
| us/gpt-6-astraazure/us/gpt-6-astra | 922K | $11 | $55 | — | |||
| Codestral-2501azure_ai/codestral-2501 | 256K | $0.3 | $0.9 | — | |||
| openai/gpt-oss-120bdeepinfra/openai/gpt-oss-120b | 131.072K | $0.05 | $0.45 | — | |||
| FW-Nemotron-Lightning-3.5-30B-A3Bazure_ai/fw-nemotron-lightning-3.5-30b-a3b | 262.144K | $0.06 | $0.22 | — | |||
| MAI-Thinking-1azure_ai/mai-thinking-1 | 256K | $2 | $8 | — | |||
| Qwen/QwQ-32Btogether_ai/qwen/qwq-32b | 131.072K | $1.2 | $1.2 | — | |||
| databricks-claude-fable-5-1databricks/databricks-claude-fable-5-1 | 1M | $10 | $50 | — | |||
| accounts/fireworks/models/toppy-m-7bfireworks_ai/accounts/fireworks/models/toppy-m-7b | 32.768K | $0.2 | $0.2 | — | |||
| databricks-gemini-3-1-flash-imagedatabricks/databricks-gemini-3-1-flash-image | 131.072K | — | — | — | |||
| databricks-gemini-3-pro-imagedatabricks/databricks-gemini-3-pro-image | 65.536K | — | — | — | |||
| accounts/fireworks/models/snorkel-mistral-7b-pairrm-dpofireworks_ai/accounts/fireworks/models/snorkel-mistral-7b-pairrm-dpo | 32.768K | $0.2 | $0.2 | — | |||
| nvidia/NVIDIA-Nemotron-Nano-9B-v2deepinfra/nvidia/nvidia-nemotron-nano-9b-v2 | 131.072K | $0.04 | $0.16 | — | |||
| databricks-gemini-3-8-flashdatabricks/databricks-gemini-3-8-flash | 1.04858M | — | — | — | |||
| databricks-gemini-3-7-flashdatabricks/databricks-gemini-3-7-flash | 1.04858M | — | — | — | |||
| MiniMax: MiniMax M2minimax/minimax-m2 | 204.8K | $0.255 | $1.02 | — | |||
| accounts/fireworks/models/rolm-ocrfireworks_ai/accounts/fireworks/models/rolm-ocr | 128K | $0.2 | $0.2 | — | |||
| databricks-gemini-3-6-flashdatabricks/databricks-gemini-3-6-flash | 1.04858M | $1.875 | $9.375 | — | |||
| databricks-gemini-3-5-flashdatabricks/databricks-gemini-3-5-flash | 1.04858M | $1.875 | $11.25 | — | |||
| accounts/fireworks/models/qwq-32bfireworks_ai/accounts/fireworks/models/qwq-32b | 131.072K | $0.9 | $0.9 | — | |||
| databricks-gemini-3-5-flash-litedatabricks/databricks-gemini-3-5-flash-lite | 1.04858M | $0.375 | $3.125 | — | |||
| databricks-glm-5-3databricks/databricks-glm-5-3 | 1.04858M | $1.4 | $4.4 | — | |||
| databricks-gpt-5-6-soldatabricks/databricks-gpt-5-6-sol | 922K | $4 | $20 | — | |||
| databricks-gpt-5-6-terradatabricks/databricks-gpt-5-6-terra | 922K | $2.5 | $15 | — | |||
| databricks-gpt-5-6-lunadatabricks/databricks-gpt-5-6-luna | 922K | $1 | $6 | — | |||
| databricks-grok-4-6databricks/databricks-grok-4-6 | 500K | $2.5 | $7.5 | — | |||
| databricks-inklingdatabricks/databricks-inkling | 1M | $1 | $4.05 | — | |||
| nvidia/Llama-3.3-Nemotron-Super-49B-v1.5deepinfra/nvidia/llama-3.3-nemotron-super-49b-v1.5 | 131.072K | $0.1 | $0.4 | — |