No provider description is available for this model yet.
Models
Every model in the catalog with source-linked pricing, context limits, provider availability, and published benchmark results.
Mistral Medium 3.5 is a dense 128B instruction-following model from Mistral AI. It supports text and image inputs with text output, and is designed for agentic workflows, coding, and complex...
Mistral's cutting-edge language model for coding released end of July 2025. Codestral specializes in low-latency, high-frequency tasks such as fill-in-the-middle (FIM), code correction and test generation. [Blog Post](https://mistral.ai/news/codestral-25-08)
Qwen3-14B is a dense 14.8B parameter causal language model from the Qwen3 series, designed for both complex reasoning and efficient dialogue. It supports seamless switching between a "thinking" mode for...
Claude Sonnet 4.5 is Anthropic’s most advanced Sonnet model to date, optimized for real-world agents and coding workflows. It delivers state-of-the-art performance on coding benchmarks such as SWE-bench Verified, with...
No provider description is available for this model yet.
No provider description is available for this model yet.
OpenAI o3-mini is a cost-efficient language model optimized for STEM reasoning tasks, particularly excelling in science, mathematics, and coding. This model supports the `reasoning_effort` parameter, which can be set to...
Claude Sonnet 4 significantly enhances the capabilities of its predecessor, Sonnet 3.7, excelling in both coding and reasoning tasks with improved precision and controllability. Achieving state-of-the-art performance on SWE-bench (72.7%),...
Gemini 2.5 Pro is Google’s state-of-the-art AI model designed for advanced reasoning, coding, mathematics, and scientific tasks. It employs “thinking” capabilities, enabling it to reason through responses with enhanced accuracy...
No provider description is available for this model yet.
GPT-5.2 Pro is OpenAI’s most advanced model, offering major improvements in agentic coding and long context performance over GPT-5 Pro. It is optimized for complex tasks that require step-by-step reasoning,...
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
Relace Apply 3 is a specialized code-patching LLM that merges AI-suggested edits straight into your source files. It can apply updates from GPT-4o, Claude, and others into your files at...
Claude Opus 4.1 is an updated version of Anthropic’s flagship model, offering improved performance in coding, reasoning, and agentic tasks. It achieves 74.5% on SWE-bench Verified and shows notable gains...
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
MoonshotAI: Kimi K2.7 Code is a coding-focused model in Moonshot AI's Kimi K2 family, built to complete end-to-end programming tasks reliably over long contexts. It uses a native multimodal mixture-of-experts...
No provider description is available for this model yet.
No provider description is available for this model yet.
*Ling-3.0-flash* is a *124B-parameter Mixture-of-Experts (MoE) model*, with approximately *5.1B parameters activated per token*. The model is designed with *token efficiency and production-scale agentic inference* as key priorities, enabling developers...
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
Claude Opus 4.8 is Anthropic's most capable generally available model in the Opus family. It supports text, image, and file inputs with text output, with reasoning support and a 1M-token...
Claude Opus 4.8 is Anthropic's most capable generally available model in the Opus family. It supports text, image, and file inputs with text output, with reasoning support and a 1M-token...
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
| Model | Creator | Inputs | Context | Input | Output | Released | Compare |
|---|---|---|---|---|---|---|---|
| gemini-2.5-flash-preview-09-2025gemini/gemini-2.5-flash-preview-09-2025 | 1.04858M | $0.3 | $2.5 | — | |||
| Mistral: Mistral Medium 3.5 (batch)mistralai/mistral-medium-3-5:batch | 262.144K | $0.75 | $3.75 | — | |||
| Mistral: Codestral 2508mistralai/codestral-2508 | 256K | $0.3 | $0.9 | — | |||
| Qwen: Qwen3 14Bqwen/qwen3-14b | 131.072K | $0.227 | $0.91 | — | |||
| Anthropic: Claude Sonnet 4.5anthropic/claude-sonnet-4.5 | 1M | $3 | $15 | — | |||
| gpt-3.5-turbo-instruct-0914azure_text/gpt-3.5-turbo-instruct-0914 | 4.097K | $1.5 | $2 | — | |||
| accounts/fireworks/models/mixtral-8x22b-instructfireworks_ai/accounts/fireworks/models/mixtral-8x22b-instruct | 65.536K | $1.2 | $1.2 | — | |||
| OpenAI: o3 Mini (batch)openai/o3-mini:batch | 200K | $0.55 | $2.2 | — | |||
| Anthropic: Claude Sonnet 4anthropic/claude-sonnet-4 | 200K | $3 | $15 | — | |||
| Google: Gemini 2.5 Pro Preview 05-06google/gemini-2.5-pro-preview-05-06 | 1.04858M | $1.25 | $10 | — | |||
| accounts/fireworks/models/mixtral-8x22bfireworks_ai/accounts/fireworks/models/mixtral-8x22b | 65.536K | $1.2 | $1.2 | — | |||
| OpenAI: GPT-5.2 Pro (batch)openai/gpt-5.2-pro:batch | 400K | $10.5 | $84 | — | |||
| gemini-2.5-flash-lite-preview-09-2025gemini/gemini-2.5-flash-lite-preview-09-2025 | 1.04858M | $0.1 | $0.4 | — | |||
| accounts/fireworks/models/mistral-small-24b-instruct-2501fireworks_ai/accounts/fireworks/models/mistral-small-24b-instruct-2501 | 32.768K | $0.9 | $0.9 | — | |||
| accounts/fireworks/models/mistral-nemo-instruct-2407fireworks_ai/accounts/fireworks/models/mistral-nemo-instruct-2407 | 128K | $0.2 | $0.2 | — | |||
| gemini-2.5-flash-litegemini/gemini-2.5-flash-lite | 1.04858M | $0.1 | $0.4 | — | |||
| @cf/meta-llama/llama-2-7b-chat-hf-loracloudflare/@cf/meta-llama/llama-2-7b-chat-hf-lora | 8.192K | — | — | — | |||
| accounts/fireworks/models/mistral-nemo-base-2407fireworks_ai/accounts/fireworks/models/mistral-nemo-base-2407 | 128K | $0.2 | $0.2 | — | |||
| Qwen/Qwen-Drive-1.0-4BQwen/Qwen-Drive-1.0-4B | Not documented | — | — | — | |||
| accounts/fireworks/models/mistral-large-3-fp8fireworks_ai/accounts/fireworks/models/mistral-large-3-fp8 | 256K | $1.2 | $1.2 | — | |||
| gemini-2.5-flashgemini/gemini-2.5-flash | 1.04858M | $0.3 | $2.5 | — | |||
| accounts/fireworks/models/mistral-7b-v0p2fireworks_ai/accounts/fireworks/models/mistral-7b-v0p2 | 32.768K | $0.2 | $0.2 | — | |||
| Relace: Relace Apply 3relace/relace-apply-3 | 256K | $0.85 | $1.25 | — | |||
| Anthropic: Claude Opus 4.1 (batch)anthropic/claude-opus-4.1:batch | 200K | $7.5 | $37.5 | — | |||
| accounts/fireworks/models/mistral-7b-instruct-v3fireworks_ai/accounts/fireworks/models/mistral-7b-instruct-v3 | 32.768K | $0.2 | $0.2 | — | |||
| gemini-2.0-flash-litegemini/gemini-2.0-flash-lite | 1.04858M | $0.075 | $0.3 | — | |||
| microsoft/Dayhoff-170M-GRS-SS-38000microsoft/Dayhoff-170M-GRS-SS-38000 | Not documented | — | — | — | |||
| nvidia/NVIDIA-NemotronLabs-AI-for-Media-Sports-Tennisnvidia/NVIDIA-NemotronLabs-AI-for-Media-Sports-Tennis | Not documented | — | — | — | |||
| gpt-3.5-turbo-0125azure/gpt-3.5-turbo-0125 | 16.384K | $0.5 | $1.5 | — | |||
| deepseek-ai/DeepSeek-R1-0528-Qwen3-8Bdeepseek-ai/DeepSeek-R1-0528-Qwen3-8B | Not documented | — | — | — | |||
| MoonshotAI: Kimi K2.7 Code (batch)moonshotai/kimi-k2.7-code:batch | 262.144K | $0.95 | $4 | — | |||
| accounts/fireworks/models/mistral-7b-instruct-v0p2fireworks_ai/accounts/fireworks/models/mistral-7b-instruct-v0p2 | 32.768K | $0.2 | $0.2 | — | |||
| deepseek-ai/DeepSeek-V4.1-Flashdeepseek-ai/DeepSeek-V4.1-Flash | Not documented | — | — | — | |||
| inclusionAI: Ling 3.0 Flashinclusionai/ling-3.0-flash | 262.144K | $0.021 | $0.063 | — | |||
| accounts/fireworks/models/mistral-7b-instruct-4kfireworks_ai/accounts/fireworks/models/mistral-7b-instruct-4k | 32.768K | $0.2 | $0.2 | — | |||
| gemini-2.0-flash-001gemini/gemini-2.0-flash-001 | 1.04858M | $0.1 | $0.4 | — | |||
| accounts/fireworks/models/mistral-7bfireworks_ai/accounts/fireworks/models/mistral-7b | 32.768K | $0.2 | $0.2 | — | |||
| accounts/fireworks/models/ministral-3-8b-instruct-2512fireworks_ai/accounts/fireworks/models/ministral-3-8b-instruct-2512 | 256K | $0.2 | $0.2 | — | |||
| gemini-2.0-flashgemini/gemini-2.0-flash | 1.04858M | $0.1 | $0.4 | — | |||
| microsoft/Dayhoff-170M-GRS-SS-86000microsoft/Dayhoff-170M-GRS-SS-86000 | Not documented | — | — | — | |||
| accounts/fireworks/models/ministral-3-3b-instruct-2512fireworks_ai/accounts/fireworks/models/ministral-3-3b-instruct-2512 | 256K | $0.1 | $0.1 | — | |||
| accounts/fireworks/models/ministral-3-14b-instruct-2512fireworks_ai/accounts/fireworks/models/ministral-3-14b-instruct-2512 | 256K | $0.2 | $0.2 | — | |||
| Anthropic: Claude Opus 4.8anthropic/claude-opus-4.8 | 1M | $5 | $25 | — | |||
| Anthropic: Claude Opus 4.8 (batch)anthropic/claude-opus-4.8:batch | 1M | $2.5 | $12.5 | — | |||
| gemini-2.5-computer-use-preview-10-2025vertex_ai-language-models/gemini-2.5-computer-use-preview-10-2025 | 128K | $1.25 | $10 | — | |||
| accounts/fireworks/models/minimax-m2fireworks_ai/accounts/fireworks/models/minimax-m2 | 4.096K | $0.3 | $1.2 | — | |||
| accounts/fireworks/models/minimax-m1-80kfireworks_ai/accounts/fireworks/models/minimax-m1-80k | 4.096K | $0.1 | $0.1 | — | |||
| gemini-robotics-er-1.5-previewgemini/gemini-robotics-er-1.5-preview | 1.04858M | $0.3 | $2.5 | — | |||
| microsoft/Dayhoff-3b-UR90-10microsoft/Dayhoff-3b-UR90-10 | Not documented | — | — | — | |||
| @cf/google/gemma-7b-it-loracloudflare/@cf/google/gemma-7b-it-lora | 3.5K | — | — | — |