No provider description is available for this model yet.
Models
Every model in the catalog with source-linked pricing, context limits, provider availability, and published benchmark results.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
GPT-5.6 Terra is a balanced model in OpenAI's GPT-5.6 series, positioned between the flagship Sol tier and the cost-efficient Luna tier. It is suited for everyday coding, reasoning, and agentic...
No provider description is available for this model yet.
MiniMax-M3 is a multimodal foundation model from MiniMax. It supports text, image, and video inputs with text output, a 1M-token context window, and is suited for long-horizon agentic work, coding,...
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
Nex-N2.5 is an agentic model built to turn goals into working, verified outcomes. Its core strength is agentic coding within a visual feedback loop: it can explore codebases, implement multi-file...
Claude Opus 5 is Anthropic’s flagship model for demanding reasoning, coding, and long-horizon agentic work. It is particularly strong at end-to-end software tasks, code review and bug finding, visual analysis...
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
Muse Glimmer 30B is a dense, open-weight multimodal model from Meta Superintelligence Labs, distilled from Muse Spark and optimized for autonomous agents on consumer hardware. It is suited for long-horizon...
No provider description is available for this model yet.
No provider description is available for this model yet.
Gemini 3.1 Flash Lite is Google’s GA high-efficiency multimodal model optimized for low-latency, high-volume workloads. It supports text, image, video, audio, and PDF inputs, and is designed for lightweight agentic...
Grok 4.3 is a reasoning model from SpaceXAI. It accepts text and image inputs with text output, and is suited for agentic workflows, instruction-following tasks, and applications requiring high factual...
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
DeepSeek-V3.2-Exp is an experimental large language model released by DeepSeek as an intermediate step between V3.1 and future architectures. It introduces DeepSeek Sparse Attention (DSA), a fine-grained sparse attention mechanism...
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
Gemini 3.5 Flash Lite is a high-efficiency model from Google with upgraded agentic capabilities. It is suited for subagents that execute focused tasks within complex, multi-agent workflows.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
GPT-5.4 Pro is OpenAI's most advanced model, building on GPT-5.4's unified architecture with enhanced reasoning capabilities for complex, high-stakes tasks. It features a 1M+ token context window (922K input, 128K...
No provider description is available for this model yet.
No provider description is available for this model yet.
Inkling is an open-weight multimodal mixture-of-experts model from Thinking Machines Lab, with 41B active parameters out of 975B total. It is designed for general-purpose reasoning, coding, agentic and tool-use systems,...
No provider description is available for this model yet.
DeepSeek V4 Flash 0731 is a sparse mixture-of-experts model from DeepSeek, with 13B active parameters out of 284B total. This re-post-trained revision is suited for coding, reasoning, and agent workflows....
No provider description is available for this model yet.
No provider description is available for this model yet.
NVIDIA Nemotron 3 Ultra is an open frontier-reasoning and orchestration model from NVIDIA, with 55B active parameters out of 550B total (MoE). Built on a hybrid Transformer-Mamba mixture-of-experts architecture, it...
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
| Model | Creator | Inputs | Context | Input | Output | Released | Compare |
|---|---|---|---|---|---|---|---|
| Qwen/Qwen3.8-Flashtogether_ai/qwen/qwen3.8-flash | 1M | $0.15 | $0.47 | — | |||
| gemma-4-31bcerebras/gemma-4-31b | 131.072K | $0.99 | $1.49 | — | |||
| accounts/fireworks/models/cogito-v1-preview-qwen-14bfireworks_ai/accounts/fireworks/models/cogito-v1-preview-qwen-14b | 131.072K | $0.2 | $0.2 | — | |||
| accounts/fireworks/models/cogito-v1-preview-llama-8bfireworks_ai/accounts/fireworks/models/cogito-v1-preview-llama-8b | 131.072K | $0.2 | $0.2 | — | |||
| anthropic/claude-opus-4.7openrouter/anthropic/claude-opus-4.7 | 1M | $5 | $25 | — | |||
| OpenAI: GPT-5.6 Terra (batch)openai/gpt-5.6-terra:batch | 1.05M | $1 | $6 | — | |||
| accounts/fireworks/models/cogito-v1-preview-llama-70bfireworks_ai/accounts/fireworks/models/cogito-v1-preview-llama-70b | 131.072K | $0.9 | $0.9 | — | |||
| MiniMax: MiniMax M3 (free)minimax/minimax-m3:free | 1.04858M | Free | Free | — | |||
| accounts/fireworks/models/cogito-v1-preview-llama-3bfireworks_ai/accounts/fireworks/models/cogito-v1-preview-llama-3b | 131.072K | $0.1 | $0.1 | — | |||
| anthropic/claude-haiku-4.5openrouter/anthropic/claude-haiku-4.5 | 200K | $1 | $5 | — | |||
| Qwen3-4B-Instruct-2507-GGUFlemonade/qwen3-4b-instruct-2507-gguf | 262.144K | — | — | — | |||
| Nex AGI: Nex-N2.5-Pro (free)nex-agi/nex-n2.5-pro:free | 262.144K | Free | Free | — | |||
| Claude Opus 5 (batch)anthropic/claude-opus-5:batch | 1M | $2.5 | $12.5 | — | |||
| microsoft/Dayhoff-170M-UR90-HL-48000microsoft/Dayhoff-170M-UR90-HL-48000 | Not documented | — | — | — | |||
| microsoft/Dayhoff-170M-UR90-HL-38000microsoft/Dayhoff-170M-UR90-HL-38000 | Not documented | — | — | — | |||
| microsoft/Dayhoff-170M-UR90-HL-32000microsoft/Dayhoff-170M-UR90-HL-32000 | Not documented | — | — | — | |||
| Meta: Muse Glimmer 30B (batch)meta/muse-glimmer-30b:batch | 131.072K | $0.175 | $0.75 | — | |||
| accounts/fireworks/models/cogito-671b-v2-p1fireworks_ai/accounts/fireworks/models/cogito-671b-v2-p1 | 163.84K | $1.2 | $1.2 | — | |||
| accounts/fireworks/models/codegemma-7bfireworks_ai/accounts/fireworks/models/codegemma-7b | 8.192K | $0.2 | $0.2 | — | |||
| Google: Gemini 3.1 Flash Lite (batch)google/gemini-3.1-flash-lite:batch | 1.04858M | $0.125 | $0.75 | — | |||
| SpaceXAI: Grok 4.3x-ai/grok-4.3 | 1M | $1.25 | $2.5 | — | |||
| qwen/qwen3-coder-plusopenrouter/qwen/qwen3-coder-plus | 997.952K | $1 | $5 | — | |||
| accounts/fireworks/models/codegemma-2bfireworks_ai/accounts/fireworks/models/codegemma-2b | 8.192K | $0.1 | $0.1 | — | |||
| accounts/fireworks/models/code-qwen-1p5-7bfireworks_ai/accounts/fireworks/models/code-qwen-1p5-7b | 65.536K | $0.2 | $0.2 | — | |||
| anthropic/claude-sonnet-4.5openrouter/anthropic/claude-sonnet-4.5 | 1M | $3 | $15 | — | |||
| DeepSeek: DeepSeek V3.2 Expdeepseek/deepseek-v3.2-exp | 163.84K | $0.27 | $0.41 | — | |||
| Gemma-3-4b-it-GGUFlemonade/gemma-3-4b-it-gguf | 128K | — | — | — | |||
| ft:gpt-4o-mini-2024-07-18openai/ft:gpt-4o-mini-2024-07-18 | 128K | $0.3 | $1.2 | — | |||
| fireworks-ai-56b-to-176bfireworks_ai/fireworks-ai-56b-to-176b | Not documented | $1.2 | $1.2 | — | |||
| accounts/fireworks/models/code-llama-7b-pythonfireworks_ai/accounts/fireworks/models/code-llama-7b-python | 16.384K | $0.2 | $0.2 | — | |||
| Google: Gemini 3.5 Flash Lite (batch)google/gemini-3.5-flash-lite:batch | 1.04858M | $0.15 | $1.25 | — | |||
| accounts/fireworks/models/code-llama-7b-instructfireworks_ai/accounts/fireworks/models/code-llama-7b-instruct | 16.384K | $0.2 | $0.2 | — | |||
| anthropic/claude-opus-4.6openrouter/anthropic/claude-opus-4.6 | 1M | $5 | $25 | — | |||
| accounts/fireworks/models/code-llama-7bfireworks_ai/accounts/fireworks/models/code-llama-7b | 16.384K | $0.2 | $0.2 | — | |||
| accounts/fireworks/models/code-llama-70b-pythonfireworks_ai/accounts/fireworks/models/code-llama-70b-python | 4.096K | $0.9 | $0.9 | — | |||
| OpenAI: GPT-5.4 Pro (batch)openai/gpt-5.4-pro:batch | 1.05M | $15 | $90 | — | |||
| anthropic/claude-opus-4.5openrouter/anthropic/claude-opus-4.5 | 200K | $5 | $25 | — | |||
| gpt-oss-120b-mxfp-GGUFlemonade/gpt-oss-120b-mxfp-gguf | 131.072K | — | — | — | |||
| Thinking Machines: Inkling (free)thinkingmachines/inkling:free | 1.04858M | Free | Free | — | |||
| accounts/fireworks/models/code-llama-70b-instructfireworks_ai/accounts/fireworks/models/code-llama-70b-instruct | 4.096K | $0.9 | $0.9 | — | |||
| DeepSeek: DeepSeek V4 Flash 0731 (batch)deepseek/deepseek-v4-flash-0731:batch | 1.04858M | $0.11 | $0.33 | — | |||
| accounts/fireworks/models/code-llama-70bfireworks_ai/accounts/fireworks/models/code-llama-70b | 4.096K | $0.9 | $0.9 | — | |||
| anthropic/claude-sonnet-4.6openrouter/anthropic/claude-sonnet-4.6 | 1M | $3 | $15 | — | |||
| NVIDIA: Nemotron 3 Ultra (free)nvidia/nemotron-3-ultra-550b-a55b:free | 1M | Free | Free | — | |||
| accounts/fireworks/models/code-llama-34b-pythonfireworks_ai/accounts/fireworks/models/code-llama-34b-python | 16.384K | $0.9 | $0.9 | — | |||
| accounts/fireworks/models/code-llama-34b-instructfireworks_ai/accounts/fireworks/models/code-llama-34b-instruct | 16.384K | $0.9 | $0.9 | — | |||
| anthropic/claude-sonnet-4openrouter/anthropic/claude-sonnet-4 | 1M | $3 | $15 | — | |||
| gpt-oss-20b-mxfp4-GGUFlemonade/gpt-oss-20b-mxfp4-gguf | 131.072K | — | — | — | |||
| gpt-5.2github_copilot/gpt-5.2 | 128K | — | — | — | |||
| accounts/fireworks/models/code-llama-34bfireworks_ai/accounts/fireworks/models/code-llama-34b | 16.384K | $0.9 | $0.9 | — |