No provider description is available for this model yet.
Models
Every model in the catalog with source-linked pricing, context limits, provider availability, and published benchmark results.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
Nex-N2.5 is an agentic model built to turn goals into working, verified outcomes. Its core strength is agentic coding within a visual feedback loop: it can explore codebases, implement multi-file...
Gemini 2.5 Flash is Google's state-of-the-art workhorse model, specifically designed for advanced reasoning, coding, mathematics, and scientific tasks. It includes built-in "thinking" capabilities, enabling it to provide responses with greater...
DeepSeek-V3.1 Terminus is an update to [DeepSeek V3.1](/deepseek/deepseek-chat-v3.1) that maintains the model's original capabilities while addressing issues reported by users, including language consistency and agent capabilities, further optimizing the model's...
OpenAI o4-mini is a compact reasoning model in the o-series, optimized for fast, cost-efficient performance while retaining strong multimodal and agentic capabilities. It supports tool use and demonstrates competitive reasoning...
Muse Glimmer 30B is a dense, open-weight multimodal model from Meta Superintelligence Labs, distilled from Muse Spark and optimized for autonomous agents on consumer hardware. It is suited for long-horizon...
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
Qwen3-235B-A22B is a 235B parameter mixture-of-experts (MoE) model developed by Qwen, activating 22B parameters per forward pass. It supports seamless switching between a "thinking" mode for complex reasoning, math, and...
o3 is a well-rounded and powerful model across domains. It sets a new standard for math, science, coding, and visual reasoning tasks. It also excels at technical writing and instruction-following....
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
DeepSeek V4 Flash 0731 is a sparse mixture-of-experts model from DeepSeek, with 13B active parameters out of 284B total. This re-post-trained revision is suited for coding, reasoning, and agent workflows....
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
Claude Opus 5 is Anthropic’s flagship model for demanding reasoning, coding, and long-horizon agentic work. It is particularly strong at end-to-end software tasks, code review and bug finding, visual analysis...
Claude Fable 5 is a Mythos-class model from Anthropic, built for autonomous knowledge work and coding. It supports text, image, and file inputs with text output, with reasoning support and...
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
GPT-5.2 Pro is OpenAI’s most advanced model, offering major improvements in agentic coding and long context performance over GPT-5 Pro. It is optimized for complex tasks that require step-by-step reasoning,...
No provider description is available for this model yet.
No provider description is available for this model yet.
Gemini 3.5 Flash is Google's high-efficiency multimodal model, bringing near-Pro level coding and reasoning at Flash-tier cost and speed. It is highly optimized for coding proficiency and parallel agentic execution...
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
Relace Apply 3 is a specialized code-patching LLM that merges AI-suggested edits straight into your source files. It can apply updates from GPT-4o, Claude, and others into your files at...
No provider description is available for this model yet.
| Model | Creator | Inputs | Context | Input | Output | Released | Compare |
|---|---|---|---|---|---|---|---|
| accounts/fireworks/models/starcoder2-15bfireworks_ai/accounts/fireworks/models/starcoder2-15b | 16.384K | $0.2 | $0.2 | — | |||
| Qwen/Qwen3.8-Flashtogether_ai/qwen/qwen3.8-flash | 1M | $0.15 | $0.47 | — | |||
| accounts/fireworks/models/starcoder-7bfireworks_ai/accounts/fireworks/models/starcoder-7b | 8.192K | $0.2 | $0.2 | — | |||
| ft:gpt-3.5-turboopenai/ft:gpt-3.5-turbo | 16.385K | $3 | $6 | — | |||
| global/gpt-5.1azure/global/gpt-5.1 | 272K | $1.25 | $10 | — | |||
| Nex AGI: Nex-N2.5-Pro (free)nex-agi/nex-n2.5-pro:free | 262.144K | Free | Free | — | |||
| Google: Gemini 2.5 Flash (batch)google/gemini-2.5-flash:batch | 1.04858M | $0.15 | $1.25 | — | |||
| DeepSeek: DeepSeek V3.1 Terminusdeepseek/deepseek-v3.1-terminus | 131.072K | $0.27 | $1 | — | |||
| OpenAI: o4 Mini (batch)openai/o4-mini:batch | 200K | $0.55 | $2.2 | — | |||
| Meta: Muse Glimmer 30B (batch)meta/muse-glimmer-30b:batch | 131.072K | $0.175 | $0.75 | — | |||
| accounts/fireworks/models/starcoder-16bfireworks_ai/accounts/fireworks/models/starcoder-16b | 8.192K | $0.2 | $0.2 | — | |||
| accounts/fireworks/models/snorkel-mistral-7b-pairrm-dpofireworks_ai/accounts/fireworks/models/snorkel-mistral-7b-pairrm-dpo | 32.768K | $0.2 | $0.2 | — | |||
| ft:davinci-002text-completion-openai/ft:davinci-002 | 16.384K | $12 | $12 | — | |||
| accounts/fireworks/models/rolm-ocrfireworks_ai/accounts/fireworks/models/rolm-ocr | 128K | $0.2 | $0.2 | — | |||
| Qwen: Qwen3 235B A22Bqwen/qwen3-235b-a22b | 131.072K | $0.455 | $1.82 | — | |||
| OpenAI: o3 (batch)openai/o3:batch | 200K | $1 | $4 | — | |||
| accounts/fireworks/models/qwq-32bfireworks_ai/accounts/fireworks/models/qwq-32b | 131.072K | $0.9 | $0.9 | — | |||
| ft:babbage-002text-completion-openai/ft:babbage-002 | 16.384K | $1.6 | $1.6 | — | |||
| global/gpt-4o-2024-11-20azure/global/gpt-4o-2024-11-20 | 128K | $2.5 | $10 | — | |||
| accounts/fireworks/models/qwen3p7-plusfireworks_ai/accounts/fireworks/models/qwen3p7-plus | 262.144K | $0.4 | $1.6 | — | |||
| accounts/fireworks/models/qwen3-vl-30b-a3b-thinkingfireworks_ai/accounts/fireworks/models/qwen3-vl-30b-a3b-thinking | 262.144K | $0.15 | $0.6 | — | |||
| meta-llama-3.1-8b-instructfriendliai/meta-llama-3.1-8b-instruct | 8.192K | $0.1 | $0.1 | — | |||
| accounts/fireworks/models/qwen3-vl-30b-a3b-instructfireworks_ai/accounts/fireworks/models/qwen3-vl-30b-a3b-instruct | 262.144K | $0.15 | $0.6 | — | |||
| DeepSeek: DeepSeek V4 Flash 0731 (batch)deepseek/deepseek-v4-flash-0731:batch | 1.04858M | $0.11 | $0.33 | — | |||
| accounts/fireworks/models/qwen3-vl-235b-a22b-thinkingfireworks_ai/accounts/fireworks/models/qwen3-vl-235b-a22b-thinking | 262.144K | $0.22 | $0.88 | — | |||
| meta-llama-3.1-70b-instructfriendliai/meta-llama-3.1-70b-instruct | 8.192K | $0.6 | $0.6 | — | |||
| chatgpt-4o-latestopenai/chatgpt-4o-latest | 128K | $5 | $15 | — | |||
| us-gov-west-1/google.gemma-4-e2bbedrock_mantle/us-gov-west-1/google.gemma-4-e2b | 128K | $0.048 | $0.096 | — | |||
| llama3:8bollama/llama3:8b | 8.192K | — | — | — | |||
| accounts/fireworks/models/qwen3-vl-235b-a22b-instructfireworks_ai/accounts/fireworks/models/qwen3-vl-235b-a22b-instruct | 262.144K | $0.22 | $0.88 | — | |||
| accounts/fireworks/models/qwen3-coder-30b-a3b-instructfireworks_ai/accounts/fireworks/models/qwen3-coder-30b-a3b-instruct | 262.144K | $0.15 | $0.6 | — | |||
| Claude Opus 5 (batch)anthropic/claude-opus-5:batch | 1M | $2.5 | $12.5 | — | |||
| Anthropic: Claude Fable 5 (batch)anthropic/claude-fable-5:batch | 1M | $5 | $25 | — | |||
| qwen3p7-plusfireworks_ai/qwen3p7-plus | 262.144K | $0.4 | $1.6 | — | |||
| accounts/fireworks/models/qwen3-8bfireworks_ai/accounts/fireworks/models/qwen3-8b | 40.96K | $0.2 | $0.2 | — | |||
| accounts/fireworks/models/qwen3-4b-instruct-2507fireworks_ai/accounts/fireworks/models/qwen3-4b-instruct-2507 | 262.144K | $0.2 | $0.2 | — | |||
| minimax-m3fireworks_ai/minimax-m3 | 512K | $0.3 | $1.2 | — | |||
| global/gpt-4o-2024-08-06azure/global/gpt-4o-2024-08-06 | 128K | $2.5 | $10 | — | |||
| OpenAI: GPT-5.2 Pro (batch)openai/gpt-5.2-pro:batch | 400K | $10.5 | $84 | — | |||
| accounts/fireworks/models/qwen3-4bfireworks_ai/accounts/fireworks/models/qwen3-4b | 40.96K | $0.2 | $0.2 | — | |||
| accounts/fireworks/models/qwen3-32bfireworks_ai/accounts/fireworks/models/qwen3-32b | 131.072K | $0.9 | $0.9 | — | |||
| Google: Gemini 3.5 Flash (batch)google/gemini-3.5-flash:batch | 1.04858M | $0.75 | $4.5 | — | |||
| minimax-m2p7fireworks_ai/minimax-m2p7 | 196.608K | $0.3 | $1.2 | — | |||
| accounts/fireworks/models/qwen3-30b-a3b-thinking-2507fireworks_ai/accounts/fireworks/models/qwen3-30b-a3b-thinking-2507 | 262.144K | $0.9 | $0.9 | — | |||
| accounts/fireworks/models/qwen3-30b-a3b-instruct-2507fireworks_ai/accounts/fireworks/models/qwen3-30b-a3b-instruct-2507 | 262.144K | $0.5 | $0.5 | — | |||
| minimax-m2p1fireworks_ai/minimax-m2p1 | 204.8K | $0.3 | $1.2 | — | |||
| global-standard/gpt-4o-miniazure/global-standard/gpt-4o-mini | 128K | $0.15 | $0.6 | — | |||
| us-gov-east-1/anthropic.claude-3-haiku-20240307-v1:0bedrock/us-gov-east-1/anthropic.claude-3-haiku-20240307-v1:0 | 200K | $0.3 | $1.5 | — | |||
| Relace: Relace Apply 3relace/relace-apply-3 | 256K | $0.85 | $1.25 | — | |||
| accounts/fireworks/models/qwen3-30b-a3bfireworks_ai/accounts/fireworks/models/qwen3-30b-a3b | 131.072K | $0.15 | $0.6 | — |