This is a series of models designed to replicate the prose quality of the Claude 3 models, specifically Sonnet(https://openrouter.ai/anthropic/claude-3.5-sonnet) and Opus(https://openrouter.ai/anthropic/claude-3-opus). The model is fine-tuned on top of [Qwen2.5 72B](https://openrouter.ai/qwen/qwen-2.5-72b-instruct).
Models
Every model in the catalog with source-linked pricing, context limits, provider availability, and published benchmark results.
No provider description is available for this model yet.
Llama 4 Maverick 17B Instruct (128E) is a high-capacity multimodal language model from Meta, built on a mixture-of-experts (MoE) architecture with 128 experts and 17 billion active parameters per forward...
No provider description is available for this model yet.
Ling-2.6-flash is an instant (instruct) model from inclusionAI with 104B total parameters and 7.4B active parameters, designed for real-world agents that require fast responses, strong execution, and high token efficiency....
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
UnslopNemo v4.1 is the latest addition from the creator of Rocinante, designed for adventure writing and role-play scenarios.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
Qwen2.5-Coder is the latest series of Code-Specific Qwen large language models (formerly known as CodeQwen). Qwen2.5-Coder brings the following improvements upon CodeQwen1.5: - Significantly improvements in **code generation**, **code reasoning**...
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
Skyfall 36B v2 is an enhanced iteration of Mistral Small 2501, specifically fine-tuned for improved creativity, nuanced writing, role-playing, and coherent storytelling.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
Relace Apply 3 is a specialized code-patching LLM that merges AI-suggested edits straight into your source files. It can apply updates from GPT-4o, Claude, and others into your files at...
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
This is Mistral AI's flagship model, Mistral Large 2 (version mistral-large-2407). It's a proprietary weights-available model and excels at reasoning, code, JSON, chat, and more. Read the launch announcement [here](https://mistral.ai/news/mistral-large-2407/)....
| Model | Creator | Inputs | Context | Input | Output | Released | Compare |
|---|---|---|---|---|---|---|---|
| Magnum v4 72Banthracite-org/magnum-v4-72b | 32.768K | $2.5 | $5 | — | |||
| FW-Kimi-K2.6azure_ai/fw-kimi-k2.6 | 262.144K | $1.045 | $4.4 | — | |||
| Meta: Llama 4 Maverickmeta-llama/llama-4-maverick | 128K | $0.2 | $0.696 | — | |||
| kimi-k3-usfireworks_ai/kimi-k3-us | 1.04858M | $3.3 | $16.5 | — | |||
| inclusionAI: Ling-2.6-flashinclusionai/ling-2.6-flash | 262.144K | $0.01 | $0.03 | — | |||
| muse-glimmer-30bfireworks_ai/muse-glimmer-30b | 131.072K | $0.35 | $1.5 | — | |||
| us-gov.nvidia.nemotron-nano-9b-v2bedrock_converse/us-gov.nvidia.nemotron-nano-9b-v2 | 128K | $0.072 | $0.276 | — | |||
| us-gov-west-1/anthropic.claude-opus-5bedrock/us-gov-west-1/anthropic.claude-opus-5 | 1M | $6 | $30 | — | |||
| TheDrummer: UnslopNemo 12Bthedrummer/unslopnemo-12b | 1.024M | $0.4 | $0.4 | — | |||
| us-gov-west-1/anthropic.claude-fable-5-1bedrock/us-gov-west-1/anthropic.claude-fable-5-1 | 1M | $12 | $60 | — | |||
| glm-5p2-fastfireworks_ai/glm-5p2-fast | 1.04858M | $2.1 | $6.6 | — | |||
| FW-MiniMax-M2.5azure_ai/fw-minimax-m2.5 | 1M | $0.33 | $1.32 | — | |||
| FW-MiniMax-M3azure_ai/fw-minimax-m3 | 512K | $0.33 | $1.32 | — | |||
| FW-Nemotron-3-Ultra-NVFP4azure_ai/fw-nemotron-3-ultra-nvfp4 | 262.144K | $0.6 | $2.4 | — | |||
| deepseek-v4-flash-0731fireworks_ai/deepseek-v4-flash-0731 | 1.04858M | $0.14 | $0.28 | — | |||
| us-gov-east-1/nvidia.nemotron-nano-9b-v2bedrock/us-gov-east-1/nvidia.nemotron-nano-9b-v2 | 128K | $0.072 | $0.276 | — | |||
| Qwen2.5 Coder 32B Instructqwen/qwen-2.5-coder-32b-instruct | 32.768K | $0.66 | $1 | — | |||
| nemotron-3-ultra-nvfp4fireworks_ai/nemotron-3-ultra-nvfp4 | 262.144K | $0.6 | $2.4 | — | |||
| accounts/fireworks/models/muse-glimmer-30bfireworks_ai/accounts/fireworks/models/muse-glimmer-30b | 131.072K | $0.35 | $1.5 | — | |||
| accounts/fireworks/models/nemotron-lightning-3p5-30b-a3bfireworks_ai/accounts/fireworks/models/nemotron-lightning-3p5-30b-a3b | 262.144K | $0.05 | $0.2 | — | |||
| grok-4.3azure_ai/grok-4.3 | 200K | $1.25 | $2.5 | — | |||
| gpt-4-o-previewgithub_copilot/gpt-4-o-preview | 64K | — | — | — | |||
| accounts/fireworks/models/nemotron-3-ultra-nvfp4fireworks_ai/accounts/fireworks/models/nemotron-3-ultra-nvfp4 | 262.144K | $0.6 | $2.4 | — | |||
| us-gov-east-1/anthropic.claude-opus-5bedrock/us-gov-east-1/anthropic.claude-opus-5 | 1M | $6 | $30 | — | |||
| us-east-1/minimax.minimax-m2.1bedrock/us-east-1/minimax.minimax-m2.1 | 196K | $0.3 | $1.2 | — | |||
| gpt-4.1github_copilot/gpt-4.1 | 128K | — | — | — | |||
| gpt-4.1-2025-04-14github_copilot/gpt-4.1-2025-04-14 | 128K | — | — | — | |||
| accounts/fireworks/models/qwen3p8-maxfireworks_ai/accounts/fireworks/models/qwen3p8-max | 262.144K | $2 | $6 | — | |||
| TheDrummer: Skyfall 36B V2thedrummer/skyfall-36b-v2 | 32.768K | $0.55 | $0.8 | — | |||
| accounts/fireworks/routers/glm-5p2-fastfireworks_ai/accounts/fireworks/routers/glm-5p2-fast | 1.04858M | $2.1 | $6.6 | — | |||
| accounts/fireworks/routers/glm-5p2-fast-usfireworks_ai/accounts/fireworks/routers/glm-5p2-fast-us | 1.04858M | $2.1 | $6.6 | — | |||
| us-gov-east-1/anthropic.claude-fable-5-1bedrock/us-gov-east-1/anthropic.claude-fable-5-1 | 1M | $12 | $60 | — | |||
| us-east-1/minimax.minimax-m2.5bedrock/us-east-1/minimax.minimax-m2.5 | 1M | $0.3 | $1.2 | — | |||
| Relace: Relace Apply 3relace/relace-apply-3 | 256K | $0.85 | $1.25 | — | |||
| accounts/fireworks/routers/kimi-k3-fastfireworks_ai/accounts/fireworks/routers/kimi-k3-fast | 1.04858M | $4.5 | $22.5 | — | |||
| us-gov-west-1/xai.grok-4.6bedrock_mantle/us-gov-west-1/xai.grok-4.6 | 500K | $2.64 | $7.92 | — | |||
| accounts/fireworks/routers/kimi-k3-usfireworks_ai/accounts/fireworks/routers/kimi-k3-us | 1.04858M | $3.3 | $16.5 | — | |||
| us-gov-west-1/google.gemma-4-e2bbedrock_mantle/us-gov-west-1/google.gemma-4-e2b | 128K | $0.048 | $0.096 | — | |||
| us-gov-west-1/google.gemma-4-26b-a4bbedrock_mantle/us-gov-west-1/google.gemma-4-26b-a4b | 256K | $0.156 | $0.48 | — | |||
| us-gov-west-1/google.gemma-4-31bbedrock_mantle/us-gov-west-1/google.gemma-4-31b | 256K | $0.168 | $0.48 | — | |||
| us-gov-west-1/openai.gpt-oss-20bbedrock_mantle/us-gov-west-1/openai.gpt-oss-20b | 131.072K | $0.084 | $0.36 | — | |||
| us-gov-west-1/openai.gpt-oss-120bbedrock_mantle/us-gov-west-1/openai.gpt-oss-120b | 131.072K | $0.18 | $0.72 | — | |||
| us-gov-east-1/xai.grok-4.6bedrock_mantle/us-gov-east-1/xai.grok-4.6 | 500K | $2.64 | $7.92 | — | |||
| us-gov-east-1/openai.gpt-oss-20bbedrock_mantle/us-gov-east-1/openai.gpt-oss-20b | 131.072K | $0.084 | $0.36 | — | |||
| us-gov-east-1/openai.gpt-oss-120bbedrock_mantle/us-gov-east-1/openai.gpt-oss-120b | 131.072K | $0.18 | $0.72 | — | |||
| FW-Inklingazure_ai/fw-inkling | 1.04858M | $1 | $4.05 | — | |||
| us-east-1/moonshotai.kimi-k2-thinkingbedrock/us-east-1/moonshotai.kimi-k2-thinking | 262.144K | $0.6 | $2.5 | — | |||
| us-east-1/moonshotai.kimi-k2.5bedrock/us-east-1/moonshotai.kimi-k2.5 | 262.144K | $0.6 | $3 | — | |||
| FW-GLM-5.2-Fastazure_ai/fw-glm-5.2-fast | 1.04858M | $2.1 | $6.6 | — | |||
| Mistral Large 2407mistralai/mistral-large-2407 | 131.072K | $2 | $6 | — |