This is a series of models designed to replicate the prose quality of the Claude 3 models, specifically Sonnet(https://openrouter.ai/anthropic/claude-3.5-sonnet) and Opus(https://openrouter.ai/anthropic/claude-3-opus). The model is fine-tuned on top of [Qwen2.5 72B](https://openrouter.ai/qwen/qwen-2.5-72b-instruct).

anthracite-org/magnum-v4-72b 32.768K context $2.5/M input $5/M output

No provider description is available for this model yet.

azure_ai/fw-kimi-k2.6 262.144K context $1.045/M input $4.4/M output

Llama 4 Maverick 17B Instruct (128E) is a high-capacity multimodal language model from Meta, built on a mixture-of-experts (MoE) architecture with 128 experts and 17 billion active parameters per forward...

meta-llama/llama-4-maverick 128K context $0.2/M input $0.696/M output

No provider description is available for this model yet.

fireworks_ai/kimi-k3-us 1.04858M context $3.3/M input $16.5/M output

Ling-2.6-flash is an instant (instruct) model from inclusionAI with 104B total parameters and 7.4B active parameters, designed for real-world agents that require fast responses, strong execution, and high token efficiency....

inclusionai/ling-2.6-flash 262.144K context $0.01/M input $0.03/M output

No provider description is available for this model yet.

fireworks_ai/muse-glimmer-30b 131.072K context $0.35/M input $1.5/M output

UnslopNemo v4.1 is the latest addition from the creator of Rocinante, designed for adventure writing and role-play scenarios.

thedrummer/unslopnemo-12b 1.024M context $0.4/M input $0.4/M output

No provider description is available for this model yet.

fireworks_ai/glm-5p2-fast 1.04858M context $2.1/M input $6.6/M output

No provider description is available for this model yet.

azure_ai/fw-minimax-m2.5 1M context $0.33/M input $1.32/M output

No provider description is available for this model yet.

azure_ai/fw-minimax-m3 512K context $0.33/M input $1.32/M output

No provider description is available for this model yet.

azure_ai/fw-nemotron-3-ultra-nvfp4 262.144K context $0.6/M input $2.4/M output

No provider description is available for this model yet.

fireworks_ai/deepseek-v4-flash-0731 1.04858M context $0.14/M input $0.28/M output

Qwen2.5-Coder is the latest series of Code-Specific Qwen large language models (formerly known as CodeQwen). Qwen2.5-Coder brings the following improvements upon CodeQwen1.5: - Significantly improvements in **code generation**, **code reasoning**...

qwen/qwen-2.5-coder-32b-instruct 32.768K context $0.66/M input $1/M output

No provider description is available for this model yet.

azure_ai/grok-4.3 200K context $1.25/M input $2.5/M output

No provider description is available for this model yet.

github_copilot/gpt-4-o-preview 64K context Input not listed Output not listed

No provider description is available for this model yet.

github_copilot/gpt-4.1 128K context Input not listed Output not listed

No provider description is available for this model yet.

github_copilot/gpt-4.1-2025-04-14 128K context Input not listed Output not listed

Skyfall 36B v2 is an enhanced iteration of Mistral Small 2501, specifically fine-tuned for improved creativity, nuanced writing, role-playing, and coherent storytelling.

thedrummer/skyfall-36b-v2 32.768K context $0.55/M input $0.8/M output

Relace Apply 3 is a specialized code-patching LLM that merges AI-suggested edits straight into your source files. It can apply updates from GPT-4o, Claude, and others into your files at...

relace/relace-apply-3 256K context $0.85/M input $1.25/M output

No provider description is available for this model yet.

azure_ai/fw-inkling 1.04858M context $1/M input $4.05/M output

No provider description is available for this model yet.

azure_ai/fw-glm-5.2-fast 1.04858M context $2.1/M input $6.6/M output

This is Mistral AI's flagship model, Mistral Large 2 (version mistral-large-2407). It's a proprietary weights-available model and excels at reasoning, code, JSON, chat, and more. Read the launch announcement [here](https://mistral.ai/news/mistral-large-2407/)....

mistralai/mistral-large-2407 131.072K context $2/M input $6/M output