4,213 models
Open weights

No provider description is available for this model yet.

google/gemma-4-12B-it-assistant Not documented context Input not listed Output not listed

No provider description is available for this model yet.

fireworks_ai/muse-glimmer-30b 131.072K context $0.35/M input $1.5/M output

No provider description is available for this model yet.

fireworks_ai/qwen3p8-max 262.144K context $2/M input $6/M output

Qwen3 Coder Flash is Alibaba's fast and cost efficient version of their proprietary Qwen3 Coder Plus. It is a powerful coding agent model specializing in autonomous programming via tool calling...

qwen/qwen3-coder-flash 1M context $0.195/M input $0.975/M output

Euryale L3.1 70B v2.2 is a model focused on creative roleplay from [Sao10k](https://ko-fi.com/sao10k). It is the successor of [Euryale L3 70B v2.1](/models/sao10k/l3-euryale-70b).

sao10k/l3.1-euryale-70b 131.072K context $0.85/M input $0.85/M output

Mistral Saba is a 24B-parameter language model specifically designed for the Middle East and South Asia, delivering accurate and contextually relevant responses while maintaining efficient performance. Trained on curated regional...

mistralai/mistral-saba 32.768K context $0.2/M input $0.6/M output

Morph's high-accuracy apply model for complex code edits. ~4,500 tokens/sec with 98% accuracy for precise code transformations. The model requires the prompt to be in the following format: <instruction>{instruction}</instruction> <code>{initial_code}</code>...

morph/morph-v3-large 262.144K context $0.9/M input $1.9/M output

Qwen3-Coder-480B-A35B-Instruct is a Mixture-of-Experts (MoE) code generation model developed by the Qwen team. It is optimized for agentic coding tasks such as function calling, tool use, and long-context reasoning over...

qwen/qwen3-coder:free 262K context Free input Free output

No provider description is available for this model yet.

fireworks_ai/kimi-k3-us 1.04858M context $3.3/M input $16.5/M output

No provider description is available for this model yet.

together_ai/together-ai-up-to-4b Not documented context $0.1/M input $0.1/M output

No provider description is available for this model yet.

azure_ai/fw-kimi-k2.7-code 262.144K context $1.05/M input $4.4/M output

No provider description is available for this model yet.

together_ai/together-ai-81.1b-110b Not documented context $1.8/M input $1.8/M output

No provider description is available for this model yet.

azure_ai/fw-kimi-k2.6 262.144K context $1.045/M input $4.4/M output

No provider description is available for this model yet.

together_ai/together-ai-8.1b-21b 1K context $0.3/M input $0.3/M output

No provider description is available for this model yet.

fireworks_ai/kimi-k3-fast 1.04858M context $4.5/M input $22.5/M output

The largest model in the Ministral 3 family, Ministral 3 14B offers frontier capabilities and performance comparable to its larger Mistral Small 3.2 24B counterpart. A powerful and efficient language...

mistralai/ministral-14b-2512 262.144K context $0.2/M input $0.2/M output
Open weights

No provider description is available for this model yet.

google/gemma-4-12B-it Not documented context Input not listed Output not listed

No provider description is available for this model yet.

scx-ai/qwen3.8-max 1M context $1.65/M input $4.99/M output

No provider description is available for this model yet.

bedrock/eu-north-1/deepseek.v3.2 163.84K context $0.74/M input $2.22/M output

No provider description is available for this model yet.

scx-ai/glm-5.2 1.04858M context $0.61/M input $1.98/M output

Venice Uncensored Dolphin Mistral 24B Venice Edition is a fine-tuned variant of Mistral-Small-24B-Instruct-2501, developed by dphn.ai in collaboration with Venice.ai. This model is designed as an “uncensored” instruct-tuned LLM, preserving...

cognitivecomputations/dolphin-mistral-24b-venice-edition:free 32.768K context Free input Free output
Open weights

No provider description is available for this model yet.

microsoft/Mage-Flow-Base Not documented context Input not listed Output not listed
Open weights

No provider description is available for this model yet.

google/gemma-4-12B Not documented context Input not listed Output not listed

NVIDIA Nemotron 3 Nano 30B A3B is a small language MoE model with highest compute efficiency and accuracy for developers to build specialized agentic AI systems. The model is fully...

nvidia/nemotron-3-nano-30b-a3b:free 256K context Free input Free output

No provider description is available for this model yet.

azure/gpt-audio-mini 128K context $0.6/M input $2.4/M output