3,424 models

KAT-Coder-Pro V2.5 is a flagship-level Agentic Coding model that can directly hand over an entire issue or an entire business workflow to it, allowing it to autonomously locate and make...

kwaipilot/kat-coder-pro-v2.5:free 256K context Free input Free output

No provider description is available for this model yet.

fireworks_ai/muse-glimmer-30b 131.072K context $0.35/M input $1.5/M output

Lunaris 8B is a versatile generalist and roleplaying model based on Llama 3. It's a strategic merge of multiple models, designed to balance creativity with improved logic and general knowledge....

sao10k/l3-lunaris-8b 8.192K context $0.04/M input $0.05/M output

NVIDIA Nemotron 3 Super is a 120B-parameter open hybrid MoE model, activating just 12B parameters for maximum compute efficiency and accuracy in complex multi-agent applications. Built on a hybrid Mamba-Transformer...

nvidia/nemotron-3-super-120b-a12b:free 262.144K context Free input Free output

KAT-Coder-Air V2.5 is a flagship-level Agentic Coding model that can directly hand over an entire issue or an entire business workflow to it, allowing it to autonomously locate and make...

kwaipilot/kat-coder-air-v2.5 256K context $0.15/M input $0.6/M output

No provider description is available for this model yet.

fireworks_ai/qwen3p8-max 262.144K context $2/M input $6/M output

No provider description is available for this model yet.

fireworks_ai/kimi-k3-us 1.04858M context $3.3/M input $16.5/M output

No provider description is available for this model yet.

azure_ai/fw-kimi-k2.7-code 262.144K context $1.05/M input $4.4/M output

No provider description is available for this model yet.

azure_ai/fw-kimi-k2.6 262.144K context $1.045/M input $4.4/M output

Hermes 3 is a generalist language model with many improvements over Hermes 2, including advanced agentic capabilities, much better roleplaying, reasoning, multi-turn conversation, long context coherence, and improvements across the...

nousresearch/hermes-3-llama-3.1-405b 131.072K context $1/M input $1/M output

No provider description is available for this model yet.

fireworks_ai/kimi-k3-fast 1.04858M context $4.5/M input $22.5/M output

Venice Uncensored Dolphin Mistral 24B Venice Edition is a fine-tuned variant of Mistral-Small-24B-Instruct-2501, developed by dphn.ai in collaboration with Venice.ai. This model is designed as an “uncensored” instruct-tuned LLM, preserving...

cognitivecomputations/dolphin-mistral-24b-venice-edition:free 32.768K context Free input Free output

Hermes 3 is a generalist language model with many improvements over [Hermes 2](/models/nousresearch/nous-hermes-2-mistral-7b-dpo), including advanced agentic capabilities, much better roleplaying, reasoning, multi-turn conversation, long context coherence, and improvements across the...

nousresearch/hermes-3-llama-3.1-70b 131.072K context $0.7/M input $0.7/M output

Reka Edge is an extremely efficient 7B multimodal vision-language model that accepts image/video+text inputs and generates text outputs. This model is optimized specifically to deliver industry-leading performance in image understanding,...

rekaai/reka-edge 16.384K context $0.1/M input $0.1/M output

No provider description is available for this model yet.

fireworks_ai/kimi-k3 1.04858M context $3/M input $15/M output

No provider description is available for this model yet.

fireworks_ai/glm-5p2-fast-us 1.04858M context $2.1/M input $6.6/M output

Euryale L3.1 70B v2.2 is a model focused on creative roleplay from [Sao10k](https://ko-fi.com/sao10k). It is the successor of [Euryale L3 70B v2.1](/models/sao10k/l3-euryale-70b).

sao10k/l3.1-euryale-70b 131.072K context $0.85/M input $0.85/M output

No provider description is available for this model yet.

azure_ai/fw-kimi-k2.5 262.144K context $0.66/M input $3.3/M output

No provider description is available for this model yet.

fireworks_ai/glm-5p2-fast 1.04858M context $2.1/M input $6.6/M output

Qwen3.5 Plus (April 2026) is a large-scale multimodal language model from Alibaba. It accepts text, image, and video input and produces text output, with a 1M token context window. This...

qwen/qwen3.5-plus-20260420 1M context $0.3/M input $1.8/M output

KAT-Coder-Pro V2 is the latest high-performance model in KwaiKAT’s KAT-Coder series, designed for complex enterprise-grade software engineering and SaaS integration. It builds on the agentic coding strengths of earlier versions,...

kwaipilot/kat-coder-pro-v2 262.144K context $0.3/M input $1.2/M output

Auto Router (Beta) is a task-aware router from OpenRouter. It classifies each request, then routes it the [most popular model](/rankings#task-spend) for that task based on aggregate spend, filtered by your...

openrouter/auto-beta 2M context Input not listed Output not listed

Rocinante 12B is designed for engaging storytelling and rich prose. Early testers have reported: - Expanded vocabulary with unique and expressive word choices - Enhanced creativity for vivid narratives -...

thedrummer/rocinante-12b 65.536K context $0.25/M input $0.5/M output

NVIDIA-Nemotron-Nano-9B-v2 is a large language model (LLM) trained from scratch by NVIDIA, and designed as a unified model for both reasoning and non-reasoning tasks. It responds to user queries and...

nvidia/nemotron-nano-9b-v2:free 128K context Free input Free output

No provider description is available for this model yet.

fireworks_ai/deepseek-v4-flash-0731 1.04858M context $0.14/M input $0.28/M output

No provider description is available for this model yet.

azure_ai/fw-inkling 1.04858M context $1/M input $4.05/M output

No provider description is available for this model yet.

azure_ai/fw-glm-5.2-fast 1.04858M context $2.1/M input $6.6/M output

Amazon Nova Micro 1.0 is a text-only model that delivers the lowest latency responses in the Amazon Nova family of models at a very low cost. With a context length...

amazon/nova-micro-v1 128K context $0.035/M input $0.14/M output

No provider description is available for this model yet.

azure_ai/fw-glm-5.2 1.04858M context $1.54/M input $4.84/M output

No provider description is available for this model yet.

azure_ai/fw-glm-5.1 202.8K context $1.54/M input $4.84/M output

Amazon Nova Lite 1.0 is a very low-cost multimodal model from Amazon that focused on fast processing of image, video, and text inputs to generate text output. Amazon Nova Lite...

amazon/nova-lite-v1 300K context $0.06/M input $0.24/M output

Mistral Small 4 is the next major release in the Mistral Small family, unifying the capabilities of several flagship Mistral models into a single system. It combines strong reasoning from...

mistralai/mistral-small-2603 262.144K context $0.15/M input $0.6/M output

The Meta Llama 3.3 multilingual large language model (LLM) is a pretrained and instruction tuned generative model in 70B (text in/text out). The Llama 3.3 instruction tuned text only model...

meta-llama/llama-3.3-70b-instruct 131.072K context $0.1/M input $0.32/M output

No provider description is available for this model yet.

ollama/mistral 8.192K context Input not listed Output not listed

No provider description is available for this model yet.

bedrock/us-east-1/deepseek.v3.2 163.84K context $0.62/M input $1.85/M output