3,239 models

Claude Opus 4 is benchmarked as the world’s best coding model, at time of release, bringing sustained performance on complex, long-running tasks and agent workflows. It sets new benchmarks in...

anthropic/claude-opus-4 200K context $15/M input $75/M output

No provider description is available for this model yet.

novita/baidu/ernie-4.5-21b-a3b 120K context $0.07/M input $0.28/M output

Mistral Medium 3 is a high-performance enterprise-grade language model designed to deliver frontier-level capabilities at significantly reduced operational cost. It balances state-of-the-art reasoning and multimodal performance with 8× lower cost...

mistralai/mistral-medium-3 131.072K context $0.4/M input $2/M output

Llama 3.2 11B Vision is a multimodal model with 11 billion parameters, designed to handle tasks combining visual and textual data. It excels in tasks such as image captioning and...

meta-llama/llama-3.2-11b-vision-instruct 131.072K context $0.345/M input $0.345/M output

Hermes 3 is a generalist language model with many improvements over Hermes 2, including advanced agentic capabilities, much better roleplaying, reasoning, multi-turn conversation, long context coherence, and improvements across the...

nousresearch/hermes-3-llama-3.1-405b:free 131.072K context Free input Free output

No provider description is available for this model yet.

gmi/deepseek-ai/deepseek-v3-0324 163.84K context $0.28/M input $0.88/M output

A balanced model in the Ministral 3 family, Ministral 3 8B is a powerful, efficient tiny language model with vision capabilities.

mistralai/ministral-8b-2512 262.144K context $0.15/M input $0.15/M output

No provider description is available for this model yet.

gmi/google/gemini-3-pro-preview 1.04858M context $2/M input $12/M output

No provider description is available for this model yet.

gmi/google/gemini-3-flash-preview 1.04858M context $0.5/M input $3/M output

No provider description is available for this model yet.

gmi/moonshotai/kimi-k2-thinking 262.144K context $0.8/M input $1.2/M output

No provider description is available for this model yet.

gmi/minimaxai/minimax-m2.1 196.608K context $0.3/M input $1.2/M output

Coder‑Large is a 32 B‑parameter offspring of Qwen 2.5‑Instruct that has been further trained on permissively‑licensed GitHub, CodeSearchNet and synthetic bug‑fix corpora. It supports a 32k context window, enabling multi‑file...

arcee-ai/coder-large 32.768K context $0.5/M input $0.8/M output

Aion-RP-Llama-3.1-8B ranks the highest in the character evaluation portion of the RPBench-Auto benchmark, a roleplaying-specific variant of Arena-Hard-Auto, where LLMs evaluate each other’s responses. It is a fine-tuned base model...

aion-labs/aion-rp-llama-3.1-8b 32.768K context $0.8/M input $1.6/M output

The latest GPT-4 Turbo model with vision capabilities. Vision requests can now use JSON mode and function calling. Training data: up to December 2023.

openai/gpt-4-turbo:batch 128K context $5/M input $15/M output

No provider description is available for this model yet.

openai/chatgpt-4o-latest 128K context $5/M input $15/M output

This model always redirects to the latest model in the GLM Flash family.

~z-ai/glm-flash-latest 1.04858M context $0.075/M input $0.25/M output

No provider description is available for this model yet.

gmi/zai-org/glm-4.7-fp8 202.752K context $0.4/M input $2/M output

No provider description is available for this model yet.

novita/zai-org/glm-4.5-air 131.072K context $0.13/M input $0.85/M output

No provider description is available for this model yet.

scx-ai/glm-5.2 1.04858M context $0.61/M input $1.98/M output

No provider description is available for this model yet.

novita/qwen/qwen3-vl-8b-instruct 131.072K context $0.08/M input $0.5/M output