4,215 models

No provider description is available for this model yet.

bedrock/us-east-1/zai.glm-5 200K context $1/M input $3.2/M output

No provider description is available for this model yet.

bedrock/us-west-2/zai.glm-5 200K context $1/M input $3.2/M output

No provider description is available for this model yet.

snowflake/claude-sonnet-4-5 200K context $3/M input $15/M output

No provider description is available for this model yet.

snowflake/claude-sonnet-4-6 200K context $3/M input $15/M output

No provider description is available for this model yet.

snowflake/claude-4-opus 200K context $5/M input $25/M output

No provider description is available for this model yet.

snowflake/claude-haiku-4-5 200K context $1/M input $5/M output

No provider description is available for this model yet.

snowflake/claude-3-7-sonnet 200K context $3/M input $15/M output

No provider description is available for this model yet.

snowflake/openai-gpt-4.1 300K context $2/M input $8/M output

No provider description is available for this model yet.

snowflake/openai-gpt-5 300K context $1.25/M input $10/M output

No provider description is available for this model yet.

snowflake/openai-gpt-5-mini 1M context $0.3/M input $1.2/M output

No provider description is available for this model yet.

snowflake/openai-gpt-5-nano 5M context $0.15/M input $0.6/M output

No provider description is available for this model yet.

snowflake/llama4-maverick 128K context $0.24/M input $0.97/M output

No provider description is available for this model yet.

tensormesh/qwen/qwen3.6-27b-fp8 262.144K context $0.32/M input $3.2/M output

No provider description is available for this model yet.

tensormesh/moonshotai/kimi-k2.6 32.768K context $0.96/M input $4/M output

No provider description is available for this model yet.

tensormesh/minimaxai/minimax-m2.5 196.608K context $0.3/M input $1.2/M output

No provider description is available for this model yet.

tensormesh/google/gemma-4-31b-it 32.768K context $0.14/M input $0.56/M output

No provider description is available for this model yet.

tensormesh/openai/gpt-oss-120b 131.072K context $0.15/M input $0.6/M output

No provider description is available for this model yet.

tensormesh/openai/gpt-oss-20b 131.072K context $0.07/M input $0.28/M output

No provider description is available for this model yet.

tencent/deepseek-v4-pro 1M context $0.435/M input $0.87/M output

No provider description is available for this model yet.

tencent/deepseek-v4-flash 1M context $0.14/M input $0.28/M output

No provider description is available for this model yet.

pinstripes/ps/glm-4.5-air 128K context $0.125/M input $0.45/M output

No provider description is available for this model yet.

pinstripes/ps/qwen3.6-35b-a3b 131.072K context $0.14/M input $0.45/M output

No provider description is available for this model yet.

pinstripes/ps/qwen3-30b-a3b 131.072K context $0.09/M input $0.2/M output

No provider description is available for this model yet.

pinstripes/ps/qwen3-coder-30b-a3b 131.072K context $0.3/M input $0.6/M output

No provider description is available for this model yet.

pinstripes/ps/deepseek-v4-flash 163.84K context $0.1/M input $0.2/M output

No provider description is available for this model yet.

pinstripes/ps/minimax-m2.7 1.00019M context $0.255/M input $0.55/M output

No provider description is available for this model yet.

darkbloom/gemma-4-26b 131.072K context $0.03/M input $0.165/M output

No provider description is available for this model yet.

darkbloom/gpt-oss-20b 131.072K context $0.015/M input $0.07/M output

No provider description is available for this model yet.

unknown/fallback_generalizations Not documented context Input not listed Output not listed

Mercury 2 is an extremely fast reasoning LLM, and the first reasoning diffusion LLM (dLLM). Instead of generating tokens sequentially, Mercury 2 produces and refines multiple tokens in parallel, achieving...

inception/mercury-2 128K context $0.25/M input $0.75/M output

A cost-efficient version of GPT Audio. The new snapshot features an upgraded decoder for more natural sounding voices and maintains better voice consistency. Input is priced at $0.60 per million...

openai/gpt-audio-mini 128K context $0.6/M input $2.4/M output

LFM2.5-2.6B is a compact reasoning model from Liquid AI. It is suited for agent workflows, data extraction, RAG, and long-context processing. Liquid advises against using it for agentic coding or...

liquid/lfm-2.5-2.6b:free 65.536K context Free input Free output

Mistral Medium 3.1 is an updated version of Mistral Medium 3, which is a high-performance enterprise-grade language model designed to deliver frontier-level capabilities at significantly reduced operational cost. It balances...

mistralai/mistral-medium-3.1 131.072K context $0.4/M input $2/M output

No provider description is available for this model yet.

replicate/openai/gpt-oss-20b Not documented context $0.09/M input $0.36/M output

Gemini 2.5 Flash is Google's state-of-the-art workhorse model, specifically designed for advanced reasoning, coding, mathematics, and scientific tasks. It includes built-in "thinking" capabilities, enabling it to provide responses with greater...

google/gemini-2.5-flash:batch 1.04858M context $0.15/M input $1.25/M output

Ministral 8B is an 8B parameter model featuring a unique interleaved sliding-window attention pattern for faster, memory-efficient inference. Designed for edge use cases, it supports up to 128k context length...

mistralai/ministral-8b 128K context $0.11/M input $0.11/M output

GLM-5 is Z.ai’s flagship open-source foundation model engineered for complex systems design and long-horizon agent workflows. Built for expert developers, it delivers production-grade performance on large-scale programming tasks, rivaling leading...

z-ai/glm-5 198K context $0.6/M input $1.92/M output