3,311 models

No provider description is available for this model yet.

aihubmix/gpt-5.6-sol-disc 1.05M context $4/M input $20/M output

Mercury 2 is an extremely fast reasoning LLM, and the first reasoning diffusion LLM (dLLM). Instead of generating tokens sequentially, Mercury 2 produces and refines multiple tokens in parallel, achieving...

inception/mercury-2 128K context $0.25/M input $0.75/M output

No provider description is available for this model yet.

aihubmix/gpt-5.6-terra 1.05M context $2/M input $12/M output

GPT-4o mini is OpenAI's newest model after [GPT-4 Omni](/models/openai/gpt-4o), supporting both text and image inputs with text outputs. As their most advanced small model, it is many multiples more affordable...

openai/gpt-4o-mini:batch 128K context $0.075/M input $0.3/M output

Nex-N2.5 is an agentic model built to turn goals into working, verified outcomes. Its core strength is agentic coding within a visual feedback loop: it can explore codebases, implement multi-file...

nex-agi/nex-n2.5-pro:free 262.144K context Free input Free output

No provider description is available for this model yet.

aihubmix/gpt-chat-latest 400K context $5/M input $30/M output

As a 30B-class SOTA model, GLM-4.7-Flash offers a new option that balances performance and efficiency. It is further optimized for agentic coding use cases, strengthening coding capabilities, long-horizon task planning,...

z-ai/glm-4.7-flash 131.072K context $0.061/M input $0.4/M output

No provider description is available for this model yet.

aihubmix/grok-4-20-reasoning 1M context $2/M input $6/M output

No provider description is available for this model yet.

aihubmix/grok-4.6 500K context $2/M input $6/M output

No provider description is available for this model yet.

aihubmix/grok-build-0.1 256K context $1/M input $2/M output

No provider description is available for this model yet.

aihubmix/hy3 256K context $0.156/M input $0.625/M output

No provider description is available for this model yet.

aihubmix/hy4-preview 1.04858M context $0.845/M input $2.535/M output

No provider description is available for this model yet.

aihubmix/kimi-k2.6 262.144K context $0.95/M input $4/M output

No provider description is available for this model yet.

aihubmix/kimi-k2.7-code-highspeed 262.144K context $1.9/M input $7.999/M output

No provider description is available for this model yet.

aihubmix/kimi-k3 1.04858M context $3/M input $15/M output

No provider description is available for this model yet.

aihubmix/longcat-2.0 1M context $0.775/M input $3.098/M output

GPT-5.5 Pro is OpenAI’s high-capability model optimized for deep reasoning and accuracy on complex, high-stakes workloads. It features a 1M+ token context window (922K input, 128K output) with support for...

openai/gpt-5.5-pro:batch 1.05M context $15/M input $90/M output

No provider description is available for this model yet.

aihubmix/mai-thinking-1 256K context $2/M input $8/M output

Gemini 3 Flash Preview is a high speed, high value thinking model designed for agentic workflows, multi turn chat, and coding assistance. It delivers near Pro level reasoning and tool...

google/gemini-3-flash-preview:batch 1.04858M context $0.25/M input $1.5/M output

No provider description is available for this model yet.

aihubmix/mimo-v2-omni 256K context $0.44/M input $2.2/M output

No provider description is available for this model yet.

aihubmix/mimo-v2-pro 1M context $1.1/M input $3.3/M output

No provider description is available for this model yet.

aihubmix/minimax-m2.7 204.8K context $0.296/M input $1.183/M output

No provider description is available for this model yet.

databricks/databricks-glm-5-2 1M context $1.4/M input $4.4/M output

No provider description is available for this model yet.

databricks/databricks-kimi-k3 1M context $3/M input $15/M output

No provider description is available for this model yet.

aihubmix/minimax-m3 1M context $0.288/M input $1.152/M output

Claude Opus 4.8 is Anthropic's most capable generally available model in the Opus family. It supports text, image, and file inputs with text output, with reasoning support and a 1M-token...

anthropic/claude-opus-4.8:batch 1M context $2.5/M input $12.5/M output

No provider description is available for this model yet.

mistral/ministral-14b-2512 262.144K context $0.2/M input $0.2/M output

No provider description is available for this model yet.

mistral/ministral-14b-latest 262.144K context $0.2/M input $0.2/M output

No provider description is available for this model yet.

aihubmix/muse-spark-1.2 1.04858M context $1.375/M input $4.675/M output

No provider description is available for this model yet.

mistral/ministral-3b-2512 131.072K context $0.1/M input $0.1/M output

No provider description is available for this model yet.

mistral/ministral-3b-latest 131.072K context $0.1/M input $0.1/M output

No provider description is available for this model yet.

mistral/mistral-medium-3 262.144K context $1.5/M input $7.5/M output

No provider description is available for this model yet.

mistral/voxtral-small-2507 32.768K context $0.1/M input $0.4/M output

No provider description is available for this model yet.

together_ai/zai-org/glm-5.3-flash 1.04858M context $0.15/M input $0.5/M output

MiniMax-M1 is a large-scale, open-weight reasoning model designed for extended context and high-efficiency inference. It leverages a hybrid Mixture-of-Experts (MoE) architecture paired with a custom "lightning attention" mechanism, allowing it...

minimax/minimax-m1 1M context $0.4/M input $2.2/M output

No provider description is available for this model yet.

zai/glm-5.3 1M context $1.4/M input $4.4/M output

No provider description is available for this model yet.

tencent/minimax-m3 1M context $0.3/M input $1.2/M output

No provider description is available for this model yet.

novita/zai-org/glm-5.3 1.04858M context $1.4/M input $4.4/M output

No provider description is available for this model yet.

novita/moonshotai/kimi-k3 1.04858M context $3/M input $15/M output

No provider description is available for this model yet.

novita/tencent/hy3 262.144K context $0.14/M input $0.58/M output

No provider description is available for this model yet.

novita/zai-org/glm-5.2 1.04858M context $1.4/M input $4.4/M output

No provider description is available for this model yet.

aihubmix/qwen3-coder-next 262.144K context $0.137/M input $0.548/M output

No provider description is available for this model yet.

novita/moonshotai/kimi-k2.7-code 262.144K context $0.95/M input $4/M output

Claude Opus 4.5 is Anthropic’s frontier reasoning model optimized for complex software engineering, agentic workflows, and long-horizon computer use. It offers strong multimodal capabilities, competitive performance across real-world coding and...

anthropic/claude-opus-4.5:batch 200K context $2.5/M input $12.5/M output

Fugu Max is the cost-performance model in Sakana AI's Fugu family. Rather than a single monolithic model, Fugu is a learned multi-agent orchestration system: a language model trained to route...

sakana/fugu-max 1M context $2/M input $6/M output

Grok 4.3 is a reasoning model from SpaceXAI. It accepts text and image inputs with text output, and is suited for agentic workflows, instruction-following tasks, and applications requiring high factual...

x-ai/grok-4.3 1M context $1.25/M input $2.5/M output