No provider description is available for this model yet.

nvidia/Kimi-K2.7-Code-DFlash Not documented context Input not listed Output not listed

Ling 3.0 Flash Fin is a finance-focused mixture-of-experts model from InclusionAI, built on Ling 3.0 Flash with 5.1B active parameters out of 124B total. It is designed for real-world investment...

inclusionai/ling-3.0-flash-fin:free 262.144K context Free input Free output

No provider description is available for this model yet.

azure/global/gpt-4o-2024-08-06 128K context $2.5/M input $10/M output

No provider description is available for this model yet.

azure/global/gpt-4o-2024-11-20 128K context $2.5/M input $10/M output

No provider description is available for this model yet.

nvidia/Privasis-Cleaner-4B Not documented context Input not listed Output not listed

No provider description is available for this model yet.

nvidia/Privasis-Cleaner-0.6B Not documented context Input not listed Output not listed

No provider description is available for this model yet.

nvidia/LocateAnything-3B Not documented context Input not listed Output not listed

No provider description is available for this model yet.

nvidia/Kimi-K2.6-Eagle3 Not documented context Input not listed Output not listed

Ox Alpha is a reasoning model designed for coding, sustained agentic work, and production workloads. It is suited for long-horizon software engineering, complex reasoning, and workflows that combine text with...

stealth/ox-alpha 1.04858M context Input not listed Output not listed

No provider description is available for this model yet.

nvidia/CUDA-Autocomplete Not documented context Input not listed Output not listed

No provider description is available for this model yet.

azure/global/gpt-5.1-chat 128K context $1.25/M input $10/M output

No provider description is available for this model yet.

gmi/zai-org/glm-4.7-fp8 202.752K context $0.4/M input $2/M output

Qwen3-Coder-Next is an open-weight causal language model optimized for coding agents and local development workflows. It uses a sparse MoE design with 80B total parameters and only 3B activated per...

qwen/qwen3-coder-next 262.144K context $0.12/M input $0.8/M output

Amazon Nova Premier is the most capable of Amazon’s multimodal models for complex reasoning tasks and for use as the best teacher for distilling custom models.

amazon/nova-premier-v1 1M context $2.5/M input $12.5/M output

A cost-efficient version of GPT Audio. The new snapshot features an upgraded decoder for more natural sounding voices and maintains better voice consistency. Input is priced at $0.60 per million...

openai/gpt-audio-mini 128K context $0.6/M input $2.4/M output

No provider description is available for this model yet.

azure/gpt-3.5-turbo-0125 16.384K context $0.5/M input $1.5/M output

No provider description is available for this model yet.

azure/gpt-35-turbo 4.097K context $0.5/M input $1.5/M output

No provider description is available for this model yet.

azure/gpt-35-turbo-0125 16.384K context $0.5/M input $1.5/M output

MiniMax-M3 is a multimodal foundation model from MiniMax. It supports text, image, and video inputs with text output, a 1M-token context window, and is suited for long-horizon agentic work, coding,...

minimax/minimax-m3 524.288K context $0.3/M input $1.2/M output

No provider description is available for this model yet.

azure/gpt-35-turbo-1106 16.384K context $1/M input $2/M output

No provider description is available for this model yet.

azure/gpt-35-turbo-16k 16.385K context $3/M input $4/M output

No provider description is available for this model yet.

azure/gpt-35-turbo-16k-0613 16.385K context $3/M input $4/M output

No provider description is available for this model yet.

azure_text/gpt-35-turbo-instruct 4.097K context $1.5/M input $2/M output

No provider description is available for this model yet.

baseten/deepseek-ai/deepseek-v3-0324 Not documented context $0.77/M input $0.77/M output

No provider description is available for this model yet.

friendliai/zai-org/glm-5.2 1.04858M context $1.4/M input $4.4/M output