2,852 models

No provider description is available for this model yet.

azure_ai/global/grok-3 131.072K context $3/M input $15/M output

No provider description is available for this model yet.

azure_ai/global/grok-3-mini 131.072K context $0.25/M input $1.27/M output

No provider description is available for this model yet.

azure_ai/gpt-5.4-mini 272K context $0.75/M input $4.5/M output

No provider description is available for this model yet.

azure_ai/grok-3-mini 131.072K context $0.25/M input $1.27/M output

No provider description is available for this model yet.

github_copilot/gpt-4.1 128K context Input not listed Output not listed

No provider description is available for this model yet.

azure_ai/grok-4 131.072K context $3/M input $15/M output

No provider description is available for this model yet.

azure_ai/grok-4-fast-non-reasoning 131.072K context $0.2/M input $0.5/M output

No provider description is available for this model yet.

azure_ai/grok-4-fast-reasoning 131.072K context $0.2/M input $0.5/M output

GLM-5V-Turbo is Z.ai’s first native multimodal agent foundation model, built for vision-based coding and agent-driven tasks. It natively handles image, video, and text inputs, excels at long-horizon planning, complex coding,...

z-ai/glm-5v-turbo 202.752K context $1.2/M input $4/M output

Seed 2.1 Turbo is a multimodal model from ByteDance Seed for coding and long-horizon agent workflows. It is suited for end-to-end software delivery, multi-step task execution, and understanding visual and...

bytedance-seed/seed-2-1-turbo 262.144K context $0.5/M input $2.5/M output

No provider description is available for this model yet.

azure_ai/grok-4-1-fast-reasoning 131.072K context $0.2/M input $0.5/M output

No provider description is available for this model yet.

azure_ai/grok-code-fast-1 131.072K context $0.2/M input $1.5/M output

No provider description is available for this model yet.

azure_ai/kimi-k2.5 262.144K context $0.6/M input $3/M output

No provider description is available for this model yet.

azure_ai/kimi-k2.6 262.144K context $0.95/M input $4/M output

No provider description is available for this model yet.

azure_ai/ministral-3b 128K context $0.04/M input $0.04/M output

No provider description is available for this model yet.

azure_ai/mistral-large-2407 128K context $2/M input $6/M output

No provider description is available for this model yet.

azure_ai/mistral-large-3 256K context $0.5/M input $1.5/M output

No provider description is available for this model yet.

azure_ai/mistral-medium-2505 131.072K context $0.4/M input $2/M output

No provider description is available for this model yet.

azure_ai/mistral-nemo 131.072K context $0.15/M input $0.15/M output

No provider description is available for this model yet.

azure_ai/gpt-5.4-2026-03-05 1.05M context $2.5/M input $15/M output

Gemini 2.5 Pro is Google’s state-of-the-art AI model designed for advanced reasoning, coding, mathematics, and scientific tasks. It employs “thinking” capabilities, enabling it to reason through responses with enhanced accuracy...

google/gemini-2.5-pro-preview-05-06 1.04858M context $1.25/M input $10/M output

Llama 3.2 3B is a 3-billion-parameter multilingual large language model, optimized for advanced natural language processing tasks like dialogue generation, reasoning, and summarization. Designed with the latest transformer architecture, it...

meta-llama/llama-3.2-3b-instruct:free 131.072K context Free input Free output

Gemini 2.5 Flash-Lite is a lightweight reasoning model in the Gemini 2.5 family, optimized for ultra-low latency and cost efficiency. It offers improved throughput, faster token generation, and better performance...

google/gemini-2.5-flash-lite:batch 1.04858M context $0.05/M input $0.2/M output

Gemini 3.6 Flash is a high-efficiency model from Google for coding, agentic workflows, and web and app development. It is designed to produce polished outputs with fewer unnecessary edits and...

google/gemini-3.6-flash:batch 1.04858M context $0.375/M input $1.875/M output

Qwen3.5 Plus (April 2026) is a large-scale multimodal language model from Alibaba. It accepts text, image, and video input and produces text output, with a 1M token context window. This...

qwen/qwen3.5-plus-20260420 1M context $0.3/M input $1.8/M output

No provider description is available for this model yet.

bedrock/moonshotai.kimi-k2-thinking 262.144K context $0.73/M input $3.03/M output

No provider description is available for this model yet.

anthropic/claude-4-opus-20250514 200K context $15/M input $75/M output

No provider description is available for this model yet.

azure_ai/grok-4.3 200K context $1.25/M input $2.5/M output

No provider description is available for this model yet.

cloudflare/@cf/openai/gpt-oss-20b 128K context $0.2/M input $0.3/M output