1,395 models

No provider description is available for this model yet.

azure_ai/phi-3.5-vision-instruct 128K context $0.13/M input $0.52/M output

No provider description is available for this model yet.

azure_ai/phi-4-multimodal-instruct 131.072K context $0.08/M input $0.32/M output

Mistral Medium 3 is a high-performance enterprise-grade language model designed to deliver frontier-level capabilities at significantly reduced operational cost. It balances state-of-the-art reasoning and multimodal performance with 8× lower cost...

mistralai/mistral-medium-3 131.072K context $0.4/M input $2/M output

Ling 3.0 Flash VL builds on Ling 3.0 Flash (124B total / 5.5B active MoE from InclusionAI), further strengthening its language capabilities while adding native visual perception and advanced visual...

inclusionai/ling-3.0-flash-vl:free 262.144K context Free input Free output

Claude Opus 4.1 is an updated version of Anthropic’s flagship model, offering improved performance in coding, reasoning, and agentic tasks. It achieves 74.5% on SWE-bench Verified and shows notable gains...

anthropic/claude-opus-4.1 200K context $15/M input $75/M output

No provider description is available for this model yet.

azure_ai/kimi-k2.5 262.144K context $0.6/M input $3/M output

No provider description is available for this model yet.

azure_ai/kimi-k2.6 262.144K context $0.95/M input $4/M output

No provider description is available for this model yet.

azure_ai/mistral-large-3 256K context $0.5/M input $1.5/M output

No provider description is available for this model yet.

azure_ai/mistral-small-2503 128K context $0.1/M input $0.3/M output

Gemini 3.6 Flash is a high-efficiency model from Google for coding, agentic workflows, and web and app development. It is designed to produce polished outputs with fewer unnecessary edits and...

google/gemini-3.6-flash:batch 1.04858M context $0.375/M input $1.875/M output

Inkling is an open-weight multimodal mixture-of-experts model from Thinking Machines Lab, with 41B active parameters out of 975B total. It is designed for general-purpose reasoning, coding, agentic and tool-use systems,...

thinkingmachines/inkling:batch 524.288K context $1/M input $4.05/M output

No provider description is available for this model yet.

bedrock/moonshotai.kimi-k2.5 262.144K context $0.6/M input $3.03/M output

No provider description is available for this model yet.

anthropic/claude-4-opus-20250514 200K context $15/M input $75/M output

No provider description is available for this model yet.

dashscope/qwen3-vl-32b-instruct 131.072K context $0.16/M input $0.64/M output

No provider description is available for this model yet.

dashscope/qwen3-vl-32b-thinking 131.072K context $0.16/M input $2.87/M output

No provider description is available for this model yet.

dashscope/qwen3-vl-plus 260.096K context Input not listed Output not listed

No provider description is available for this model yet.

dashscope/qwen3.5-plus 991.808K context Input not listed Output not listed

No provider description is available for this model yet.

dashscope/qwen3.7-plus 991.808K context Input not listed Output not listed

No provider description is available for this model yet.

github_copilot/gpt-5.1 128K context Input not listed Output not listed

No provider description is available for this model yet.

fireworks_ai/kimi-k2p6 262.144K context $0.95/M input $4/M output

No provider description is available for this model yet.

fireworks_ai/kimi-k2p6-fast 262.144K context $2/M input $8/M output

No provider description is available for this model yet.

fireworks_ai/kimi-k2p7-code 262.144K context $0.95/M input $4/M output