Qwen3-235B-A22B is a 235B parameter mixture-of-experts (MoE) model developed by Qwen, activating 22B parameters per forward pass. It supports seamless switching between a "thinking" mode for complex reasoning, math, and...

qwen/qwen3-235b-a22b 131.072K context $0.455/M input $1.82/M output

No provider description is available for this model yet.

databricks/databricks-glm-5-3-flash 1.04858M context Input not listed Output not listed

No provider description is available for this model yet.

gemini/gemma-4-26b-a4b-it 262.144K context Input not listed Output not listed

No provider description is available for this model yet.

gemini/gemma-4-31b-it 262.144K context Input not listed Output not listed

No provider description is available for this model yet.

moonshot/kimi-k2.7-code 262.144K context $0.95/M input $4/M output

No provider description is available for this model yet.

together_ai/zai-org/glm-5.3 1.04858M context $1.4/M input $4.4/M output

No provider description is available for this model yet.

zai/glm-5.3-flash 1.04858M context $0.15/M input $0.5/M output

No provider description is available for this model yet.

gemini/gemini-omni-1.1-flash 131.072K context $1.5/M input $9/M output

No provider description is available for this model yet.

xai/grok-4.20 1M context $1.25/M input $2.5/M output

No provider description is available for this model yet.

xai/grok-4.20-reasoning 1M context $1.25/M input $2.5/M output

No provider description is available for this model yet.

xai/grok-4.20-reasoning-latest 1M context $1.25/M input $2.5/M output

No provider description is available for this model yet.

xai/grok-4.20-non-reasoning 1M context $1.25/M input $2.5/M output

No provider description is available for this model yet.

deepinfra/qwen/qwen3.5-27b 262.144K context $0.26/M input $2.6/M output

No provider description is available for this model yet.

deepinfra/qwen/qwen3.6-35b-a3b 262.144K context $0.1/M input $0.95/M output

No provider description is available for this model yet.

novita/minimax/minimax-m2 204.8K context $0.3/M input $1.2/M output

GLM-5 is Z.ai’s flagship open-source foundation model engineered for complex systems design and long-horizon agent workflows. Built for expert developers, it delivers production-grade performance on large-scale programming tasks, rivaling leading...

z-ai/glm-5 198K context $0.6/M input $1.92/M output

No provider description is available for this model yet.

groq/qwen/qwen3.8-27b 131.042K context $0.8/M input $4/M output

No provider description is available for this model yet.

mistral/mistral-medium-3.5 262.144K context $1.5/M input $7.5/M output

No provider description is available for this model yet.

mistral/mistral-vibe-cli-latest 262.144K context $1.5/M input $7.5/M output

Grok 4.3 is a reasoning model from SpaceXAI. It accepts text and image inputs with text output, and is suited for agentic workflows, instruction-following tasks, and applications requiring high factual...

x-ai/grok-4.3 1M context $1.25/M input $2.5/M output

No provider description is available for this model yet.

deepinfra/thinkingmachines/inkling 524.288K context $0.95/M input $4.05/M output

No provider description is available for this model yet.

deepinfra/moonshotai/kimi-k2.6 262.144K context $0.75/M input $3.5/M output

No provider description is available for this model yet.

deepinfra/qwen/qwen3.7-max 256K context $2.5/M input $7.5/M output

No provider description is available for this model yet.

deepinfra/bytedance/seed-2.0-mini 256K context $0.1/M input $0.4/M output

No provider description is available for this model yet.

deepinfra/qwen/qwen3.8-2.4t-a95b 262.144K context $2/M input $6/M output

Qwen2.5-VL is proficient in recognizing common objects such as flowers, birds, fish, and insects. It is also highly capable of analyzing texts, charts, icons, graphics, and layouts within images.

qwen/qwen2.5-vl-72b-instruct 128K context $0.8/M input $1/M output

No provider description is available for this model yet.

deepinfra/minimaxai/minimax-m3 524.288K context $0.28/M input $1.1/M output

No provider description is available for this model yet.

deepinfra/google/gemini-3.7-flash 1M context $0.75/M input $3.75/M output

No provider description is available for this model yet.

deepinfra/stepfun-ai/step-3.7-flash 262.144K context $0.2/M input $1.15/M output

No provider description is available for this model yet.

deepinfra/qwen/qwen3.5-35b-a3b 262.144K context $0.14/M input $1/M output

No provider description is available for this model yet.

deepinfra/bytedance/seed-1.8 256K context $0.25/M input $2/M output

OpenAI o4-mini-high is the same model as [o4-mini](/openai/o4-mini) with reasoning_effort set to high. OpenAI o4-mini is a compact reasoning model in the o-series, optimized for fast, cost-efficient performance while retaining...

openai/o4-mini-high:batch 200K context $0.55/M input $2.2/M output

No provider description is available for this model yet.

deepinfra/tencent/hy3 262.144K context $0.14/M input $0.58/M output

No provider description is available for this model yet.

deepinfra/bytedance/seed-2.0-pro 256K context $0.5/M input $3/M output

No provider description is available for this model yet.

deepinfra/zai-org/glm-5 202.752K context $0.6/M input $2.08/M output