No provider description is available for this model yet.
Models
Every model in the catalog with source-linked pricing, context limits, provider availability, and published benchmark results.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
Inkling Small is an open-weight multimodal mixture-of-experts model from Thinking Machines Lab, with 12B active parameters out of 276B total. It is positioned as the smaller, more efficient member of...
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
| Model | Creator | Inputs | Context | Input | Output | Released | Compare |
|---|---|---|---|---|---|---|---|
| Sao10K/L3.3-70B-Euryale-v2.3deepinfra/sao10k/l3.3-70b-euryale-v2.3 | 131.072K | $0.65 | $0.75 | — | |||
| qwen3-32b-fp8lambda_ai/qwen3-32b-fp8 | 131.072K | $0.05 | $0.1 | — | |||
| qwen25-coder-32b-instructlambda_ai/qwen25-coder-32b-instruct | 131.072K | $0.05 | $0.1 | — | |||
| Sao10K/L3.1-70B-Euryale-v2.2deepinfra/sao10k/l3.1-70b-euryale-v2.2 | 131.072K | $0.65 | $0.75 | — | |||
| us/o3-mini-2025-01-31azure/us/o3-mini-2025-01-31 | 200K | $1.21 | $4.84 | — | |||
| us/gpt-5.6-terraazure/us/gpt-5.6-terra | 1.05M | $2.75 | $16.5 | — | |||
| moonshotai/Kimi-K3together_ai/moonshotai/kimi-k3 | 1.04858M | $3 | $15 | — | |||
| gpt-5azure/gpt-5 | 272K | $1.25 | $10 | — | |||
| llama3.3-70b-instruct-fp8lambda_ai/llama3.3-70b-instruct-fp8 | 131.072K | $0.12 | $0.3 | — | |||
| llama3.2-3b-instructlambda_ai/llama3.2-3b-instruct | 131.072K | $0.015 | $0.025 | — | |||
| Qwen/Qwen3-Next-80B-A3B-Thinkingdeepinfra/qwen/qwen3-next-80b-a3b-thinking | 262.144K | $0.14 | $1.4 | — | |||
| llama3.2-11b-vision-instructlambda_ai/llama3.2-11b-vision-instruct | 131.072K | $0.015 | $0.025 | — | |||
| llama3.1-nemotron-70b-instruct-fp8lambda_ai/llama3.1-nemotron-70b-instruct-fp8 | 131.072K | $0.12 | $0.3 | — | |||
| Qwen/Qwen3-Next-80B-A3B-Instructdeepinfra/qwen/qwen3-next-80b-a3b-instruct | 262.144K | $0.14 | $1.4 | — | |||
| us/o3-2025-04-16azure/us/o3-2025-04-16 | 200K | $2.2 | $8.8 | — | |||
| llama3.1-8b-instructlambda_ai/llama3.1-8b-instruct | 131.072K | $0.025 | $0.04 | — | |||
| llama3.1-70b-instruct-fp8lambda_ai/llama3.1-70b-instruct-fp8 | 131.072K | $0.12 | $0.3 | — | |||
| Qwen/Qwen3-Coder-480B-A35B-Instruct-Turbodeepinfra/qwen/qwen3-coder-480b-a35b-instruct-turbo | 262.144K | $0.29 | $1.2 | — | |||
| llama3.1-405b-instruct-fp8lambda_ai/llama3.1-405b-instruct-fp8 | 131.072K | $0.8 | $0.8 | — | |||
| llama-4-maverick-17b-128e-instruct-fp8lambda_ai/llama-4-maverick-17b-128e-instruct-fp8 | 131.072K | $0.05 | $0.1 | — | |||
| Qwen/Qwen3-Coder-480B-A35B-Instructdeepinfra/qwen/qwen3-coder-480b-a35b-instruct | 262.144K | $0.4 | $1.6 | — | |||
| Thinking Machines: Inkling Small (batch)thinkingmachines/inkling-small:batch | 524.288K | $0.5 | $1.2 | — | |||
| us/gpt-5.6-solazure/us/gpt-5.6-sol | 1.05M | $5.5 | $33 | — | |||
| lfm-7blambda_ai/lfm-7b | 131.072K | $0.025 | $0.04 | — | |||
| lfm-40blambda_ai/lfm-40b | 131.072K | $0.1 | $0.2 | — | |||
| Qwen/Qwen3-235B-A22B-Thinking-2507deepinfra/qwen/qwen3-235b-a22b-thinking-2507 | 262.144K | $0.3 | $2.9 | — | |||
| hermes3-8blambda_ai/hermes3-8b | 131.072K | $0.025 | $0.04 | — | |||
| hermes3-70blambda_ai/hermes3-70b | 131.072K | $0.12 | $0.3 | — | |||
| Qwen/Qwen3-235B-A22B-Instruct-2507deepinfra/qwen/qwen3-235b-a22b-instruct-2507 | 262.144K | $0.09 | $0.6 | — | |||
| qwen-3-32bcerebras/qwen-3-32b | 128K | $0.4 | $0.8 | — | |||
| accounts/fireworks/models/nvidia-nemotron-nano-9b-v2fireworks_ai/accounts/fireworks/models/nvidia-nemotron-nano-9b-v2 | 131.072K | $0.2 | $0.2 | — | |||
| hermes3-405blambda_ai/hermes3-405b | 131.072K | $0.8 | $0.8 | — | |||
| accounts/fireworks/models/nvidia-nemotron-nano-12b-v2fireworks_ai/accounts/fireworks/models/nvidia-nemotron-nano-12b-v2 | 131.072K | $0.2 | $0.2 | — | |||
| accounts/fireworks/models/mistral-nemo-instruct-2407fireworks_ai/accounts/fireworks/models/mistral-nemo-instruct-2407 | 128K | $0.2 | $0.2 | — | |||
| deepseek-v3-0324lambda_ai/deepseek-v3-0324 | 131.072K | $0.2 | $0.6 | — | |||
| Qwen/Qwen2.5-VL-32B-Instructdeepinfra/qwen/qwen2.5-vl-32b-instruct | 128K | $0.2 | $0.6 | — | |||
| accounts/fireworks/models/mistral-nemo-base-2407fireworks_ai/accounts/fireworks/models/mistral-nemo-base-2407 | 128K | $0.2 | $0.2 | — | |||
| accounts/fireworks/models/mistral-large-3-fp8fireworks_ai/accounts/fireworks/models/mistral-large-3-fp8 | 256K | $1.2 | $1.2 | — | |||
| deepseek-r1-671blambda_ai/deepseek-r1-671b | 131.072K | $0.8 | $0.8 | — | |||
| accounts/fireworks/models/phi-3-mini-128k-instructfireworks_ai/accounts/fireworks/models/phi-3-mini-128k-instruct | 131.072K | $0.1 | $0.1 | — | |||
| accounts/fireworks/models/qwen-v2p5-7bfireworks_ai/accounts/fireworks/models/qwen-v2p5-7b | 131.072K | $0.2 | $0.2 | — | |||
| accounts/fireworks/models/qwen2p5-14bfireworks_ai/accounts/fireworks/models/qwen2p5-14b | 131.072K | $0.2 | $0.2 | — | |||
| accounts/fireworks/models/qwen2p5-32bfireworks_ai/accounts/fireworks/models/qwen2p5-32b | 131.072K | $0.9 | $0.9 | — | |||
| accounts/fireworks/models/qwen2p5-72bfireworks_ai/accounts/fireworks/models/qwen2p5-72b | 131.072K | $0.9 | $0.9 | — | |||
| accounts/fireworks/models/qwen2p5-coder-32b-instruct-128kfireworks_ai/accounts/fireworks/models/qwen2p5-coder-32b-instruct-128k | 131.072K | $0.9 | $0.9 | — | |||
| accounts/fireworks/models/qwen2p5-vl-32b-instructfireworks_ai/accounts/fireworks/models/qwen2p5-vl-32b-instruct | 128K | $0.9 | $0.9 | — | |||
| accounts/fireworks/models/qwen2p5-vl-3b-instructfireworks_ai/accounts/fireworks/models/qwen2p5-vl-3b-instruct | 128K | $0.2 | $0.2 | — | |||
| accounts/fireworks/models/qwen2p5-vl-72b-instructfireworks_ai/accounts/fireworks/models/qwen2p5-vl-72b-instruct | 128K | $0.9 | $0.9 | — | |||
| accounts/fireworks/models/qwen2p5-vl-7b-instructfireworks_ai/accounts/fireworks/models/qwen2p5-vl-7b-instruct | 128K | $0.2 | $0.2 | — | |||
| accounts/fireworks/models/ministral-3-8b-instruct-2512fireworks_ai/accounts/fireworks/models/ministral-3-8b-instruct-2512 | 256K | $0.2 | $0.2 | — |