Inkling is an open-weight multimodal mixture-of-experts model from Thinking Machines Lab, with 41B active parameters out of 975B total. It is designed for general-purpose reasoning, coding, agentic and tool-use systems,...
Models
Every model in the catalog with source-linked pricing, context limits, provider availability, and published benchmark results.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
GPT-5.4 mini brings the core capabilities of GPT-5.4 to a faster, more efficient model optimized for high-throughput workloads. It supports text and image inputs with strong performance across reasoning, coding,...
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
GPT-5.4 Pro is OpenAI's most advanced model, building on GPT-5.4's unified architecture with enhanced reasoning capabilities for complex, high-stakes tasks. It features a 1M+ token context window (922K input, 128K...
No provider description is available for this model yet.
No provider description is available for this model yet.
GPT-5-Codex is a specialized version of GPT-5 optimized for software engineering and coding workflows. It is designed for both interactive development sessions and long, independent execution of complex engineering tasks....
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
GPT-5.1 Chat (AKA Instant is the fast, lightweight member of the 5.1 family, optimized for low-latency chat while retaining strong general intelligence. It uses adaptive reasoning to selectively “think” on...
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
Grok 4.3 is a reasoning model from SpaceXAI. It accepts text and image inputs with text output, and is suited for agentic workflows, instruction-following tasks, and applications requiring high factual...
No provider description is available for this model yet.
No provider description is available for this model yet.
| Model | Creator | Inputs | Context | Input | Output | Released | Compare |
|---|---|---|---|---|---|---|---|
| Thinking Machines: Inkling (free)thinkingmachines/inkling:free | 1.04858M | Free | Free | — | |||
| qwen/qwen3.5-397b-a17bnovita/qwen/qwen3.5-397b-a17b | 262.144K | $0.6 | $3.6 | — | |||
| deepseek/deepseek-ocr-2novita/deepseek/deepseek-ocr-2 | 8.192K | $0.03 | $0.03 | — | |||
| moonshotai/kimi-k2.5novita/moonshotai/kimi-k2.5 | 262.144K | $0.6 | $3 | — | |||
| qwen/qwen3.6-35b-a3bnovita/qwen/qwen3.6-35b-a3b | 262.144K | $0.248 | $1.485 | — | |||
| google/gemma-4-31B-itwandb/google/gemma-4-31b-it | 262.144K | $0.1 | $0.34 | — | |||
| MiniMaxAI/MiniMax-M3wandb/minimaxai/minimax-m3 | 262.144K | $0.23 | $0.96 | — | |||
| moonshotai/Kimi-K2.7-Codewandb/moonshotai/kimi-k2.7-code | 262.144K | $0.71 | $3.5 | — | |||
| moonshotai/Kimi-K2.6wandb/moonshotai/kimi-k2.6 | 262.144K | $0.65 | $3.41 | — | |||
| OpenAI: GPT-5.4 Mini (batch)openai/gpt-5.4-mini:batch | 400K | $0.375 | $2.25 | — | |||
| Qwen/Qwen3.8-27Bwandb/qwen/qwen3.8-27b | 262.144K | $0.4 | $3 | — | |||
| Qwen/Qwen3.6-35B-A3Bwandb/qwen/qwen3.6-35b-a3b | 262.144K | $0.25 | $1.25 | — | |||
| Qwen/Qwen3.6-27Bwandb/qwen/qwen3.6-27b | 262.144K | $0.6 | $3.6 | — | |||
| OpenAI: GPT-5.4 Pro (batch)openai/gpt-5.4-pro:batch | 1.05M | $15 | $90 | — | |||
| Qwen/Qwen3.5-35B-A3Bwandb/qwen/qwen3.5-35b-a3b | 262.144K | $0.25 | $1.25 | — | |||
| Qwen/Qwen3.8-27Bdeepinfra/qwen/qwen3.8-27b | 262.144K | $0.4 | $3 | — | |||
| OpenAI: GPT-5 Codex (batch)openai/gpt-5-codex:batch | 400K | $0.625 | $5 | — | |||
| google/gemma-4-31B-it-Ultradeepinfra/google/gemma-4-31b-it-ultra | 131.072K | $0.27 | $0.76 | — | |||
| moonshotai/Kimi-K2.5deepinfra/moonshotai/kimi-k2.5 | 262.144K | $0.45 | $2.25 | — | |||
| anthropic/claude-opus-4-8deepinfra/anthropic/claude-opus-4-8 | 1M | $5 | $25 | — | |||
| anthropic/claude-sonnet-4-6deepinfra/anthropic/claude-sonnet-4-6 | 1M | $3 | $15 | — | |||
| google/gemini-3.5-flashdeepinfra/google/gemini-3.5-flash | 1M | $1.5 | $9 | — | |||
| OpenAI: GPT-5.1 Chatopenai/gpt-5.1-chat | 128K | $1.25 | $10 | — | |||
| XiaomiMiMo/MiMo-V2.5deepinfra/xiaomimimo/mimo-v2.5 | 262.144K | $0.4 | $2 | — | |||
| google/gemma-4-31B-it-turbodeepinfra/google/gemma-4-31b-it-turbo | 262.144K | $0.09 | $0.34 | — | |||
| databricks-glm-5-3-flashdatabricks/databricks-glm-5-3-flash | 1.04858M | — | — | — | |||
| gemma-4-26b-a4b-itgemini/gemma-4-26b-a4b-it | 262.144K | — | — | — | |||
| gemma-4-31b-itgemini/gemma-4-31b-it | 262.144K | — | — | — | |||
| kimi-k2.7-codemoonshot/kimi-k2.7-code | 262.144K | $0.95 | $4 | — | |||
| thinkingmachines/Inkling-Smalldeepinfra/thinkingmachines/inkling-small | 524.288K | $0.45 | $1.2 | — | |||
| meta-models/Muse-Glimmer-30Bdeepinfra/meta-models/muse-glimmer-30b | 131.072K | $0.3 | $1.2 | — | |||
| glm-5.3-flashzai/glm-5.3-flash | 1.04858M | $0.15 | $0.5 | — | |||
| gemini-omni-1.1-flashgemini/gemini-omni-1.1-flash | 131.072K | $1.5 | $9 | — | |||
| grok-4.20xai/grok-4.20 | 1M | $1.25 | $2.5 | — | |||
| grok-4.20-reasoningxai/grok-4.20-reasoning | 1M | $1.25 | $2.5 | — | |||
| Qwen/Qwen3-VL-235B-A22B-Instructdeepinfra/qwen/qwen3-vl-235b-a22b-instruct | 262.144K | $0.2 | $0.88 | — | |||
| Qwen/Qwen3-VL-30B-A3B-Instructdeepinfra/qwen/qwen3-vl-30b-a3b-instruct | 262.144K | $0.15 | $0.6 | — | |||
| grok-4.20-reasoning-latestxai/grok-4.20-reasoning-latest | 1M | $1.25 | $2.5 | — | |||
| grok-4.20-non-reasoningxai/grok-4.20-non-reasoning | 1M | $1.25 | $2.5 | — | |||
| Qwen/Qwen3.5-27Bdeepinfra/qwen/qwen3.5-27b | 262.144K | $0.26 | $2.6 | — | |||
| Qwen/Qwen3.6-35B-A3Bdeepinfra/qwen/qwen3.6-35b-a3b | 262.144K | $0.1 | $0.95 | — | |||
| grok-4.20-non-reasoning-latestxai/grok-4.20-non-reasoning-latest | 1M | $1.25 | $2.5 | — | |||
| nvidia/Nemotron-Content-Safety-3.5deepinfra/nvidia/nemotron-content-safety-3.5 | 131.072K | $0.2 | $0.2 | — | |||
| qwen/qwen3.8-27bgroq/qwen/qwen3.8-27b | 131.042K | $0.8 | $4 | — | |||
| mistral-medium-3.5mistral/mistral-medium-3.5 | 262.144K | $1.5 | $7.5 | — | |||
| mistral-vibe-cli-latestmistral/mistral-vibe-cli-latest | 262.144K | $1.5 | $7.5 | — | |||
| anthropic/claude-opus-5deepinfra/anthropic/claude-opus-5 | 1M | $5 | $25 | — | |||
| SpaceXAI: Grok 4.3x-ai/grok-4.3 | 1M | $1.25 | $2.5 | — | |||
| thinkingmachines/Inklingdeepinfra/thinkingmachines/inkling | 524.288K | $0.95 | $4.05 | — | |||
| moonshotai/Kimi-K2.6deepinfra/moonshotai/kimi-k2.6 | 262.144K | $0.75 | $3.5 | — |