Multimodal Kimi model with 1M context and toggleable max-effort thinking for long-horizon agent work
Model catalog
Models
Every model in the catalog with source-linked pricing, context limits, provider availability, and published benchmark results.
63 providers
Multimodal MoE reasoning model (975B total, 41B active) for text, image, and audio
21 providers
Open flagship GLM for long-horizon coding agents and million-token context work
72 providers
Qwen vision-language model for visual reasoning, documents, and agent tasks
16 providers
Qwen instruction model for multilingual chat, reasoning, and tool use
13 providers
Open GPT reasoning model for self-hosted agents and controllable deployments
53 providers
Open GPT reasoning model for self-hosted agents and controllable deployments
32 providers
Open multimodal Llama for strong reasoning with efficient everyday serving
6 providers
Popular open Llama workhorse for multilingual chat, coding, and self-hosting
25 providers
| Model | Creator | Inputs | Context | Input | Output | Released | Compare |
|---|---|---|---|---|---|---|---|
| Kimi K3moonshotai/kimi-k3 | 1.04858M | $3 | $15 | 2026-07-16 | |||
| Inklingthinkingmachines/inkling | 1.04858M | $1.87 | $4.68 | 2026-07-15 | |||
| GLM-5.2zhipuai/glm-5.2 | 1M | $1.4 | $4.4 | 2026-06-13 | |||
| Qwen3.5 122B-A10Balibaba/qwen3.5-122b-a10b | 262.144K | $0.4 | $3.2 | 2026-02-23 | |||
| Qwen3-Next 80B-A3B Instructalibaba/qwen3-next-80b-a3b-instruct | 131.072K | $0.5 | $2 | 2025-09 | |||
| GPT OSS 120Bopenai/gpt-oss-120b | 131.072K | $0.03 | $0.17 | 2025-08-05 | |||
| GPT OSS 20Bopenai/gpt-oss-20b | 131.072K | $0.02 | $0.1 | 2025-08-05 | |||
| Llama 4 Maverick 17B Instructmeta/llama-4-maverick-17b-instruct | 1M | $0.14 | $0.59 | 2025-04-05 | |||
| Llama-3.3-70B-Instructmeta/llama-3.3-70b-instruct | 128K | $0.1 | $0.32 | 2024-12-06 |