Kimi K3 is a 2.8T parameter open-weight multimodal reasoning model from Moonshot AI. It is suited for complex coding, knowledge work, and long-horizon agentic workflows, and is particularly strong at...
Models
Every model in the catalog with source-linked pricing, context limits, provider availability, and published benchmark results.
Claude Opus 5 is Anthropic’s flagship model for demanding reasoning, coding, and long-horizon agentic work. It is particularly strong at end-to-end software tasks, code review and bug finding, visual analysis...
Kimi K2.6 is Moonshot AI's next-generation multimodal model, designed for long-horizon coding, coding-driven UI/UX generation, and multi-agent orchestration. It handles complex end-to-end coding tasks across Python, Rust, and Go, and...
Sonnet 5 is Anthropic's most capable Sonnet-class model, with frontier performance across coding, agents, and professional work. It supports adaptive thinking with selectable reasoning effort levels (low, medium, high, max,...
Gemini 3.7 Flash is a multimodal model from Google for fast agentic workflows, coding, and complex multi-step reasoning. It is designed for tasks that require responsive performance and reliable multi-step...
Gemini 3.5 Flash is Google's high-efficiency multimodal model, bringing near-Pro level coding and reasoning at Flash-tier cost and speed. It is highly optimized for coding proficiency and parallel agentic execution...
Gemini 3.6 Flash is a high-efficiency model from Google for coding, agentic workflows, and web and app development. It is designed to produce polished outputs with fewer unnecessary edits and...
MoonshotAI: Kimi K2.7 Code is a coding-focused model in Moonshot AI's Kimi K2 family, built to complete end-to-end programming tasks reliably over long contexts. It uses a native multimodal mixture-of-experts...
Kimi K2.5 is Moonshot AI's native multimodal model, delivering state-of-the-art visual coding capability and a self-directed agent swarm paradigm. Built on Kimi K2 with continued pretraining over approximately 15T mixed...
GPT-5.4 is OpenAI’s latest frontier model, unifying the Codex and GPT lines into a single system. It features a 1M+ token context window (922K input, 128K output) with support for...
GPT-5.1 is the latest frontier-grade model in the GPT-5 series, offering stronger general-purpose reasoning, improved instruction adherence, and a more natural conversational style compared to GPT-5. It uses adaptive reasoning...
DeepSeek V4 Pro is a large-scale Mixture-of-Experts model from DeepSeek with 1.6T total parameters and 49B activated parameters, supporting a 1M-token context window. It is designed for advanced reasoning, coding,...
| Model | Creator | Score | Inputs | Context | Input | Output | Released | Compare |
|---|---|---|---|---|---|---|---|---|
| MoonshotAI: Kimi K3moonshotai/kimi-k3 | 1317.0 | 1.04858M | $2.34 | $11.7 | 2026-07-16 | |||
| Claude Opus 5anthropic/claude-opus-5 | 1277.0 | 1M | $5 | $25 | 2026-07-24 | |||
| MoonshotAI: Kimi K2.6moonshotai/kimi-k2.6 | 1268.0 | 262.144K | $0.95 | $4 | 2026-04-21 | |||
| Anthropic: Claude Sonnet 5anthropic/claude-sonnet-5 | 1265.0 | 1M | $2 | $10 | 2026-06-30 | |||
| Google: Gemini 3.7 Flashgoogle/gemini-3.7-flash | 1243.0 | 1.04858M | $0.75 | $3.75 | 2026-08-13 | |||
| Google: Gemini 3.5 Flashgoogle/gemini-3.5-flash | 1223.0 | 1.04858M | $1.5 | $9 | 2026-05-19 | |||
| Google: Gemini 3.6 Flashgoogle/gemini-3.6-flash | 1217.0 | 1.04858M | $0.75 | $3.75 | 2026-07-21 | |||
| MoonshotAI: Kimi K2.7 Codemoonshotai/kimi-k2.7-code | 1195.0 | 262.144K | $0.71 | $3.5 | 2026-06-12 | |||
| MoonshotAI: Kimi K2.5moonshotai/kimi-k2.5 | 1156.0 | 262.144K | $0.45 | $2.25 | 2026-01 | |||
| OpenAI: GPT-5.4openai/gpt-5.4 | 1079.0 | 1.05M | $2.5 | $15 | 2026-03-05 | |||
| OpenAI: GPT-5.1openai/gpt-5.1 | 1029.0 | 400K | $1.25 | $10 | 2025-11-13 | |||
| DeepSeek: DeepSeek V4 Pro 0423deepseek/deepseek-v4-pro | 1000.0 | 1.024M | $0.86 | $1.72 | 2026-04-24 |