Quality-first multi-agent model for hard research, analysis, and competitions
Models
Every model in the catalog with source-linked pricing, context limits, provider availability, and published benchmark results.
Qwen frontier model tuned for agent frameworks, coding assistants, and long tasks
Google's proven reasoning model for coding, math, and multimodal analysis
Fast Gemini workhorse for multimodal apps where latency and price matter
Late GLM-4 workhorse for coding agents, reasoning, and structured tasks
Flagship Qwen3 model for coding agents, complex reasoning, and tool use
Mistral's largest general model for enterprise agents, coding, and multilingual reasoning
Classic open reasoning model for transparent math, coding, and deliberate problem solving
Hybrid-reasoning GLM release that made the 4.5 line broadly useful
GPT model for general reasoning, writing, coding, and tool-assisted tasks
GPT model for general reasoning, writing, coding, and tool-assisted tasks
Mistral model for multilingual chat, reasoning, and tool-assisted workflows
Mistral's coding-agent model for repository work, terminal tasks, and software fixes
Compact GPT model for low-latency assistance and high-volume workloads
GPT model for general reasoning, writing, coding, and tool-assisted tasks
Lighter GLM-4.5 variant for fast coding assistance and cheaper agents
Flagship Mistral model for advanced reasoning, coding, and multilingual work
Popular open Llama workhorse for multilingual chat, coding, and self-hosting
Small omni GPT for cheap multimodal assistance and production-scale traffic
Lean Gemini 2.5 lane for cheap multimodal traffic and quick agents
| Model | Creator | Score | Inputs | Context | Input | Output | Released | Compare |
|---|---|---|---|---|---|---|---|---|
| Fugu Ultrasakana/fugu-ultra | 58.7 | 1M | $5 | $30 | 2026-06-15 | |||
| Qwen3.7 Maxalibaba/qwen3.7-max | 53.5 | 1M | $2.5 | $7.5 | 2026-05-21 | |||
| Gemini 2.5 Progoogle/gemini-2.5-pro | 42.8 | 1.04858M | $1.25 | $10 | 2025-06-17 | |||
| Gemini 2.5 Flashgoogle/gemini-2.5-flash | 39.4 | 1.04858M | $0.3 | $2.5 | 2025-06-17 | |||
| GLM-4.6zhipuai/glm-4.6 | 38.4 | 204.8K | $0.6 | $2.2 | 2025-09-30 | |||
| Qwen3 Maxalibaba/qwen3-max | 38.3 | 262.144K | $1.2 | $6 | 2025-09-23 | |||
| Mistral Large 3mistral/mistral-large-2512 | 36.2 | 262.144K | $0.5 | $1.5 | 2024-11-01 | |||
| DeepSeek-R1deepseek/deepseek-r1 | 35.7 | 128K | $0.7 | $2.5 | 2025-01-20 | |||
| GLM-4.5zhipuai/glm-4.5 | 34.8 | 131.072K | $0.6 | $2.2 | 2025-07-28 | |||
| GPT-4o (2024-11-20)openai/gpt-4o-2024-11-20 | 33.3 | 128K | $2.5 | $10 | 2024-11-20 | |||
| GPT-4o (2024-08-06)openai/gpt-4o-2024-08-06 | 33.1 | 128K | $2.5 | $10 | 2024-08-06 | |||
| Mistral Medium 3mistral/mistral-medium-2505 | 33.1 | 131.072K | $0.4 | $2 | 2025-05-07 | |||
| Devstral 2mistral/devstral-2512 | 33.1 | 262.144K | $0.4 | $2 | 2025-12-09 | |||
| GPT-4 Turboopenai/gpt-4-turbo | 31.9 | 128K | $10 | $30 | 2023-11-06 | |||
| GPT-4o (2024-05-13)openai/gpt-4o-2024-05-13 | 30.9 | 128K | $5 | $15 | 2024-05-13 | |||
| GLM-4.5-Airzhipuai/glm-4.5-air | 30.6 | 131.072K | $0.2 | $1.1 | 2025-07-28 | |||
| Mistral Large 2.1mistral/mistral-large-2411 | 29.2 | 131.072K | $2 | $6 | 2024-11-18 | |||
| Llama-3.3-70B-Instructmeta/llama-3.3-70b-instruct | 26.0 | 128K | $0.1 | $0.32 | 2024-12-06 | |||
| GPT-4o miniopenai/gpt-4o-mini | 22.9 | 128K | $0.15 | $0.6 | 2024-07-18 | |||
| Gemini 2.5 Flash-Litegoogle/gemini-2.5-flash-lite | 19.3 | 1.04858M | $0.1 | $0.4 | 2025-06-17 |