Quality-first multi-agent model for hard research, analysis, and competitions
11 providers
Every model in the catalog with source-linked pricing, context limits, provider availability, and published benchmark results.
Quality-first multi-agent model for hard research, analysis, and competitions
Qwen frontier model tuned for agent frameworks, coding assistants, and long tasks
Google's proven reasoning model for coding, math, and multimodal analysis
Coding-optimized GPT model for repository edits, reviews, and agentic software work
StepFun flash lane for quick multimodal reasoning and coding assistance
Fast Gemini workhorse for multimodal apps where latency and price matter
StepFun flash model for efficient multimodal reasoning, coding, and tool use
Flagship Qwen3 model for coding agents, complex reasoning, and tool use
| Model | Creator | Score | Inputs | Context | Input | Output | Released | Compare |
|---|---|---|---|---|---|---|---|---|
| Fugu Ultrasakana/fugu-ultra | 58.7 | 1M | $5 | $30 | 2026-06-15 | |||
| Qwen3.7 Maxalibaba/qwen3.7-max | 53.5 | 1M | $2.5 | $7.5 | 2026-05-21 | |||
| Gemini 2.5 Progoogle/gemini-2.5-pro | 42.8 | 1.04858M | $1.25 | $10 | 2025-06-17 | |||
| GPT-5-Codexopenai/gpt-5-codex | 40.9 | 400K | $1.1 | $9 | 2025-09-15 | |||
| Step 3.5 Flashstepfun/step-3.5-flash | 40.4 | 256K | $0.1 | $0.3 | 2026-01-29 | |||
| Gemini 2.5 Flashgoogle/gemini-2.5-flash | 39.4 | 1.04858M | $0.3 | $2.5 | 2025-06-17 | |||
| Step 3.5 Flash 2603stepfun/step-3.5-flash-2603 | 38.5 | 256K | $0.1 | $0.3 | 2026-04-02 | |||
| Qwen3 Maxalibaba/qwen3-max | 38.3 | 262.144K | $1.2 | $6 | 2025-09-23 |