Quality-first multi-agent model for hard research, analysis, and competitions
11 providers
Every model in the catalog with source-linked pricing, context limits, provider availability, and published benchmark results.
Quality-first multi-agent model for hard research, analysis, and competitions
Qwen frontier model tuned for agent frameworks, coding assistants, and long tasks
Google's proven reasoning model for coding, math, and multimodal analysis
Coding-optimized GPT model for repository edits, reviews, and agentic software work
Newer StepFun flash model for faster agents, coding, and multimodal prompts
Fast Gemini workhorse for multimodal apps where latency and price matter
Fast Mistral production model for chat, extraction, and cost-sensitive agents
| Model | Creator | Score | Inputs | Context | Input | Output | Released | Compare |
|---|---|---|---|---|---|---|---|---|
| Fugu Ultrasakana/fugu-ultra | 58.7 | 1M | $5 | $30 | 2026-06-15 | |||
| Qwen3.7 Maxalibaba/qwen3.7-max | 53.5 | 1M | $2.5 | $7.5 | 2026-05-21 | |||
| Gemini 2.5 Progoogle/gemini-2.5-pro | 42.8 | 1.04858M | $1.25 | $10 | 2025-06-17 | |||
| GPT-5-Codexopenai/gpt-5-codex | 40.9 | 400K | $1.1 | $9 | 2025-09-15 | |||
| Step 3.7 Flashstepfun/step-3.7-flash | 40.0 | 256K | $0.185 | $1.11 | 2026-05-29 | |||
| Gemini 2.5 Flashgoogle/gemini-2.5-flash | 39.4 | 1.04858M | $0.3 | $2.5 | 2025-06-17 | |||
| Mistral Small 4mistral/mistral-small-2603 | 38.0 | 256K | $0.15 | $0.6 | 2026-03-16 |