Quality-first multi-agent model for hard research, analysis, and competitions
Models
Every model in the catalog with source-linked pricing, context limits, provider availability, and published benchmark results.
Google's proven reasoning model for coding, math, and multimodal analysis
Coding-optimized GPT model for repository edits, reviews, and agentic software work
Newer StepFun flash model for faster agents, coding, and multimodal prompts
Fast Gemini workhorse for multimodal apps where latency and price matter
Mistral's largest general model for enterprise agents, coding, and multilingual reasoning
Compact GPT model for low-latency assistance and high-volume workloads
Small omni GPT for cheap multimodal assistance and production-scale traffic
Deeper Sonar search model with broader retrieval and stronger synthesis
Lean Gemini 2.5 lane for cheap multimodal traffic and quick agents
| Model | Creator | Score | Inputs | Context | Input | Output | Released | Compare |
|---|---|---|---|---|---|---|---|---|
| Fugu Ultrasakana/fugu-ultra | 58.7 | 1M | $5 | $30 | 2026-06-15 | |||
| Gemini 2.5 Progoogle/gemini-2.5-pro | 42.8 | 1.04858M | $1.25 | $10 | 2025-06-17 | |||
| GPT-5-Codexopenai/gpt-5-codex | 40.9 | 400K | $1.1 | $9 | 2025-09-15 | |||
| Step 3.7 Flashstepfun/step-3.7-flash | 40.0 | 256K | $0.185 | $1.11 | 2026-05-29 | |||
| Gemini 2.5 Flashgoogle/gemini-2.5-flash | 39.4 | 1.04858M | $0.3 | $2.5 | 2025-06-17 | |||
| Mistral Large 3mistral/mistral-large-2512 | 36.2 | 262.144K | $0.5 | $1.5 | 2024-11-01 | |||
| GPT-4 Turboopenai/gpt-4-turbo | 31.9 | 128K | $10 | $30 | 2023-11-06 | |||
| GPT-4o miniopenai/gpt-4o-mini | 22.9 | 128K | $0.15 | $0.6 | 2024-07-18 | |||
| Sonar Properplexity/sonar-pro | 22.6 | 200K | $3 | $15 | 2024-01-01 | |||
| Gemini 2.5 Flash-Litegoogle/gemini-2.5-flash-lite | 19.3 | 1.04858M | $0.1 | $0.4 | 2025-06-17 |