16 models
Ranked by SciCode
Reasoning Tools JSON 60.1

Multi-agent model for routing expert agents across complex analytical tasks

sakana/fugu 2026-06-15 1M context Input not listed Output not listed
1 provider
Reasoning Tools JSON 58.7

Fugu Ultra is the higher-performance model in Sakana AI's Fugu family. Rather than a single monolithic model, Fugu is a learned multi-agent orchestration system: a language model trained to route...

sakana/fugu-ultra 2026-06-15 1M context $5/M input $30/M output
11 providers
Reasoning Tools 53.5

Qwen frontier model tuned for agent frameworks, coding assistants, and long tasks

alibaba/qwen3.7-max 2026-05-21 1M context $2.5/M input $7.5/M output
31 providers
Reasoning Tools JSON 42.8

Gemini 2.5 Pro is Google’s state-of-the-art AI model designed for advanced reasoning, coding, mathematics, and scientific tasks. It employs “thinking” capabilities, enabling it to reason through responses with enhanced accuracy...

google/gemini-2.5-pro 2025-06-17 1.04858M context $1.25/M input $10/M output
29 providers
Reasoning Tools JSON 40.9

Coding-optimized GPT model for repository edits, reviews, and agentic software work

openai/gpt-5-codex 2025-09-15 400K context $1.1/M input $9/M output
13 providers
Reasoning Tools JSON 39.4

Gemini 2.5 Flash is Google's state-of-the-art workhorse model, specifically designed for advanced reasoning, coding, mathematics, and scientific tasks. It includes built-in "thinking" capabilities, enabling it to provide responses with greater...

google/gemini-2.5-flash 2025-06-17 1.04858M context $0.3/M input $2.5/M output
30 providers
38.3

Flagship Qwen3 model for coding agents, complex reasoning, and tool use

alibaba/qwen3-max 2025-09-23 262.144K context $1.2/M input $6/M output
20 providers
Tools 33.3

The 2024-11-20 version of GPT-4o offers a leveled-up creative writing ability with more natural, engaging, and tailored writing to improve relevance & readability. It’s also better at working with uploaded...

openai/gpt-4o-2024-11-20 2024-11-20 128K context $2.5/M input $10/M output
8 providers
Tools 33.1

The 2024-08-06 version of GPT-4o offers improved performance in structured outputs, with the ability to supply a JSON schema in the respone_format. Read more [here](https://openai.com/index/introducing-structured-outputs-in-the-api/). GPT-4o ("o" for "omni") is...

openai/gpt-4o-2024-08-06 2024-08-06 128K context $2.5/M input $10/M output
7 providers
33.1

Mistral model for multilingual chat, reasoning, and tool-assisted workflows

mistral/mistral-medium-2505 2025-05-07 131.072K context $0.4/M input $2/M output
10 providers
31.9

The latest GPT-4 Turbo model with vision capabilities. Vision requests can now use JSON mode and function calling. Training data: up to December 2023.

openai/gpt-4-turbo 2023-11-06 128K context $10/M input $30/M output
15 providers
Tools JSON 30.9

GPT-4o ("o" for "omni") is OpenAI's latest AI model, supporting both text and image inputs with text outputs. It maintains the intelligence level of [GPT-4 Turbo](/models/openai/gpt-4-turbo) while being twice as...

openai/gpt-4o-2024-05-13 2024-05-13 128K context $5/M input $15/M output
5 providers
JSON 22.9

Sonar is lightweight, affordable, fast, and simple to use — now featuring citations and the ability to customize sources. It is designed for companies seeking to integrate lightweight question-and-answer features...

perplexity/sonar 2024-01-01 127.072K context $1/M input $1/M output
9 providers
22.9

GPT-4o mini is OpenAI's newest model after [GPT-4 Omni](/models/openai/gpt-4o), supporting both text and image inputs with text outputs. As their most advanced small model, it is many multiples more affordable...

openai/gpt-4o-mini 2024-07-18 128K context $0.15/M input $0.6/M output
23 providers
JSON 22.6

Note: Sonar Pro pricing includes Perplexity search pricing. See [details here](https://docs.perplexity.ai/guides/pricing#detailed-pricing-breakdown-for-sonar-reasoning-pro-and-sonar-pro) For enterprises seeking more advanced capabilities, the Sonar Pro API can handle in-depth, multi-step queries with added extensibility, like...

perplexity/sonar-pro 2024-01-01 200K context $3/M input $15/M output
9 providers
Reasoning Tools JSON 19.3

Gemini 2.5 Flash-Lite is a lightweight reasoning model in the Gemini 2.5 family, optimized for ultra-low latency and cost efficiency. It offers improved throughput, faster token generation, and better performance...

google/gemini-2.5-flash-lite 2025-06-17 1.04858M context $0.1/M input $0.4/M output
18 providers