Reasoning Tools JSON

GPT-5.1-Codex-Max is OpenAI’s latest agentic coding model, designed for long-running, high-context software development tasks. It is based on an updated version of the 5.1 reasoning stack and trained on agentic...

openai/gpt-5.1-codex-max 2025-11-13 400K context $1.25/M input $10/M output
15 providers
Reasoning

GPT-5.1-Codex is a specialized version of GPT-5.1 optimized for software engineering and coding workflows. It is designed for both interactive development sessions and long, independent execution of complex engineering tasks....

openai/gpt-5.1-codex 2025-11-13 400K context $1.25/M input $10/M output
19 providers
Reasoning

GPT-5.1-Codex-Mini is a smaller and faster version of GPT-5.1-Codex

openai/gpt-5.1-codex-mini 2025-11-13 400K context $0.25/M input $2/M output
15 providers
Tools Open weights

Kimi K2 Thinking is Moonshot AI’s most advanced open reasoning model to date, extending the K2 series into agentic, long-horizon reasoning. Built on the trillion-parameter Mixture-of-Experts (MoE) architecture introduced in...

moonshotai/kimi-k2-thinking 2025-11-06 262.144K context $0.6/M input $2.5/M output
22 providers
Reasoning Tools JSON

Flagship Claude model for deep reasoning, coding, and long-horizon agents

anthropic/claude-opus-4-5-20251101 2025-11-01 200K context $5/M input $25/M output
17 providers
Reasoning Open weights

gpt-oss-safeguard-20b is a safety reasoning model from OpenAI built upon gpt-oss-20b. This open-weight, 21B-parameter Mixture-of-Experts (MoE) model offers lower latency for safety tasks like content classification, LLM filtering, and trust...

openai/gpt-oss-safeguard-20b 2025-10-29 131.072K context $0.075/M input $0.3/M output
7 providers
Reasoning Tools JSON Open weights

Safety model for policy screening, moderation, and risk-aware routing workflows

openai/gpt-oss-safeguard-120b 2025-10-29 131.072K context $0.15/M input $0.6/M output
4 providers
Reasoning Tools Open weights

Nemotron multimodal model for visual reasoning and agentic AI workflows

nvidia/nemotron-nano-12b-v2-vl 2025-10-28 128K context $0.2/M input $0.6/M output
3 providers
Reasoning Tools JSON

ByteDance Seed model for long-context reasoning, instruction following, and tool-assisted tasks

bytedance-seed/seed-1-6 2025-10-15 256K context $0.119/M input $1.187/M output
4 providers
Tools JSON

Fast Claude model for responsive assistance, classification, and lightweight agents

anthropic/claude-haiku-4-5-20251001 2025-10-15 200K context $1/M input $5/M output
20 providers
Reasoning

GPT-5 Pro is OpenAI’s most advanced model, offering major improvements in reasoning, code quality, and user experience. It is optimized for complex tasks that require step-by-step reasoning, instruction following, and...

openai/gpt-5-pro 2025-10-06 400K context $15/M input $120/M output
15 providers
Reasoning Tools JSON

Balanced Claude model for coding, analysis, agent workflows, and cost control

anthropic/claude-sonnet-4-5-20250929 2025-09-29 200K context $3/M input $15/M output
18 providers

Flagship Qwen3 model for coding agents, complex reasoning, and tool use

alibaba/qwen3-max 2025-09-23 262.144K context $1.2/M input $6/M output
20 providers
Reasoning Tools JSON

Coding-optimized GPT model for repository edits, reviews, and agentic software work

openai/gpt-5-codex 2025-09-15 400K context $1.1/M input $9/M output
13 providers
Reasoning Tools Open weights

Efficient Qwen thinking model for local reasoning, math, and coding agents

alibaba/qwen3-next-80b-a3b-thinking 2025-09 131.072K context $0.5/M input $6/M output
10 providers

Gemini 2.5 Flash Image, a.k.a. "Nano Banana," is now generally available. It is a state of the art image generation model with contextual understanding. It is capable of image generation,...

google/gemini-2.5-flash-image 2025-08-26 32.768K context $0.3/M input $2.5/M output
8 providers
Reasoning Tools Open weights

Hybrid-reasoning DeepSeek model with thinking and non-thinking modes

deepseek/deepseek-v3.1 2025-08-21 131.072K context $0.19/M input $0.71/M output
10 providers
Reasoning Tools Open weights

Compact Nemotron model for efficient reasoning and deployable AI agents

nvidia/nemotron-nano-9b-v2 2025-08-18 131.072K context $0.06/M input $0.23/M output
4 providers
Reasoning

GPT-5 Mini is a compact version of GPT-5, designed to handle lighter-weight reasoning tasks. It provides the same instruction-following and safety-tuning benefits as GPT-5, but with reduced latency and cost....

openai/gpt-5-mini 2025-08-07 400K context $0.25/M input $2/M output
29 providers
Reasoning

GPT-5-Nano is the smallest and fastest variant in the GPT-5 system, optimized for developer tools, rapid interactions, and ultra-low latency environments. While limited in reasoning depth compared to its larger...

openai/gpt-5-nano 2025-08-07 400K context $0.05/M input $0.4/M output
26 providers
Tools

GPT-5 is OpenAI’s most advanced model, offering major improvements in reasoning, code quality, and user experience. It is optimized for complex tasks that require step-by-step reasoning, instruction following, and accuracy...

openai/gpt-5 2025-08-07 400K context $1.25/M input $10/M output
30 providers
Reasoning Tools JSON Open weights

gpt-oss-120b is an open-weight, 117B-parameter Mixture-of-Experts (MoE) language model from OpenAI designed for high-reasoning, agentic, and general-purpose production use cases. It activates 5.1B parameters per forward pass and is optimized...

openai/gpt-oss-120b 2025-08-05 131.072K context $0.037/M input $0.17/M output
53 providers
Reasoning Tools JSON

Flagship Claude model for deep reasoning, coding, and long-horizon agents

anthropic/claude-opus-4-1-20250805 2025-08-05 200K context $15/M input $75/M output
13 providers
Reasoning Tools Open weights

gpt-oss-20b is an open-weight 21B parameter model released by OpenAI under the Apache 2.0 license. It uses a Mixture-of-Experts (MoE) architecture with 3.6B active parameters per forward pass, optimized for...

openai/gpt-oss-20b 2025-08-05 131.072K context $0.03/M input $0.13/M output
32 providers
Reasoning Tools JSON

Gemini 2.5 Pro is Google’s state-of-the-art AI model designed for advanced reasoning, coding, mathematics, and scientific tasks. It employs “thinking” capabilities, enabling it to reason through responses with enhanced accuracy...

google/gemini-2.5-pro 2025-06-17 1.04858M context $1.25/M input $10/M output
29 providers
Reasoning Tools JSON

Gemini 2.5 Flash is Google's state-of-the-art workhorse model, specifically designed for advanced reasoning, coding, mathematics, and scientific tasks. It includes built-in "thinking" capabilities, enabling it to provide responses with greater...

google/gemini-2.5-flash 2025-06-17 1.04858M context $0.3/M input $2.5/M output
30 providers
Reasoning Tools JSON

Gemini 2.5 Flash-Lite is a lightweight reasoning model in the Gemini 2.5 family, optimized for ultra-low latency and cost efficiency. It offers improved throughput, faster token generation, and better performance...

google/gemini-2.5-flash-lite 2025-06-17 1.04858M context $0.1/M input $0.4/M output
18 providers
Reasoning Tools

The o-series of models are trained with reinforcement learning to think before they answer and perform complex reasoning. The o3-pro model uses more compute to think harder and provide consistently...

openai/o3-pro 2025-06-10 200K context $20/M input $80/M output
8 providers
Reasoning Tools

Balanced Claude model for coding, analysis, agent workflows, and cost control

anthropic/claude-sonnet-4-20250514 2025-05-22 200K context $3/M input $15/M output
11 providers
Reasoning Tools

Flagship Claude model for deep reasoning, coding, and long-horizon agents

anthropic/claude-opus-4-20250514 2025-05-22 200K context $15/M input $75/M output
8 providers
Reasoning Tools JSON

OpenAI o4-mini is a compact reasoning model in the o-series, optimized for fast, cost-efficient performance while retaining strong multimodal and agentic capabilities. It supports tool use and demonstrates competitive reasoning...

openai/o4-mini 2025-04-16 200K context $1.1/M input $4.4/M output
20 providers
Reasoning

o3 is a well-rounded and powerful model across domains. It sets a new standard for math, science, coding, and visual reasoning tasks. It also excels at technical writing and instruction-following....

openai/o3 2025-04-16 200K context $2/M input $8/M output
20 providers
Reasoning Tools

Mistral reasoning model for transparent analysis, math, and complex decisions

mistral/magistral-medium-latest 2025-03-17 128K context $2/M input $5/M output
6 providers
Reasoning Tools Open weights

DeepSeek R1 is here: Performance on par with [OpenAI o1](/openai/o1), but open-sourced and with fully open reasoning tokens. It's 671B parameters in size, with 37B active in an inference pass....

deepseek/deepseek-r1 2025-01-20 64K context $0.7/M input $2.5/M output
14 providers
Reasoning Tools JSON

OpenAI o3-mini is a cost-efficient language model optimized for STEM reasoning tasks, particularly excelling in science, mathematics, and coding. This model supports the `reasoning_effort` parameter, which can be set to...

openai/o3-mini 2024-12-20 200K context $1.1/M input $4.4/M output
19 providers
Tools Open weights

Popular open Llama workhorse for multilingual chat, coding, and self-hosting

meta/llama-3.3-70b-instruct 2024-12-06 128K context $0.1/M input $0.32/M output
25 providers
Reasoning

The latest and strongest model family from OpenAI, o1 is designed to spend more time thinking before responding. The o1 model series is trained with large-scale reinforcement learning to reason...

openai/o1 2024-12-05 200K context $15/M input $60/M output
16 providers
Open weights

Efficient Mistral-NVIDIA open model for multilingual chat and local deployment

mistral/mistral-nemo 2024-07-01 128K context $0.15/M input $0.15/M output
11 providers
Reasoning Tools

Research model for long-horizon investigation, synthesis, and analytical reports

openai/o3-deep-research 2024-06-26 200K context $9/M input $36/M output
6 providers
JSON

Note: Sonar Pro pricing includes Perplexity search pricing. See [details here](https://docs.perplexity.ai/guides/pricing#detailed-pricing-breakdown-for-sonar-reasoning-pro-and-sonar-pro) Sonar Reasoning Pro is a premier reasoning model powered by DeepSeek R1 with Chain of Thought (CoT). Designed for...

perplexity/sonar-reasoning-pro 2024-01-01 128K context $2/M input $8/M output
9 providers