166 models
Tools Open weights

Kimi K2 Thinking is Moonshot AI’s most advanced open reasoning model to date, extending the K2 series into agentic, long-horizon reasoning. Built on the trillion-parameter Mixture-of-Experts (MoE) architecture introduced in...

moonshotai/kimi-k2-thinking 2025-11-06 262.144K context $0.6/M input $2.5/M output
22 providers
Reasoning Tools JSON

Flagship Claude model for deep reasoning, coding, and long-horizon agents

anthropic/claude-opus-4-5-20251101 2025-11-01 200K context $5/M input $25/M output
17 providers
Reasoning Tools Open weights

Efficient open MiniMax model built for coding agents and tool-heavy workflows

minimax/MiniMax-M2 2025-10-27 204.8K context $0.3/M input $1.2/M output
14 providers
Reasoning

GPT-5 Pro is OpenAI’s most advanced model, offering major improvements in reasoning, code quality, and user experience. It is optimized for complex tasks that require step-by-step reasoning, instruction following, and...

openai/gpt-5-pro 2025-10-06 400K context $15/M input $120/M output
15 providers
Reasoning Tools Open weights

Late GLM-4 workhorse for coding agents, reasoning, and structured tasks

zhipuai/glm-4.6 2025-09-30 204.8K context $0.6/M input $2.2/M output
18 providers
Open weights

Gemma 3 27B tuned by AI Singapore for Southeast Asian languages and instruction following

aisingapore/gemma-sea-lion-v4-27b-it 2025-09-23 128K context Input not listed Output not listed
1 provider

Flagship Qwen3 model for coding agents, complex reasoning, and tool use

alibaba/qwen3-max 2025-09-23 262.144K context $1.2/M input $6/M output
20 providers
Reasoning Open weights

Qwen vision-language thinking model for visual reasoning, documents, and agent tasks

alibaba/qwen3-vl-235b-a22b-thinking 2025-09-23 131.072K context $0.4/M input $4/M output
8 providers
Open weights

Qwen vision-language instruct model for visual reasoning, documents, and agent tasks

alibaba/qwen3-vl-235b-a22b-instruct 2025-09-23 131.072K context $0.2/M input $0.88/M output
12 providers
Tools Open weights

Qwen instruction model for multilingual chat, reasoning, and tool use

alibaba/qwen3-next-80b-a3b-instruct 2025-09 131.072K context $0.5/M input $2/M output
13 providers
Reasoning Tools Open weights

Efficient Qwen thinking model for local reasoning, math, and coding agents

alibaba/qwen3-next-80b-a3b-thinking 2025-09 131.072K context $0.5/M input $6/M output
10 providers

Gemini 2.5 Flash Image, a.k.a. "Nano Banana," is now generally available. It is a state of the art image generation model with contextual understanding. It is capable of image generation,...

google/gemini-2.5-flash-image 2025-08-26 32.768K context $0.3/M input $2.5/M output
8 providers
Reasoning Tools Open weights

Compact Nemotron model for efficient reasoning and deployable AI agents

nvidia/nemotron-nano-9b-v2 2025-08-18 131.072K context $0.06/M input $0.23/M output
4 providers
Tools

GPT-5 is OpenAI’s most advanced model, offering major improvements in reasoning, code quality, and user experience. It is optimized for complex tasks that require step-by-step reasoning, instruction following, and accuracy...

openai/gpt-5 2025-08-07 400K context $1.25/M input $10/M output
30 providers
Reasoning

GPT-5 Mini is a compact version of GPT-5, designed to handle lighter-weight reasoning tasks. It provides the same instruction-following and safety-tuning benefits as GPT-5, but with reduced latency and cost....

openai/gpt-5-mini 2025-08-07 400K context $0.25/M input $2/M output
29 providers
Reasoning

GPT-5-Nano is the smallest and fastest variant in the GPT-5 system, optimized for developer tools, rapid interactions, and ultra-low latency environments. While limited in reasoning depth compared to its larger...

openai/gpt-5-nano 2025-08-07 400K context $0.05/M input $0.4/M output
26 providers
Reasoning Tools JSON Open weights

gpt-oss-120b is an open-weight, 117B-parameter Mixture-of-Experts (MoE) language model from OpenAI designed for high-reasoning, agentic, and general-purpose production use cases. It activates 5.1B parameters per forward pass and is optimized...

openai/gpt-oss-120b 2025-08-05 131.072K context $0.037/M input $0.17/M output
53 providers
Reasoning Tools Open weights

gpt-oss-20b is an open-weight 21B parameter model released by OpenAI under the Apache 2.0 license. It uses a Mixture-of-Experts (MoE) architecture with 3.6B active parameters per forward pass, optimized for...

openai/gpt-oss-20b 2025-08-05 131.072K context $0.03/M input $0.13/M output
32 providers

Qwen coding model for software agents, repository edits, and code reasoning

alibaba/qwen3-coder-flash 2025-07-28 1M context $0.3/M input $1.5/M output
12 providers

Hosted Qwen coder for software agents, repo edits, and long-context code

alibaba/qwen3-coder-plus 2025-07-23 1.04858M context $1/M input $5/M output
14 providers
Tools Open weights

Updated large open Qwen3 MoE instruct model for multilingual chat, coding, and tool use

alibaba/qwen3-235b-a22b-instruct-2507 2025-07-21 262.144K context $0.069/M input $0.455/M output
7 providers
Tools Open weights

Instruct model with native audio input for speech understanding and tool use

mistral/voxtral-small-latest 2025-07-15 32K context $0.1/M input $0.3/M output
3 providers
Reasoning Tools

The o-series of models are trained with reinforcement learning to think before they answer and perform complex reasoning. The o3-pro model uses more compute to think harder and provide consistently...

openai/o3-pro 2025-06-10 200K context $20/M input $80/M output
8 providers

Mistral model for multilingual chat, reasoning, and tool-assisted workflows

mistral/mistral-medium-2505 2025-05-07 131.072K context $0.4/M input $2/M output
10 providers
Reasoning Tools JSON

OpenAI o4-mini is a compact reasoning model in the o-series, optimized for fast, cost-efficient performance while retaining strong multimodal and agentic capabilities. It supports tool use and demonstrates competitive reasoning...

openai/o4-mini 2025-04-16 200K context $1.1/M input $4.4/M output
20 providers
Reasoning

o3 is a well-rounded and powerful model across domains. It sets a new standard for math, science, coding, and visual reasoning tasks. It also excels at technical writing and instruction-following....

openai/o3 2025-04-16 200K context $2/M input $8/M output
20 providers
Tools Open weights

Nemotron model for efficient reasoning, coding, and specialized AI agents

nvidia/llama-3.1-nemotron-70b-instruct 2025-04-15 128K context Input not listed Output not listed
2 providers
Tools JSON

GPT-4.1 is a flagship large language model optimized for advanced instruction following, real-world software engineering, and long-context reasoning. It supports a 1 million token context window and outperforms GPT-4o and...

openai/gpt-4.1 2025-04-14 1.04758M context $2/M input $8/M output
28 providers

GPT-4.1 Mini is a mid-sized model delivering performance competitive with GPT-4o at substantially lower latency and cost. It retains a 1 million token context window and scores 45.1% on hard...

openai/gpt-4.1-mini 2025-04-14 1.04758M context $0.4/M input $1.6/M output
25 providers

For tasks that demand low latency, GPT‑4.1 nano is the fastest and cheapest model in the GPT-4.1 series. It delivers exceptional performance at a small size with its 1 million...

openai/gpt-4.1-nano 2025-04-14 1.04758M context $0.1/M input $0.4/M output
20 providers
Tools Open weights

Smaller Qwen coder for efficient local agents and repo-level fixes

alibaba/qwen3-coder-30b-a3b-instruct 2025-04 262.144K context $0.45/M input $2.25/M output
13 providers
Tools Open weights

Open Qwen coding heavyweight for repository reasoning and agentic engineering

alibaba/qwen3-coder-480b-a35b-instruct 2025-04 262.144K context $1.5/M input $7.5/M output
8 providers
Tools JSON Open weights

March 2025 checkpoint of DeepSeek-V3 with improved reasoning and coding

deepseek/deepseek-v3-0324 2025-03-24 163.84K context $0.2/M input $0.8/M output
9 providers

The o1 series of models are trained with reinforcement learning to think before they answer and perform complex reasoning. The o1-pro model uses more compute to think harder and provide...

openai/o1-pro 2025-03-19 200K context $150/M input $600/M output
7 providers
Reasoning Tools

Mistral reasoning model for transparent analysis, math, and complex decisions

mistral/magistral-medium-latest 2025-03-17 128K context $2/M input $5/M output
6 providers
Reasoning Tools Open weights

Cohere command model for multilingual enterprise agents, tools, and chat

cohere/command-a-03-2025 2025-03-13 256K context $2.5/M input $10/M output
6 providers
Reasoning Tools

Qwen reasoning model for deliberate problem solving, math, and coding

alibaba/qwq-plus 2025-03-05 131.072K context $0.8/M input $2.4/M output
4 providers
Reasoning

Sonar Deep Research is a research-focused model designed for multi-step retrieval, synthesis, and reasoning across complex topics. It autonomously searches, reads, and evaluates sources, refining its approach as it gathers...

perplexity/sonar-deep-research 2025-02-01 128K context $2/M input $8/M output
4 providers
Reasoning Tools Open weights

DeepSeek R1 is here: Performance on par with [OpenAI o1](/openai/o1), but open-sourced and with fully open reasoning tokens. It's 671B parameters in size, with 37B active in an inference pass....

deepseek/deepseek-r1 2025-01-20 64K context $0.7/M input $2.5/M output
14 providers
Tools Open weights

Open DeepSeek MoE chat model for coding, math, and general reasoning

deepseek/deepseek-v3 2024-12-26 131.072K context $0.27/M input $1.12/M output
8 providers
Reasoning Tools JSON

OpenAI o3-mini is a cost-efficient language model optimized for STEM reasoning tasks, particularly excelling in science, mathematics, and coding. This model supports the `reasoning_effort` parameter, which can be set to...

openai/o3-mini 2024-12-20 200K context $1.1/M input $4.4/M output
19 providers
Tools Open weights

Popular open Llama workhorse for multilingual chat, coding, and self-hosting

meta/llama-3.3-70b-instruct 2024-12-06 128K context $0.1/M input $0.32/M output
25 providers
Reasoning

The latest and strongest model family from OpenAI, o1 is designed to spend more time thinking before responding. The o1 model series is trained with large-scale reinforcement learning to reason...

openai/o1 2024-12-05 200K context $15/M input $60/M output
16 providers
Tools

Efficient model for low-latency assistance, extraction, and routine automation

amazon/nova-micro 2024-12-03 128K context $0.035/M input $0.14/M output
3 providers
Tools

Efficient model for low-latency assistance, extraction, and routine automation

amazon/nova-lite 2024-12-03 300K context $0.06/M input $0.24/M output
3 providers
Tools

Flagship model for demanding analysis, coding, and production agent workflows

amazon/nova-pro 2024-12-03 300K context $0.8/M input $3.2/M output
3 providers
JSON Open weights

Command R7B (12-2024) is a small, fast update of the Command R+ model, delivered in December 2024. It excels at RAG, tool use, agents, and similar tasks requiring complex reasoning...

cohere/command-r7b-12-2024 2024-12-02 128K context $0.037/M input $0.15/M output
5 providers
Tools

The 2024-11-20 version of GPT-4o offers a leveled-up creative writing ability with more natural, engaging, and tailored writing to improve relevance & readability. It’s also better at working with uploaded...

openai/gpt-4o-2024-11-20 2024-11-20 128K context $2.5/M input $10/M output
8 providers
Tools JSON Open weights

Open coding-focused Qwen model for code generation, repair, and repository reasoning

alibaba/qwen2.5-coder-32b-instruct 2024-11-12 131.072K context $0.06/M input $0.2/M output
2 providers
Open weights

Flagship Mistral model for advanced reasoning, coding, and multilingual work

mistral/mistral-large-latest 2024-11-01 262.144K context $0.5/M input $1.5/M output
7 providers