2,824 models
Reasoning

Codex GPT for repository edits, code review, and practical software agents

openai/gpt-5.1-codex 2025-11-13 400K context $1.07/M input $8.5/M output
19 providers
Reasoning Tools JSON

Sharper GPT-5 generation for coding, product work, and tool-assisted tasks

openai/gpt-5.1 2025-11-13 400K context $1.25/M input $10/M output
28 providers
Reasoning Tools JSON

Coding-optimized GPT model for repository edits, reviews, and agentic software work

openai/gpt-5.1-codex-max 2025-11-13 400K context $1.1/M input $9/M output
15 providers
Reasoning Tools JSON

Chat-tuned GPT-5.1 for polished assistants, writing, and product conversations

openai/gpt-5.1-chat-latest 2025-11-13 128K context $1.25/M input $10/M output
7 providers
Reasoning

Coding-optimized GPT model for repository edits, reviews, and agentic software work

openai/gpt-5.1-codex-mini 2025-11-13 400K context $0.22/M input $1.8/M output
15 providers
Reasoning Tools Open weights

Kimi reasoning model for long-horizon research, planning, and tool use

moonshotai/kimi-k2-thinking-turbo 2025-11-06 262.144K context $1.15/M input $8/M output
3 providers
Tools Open weights

Thinking Kimi model for slower research passes, planning, and hard technical questions

moonshotai/kimi-k2-thinking 2025-11-06 262.144K context $0.4/M input $2.5/M output
22 providers
Reasoning Tools JSON

Flagship Claude model for deep reasoning, coding, and long-horizon agents

anthropic/claude-opus-4-5-20251101 2025-11-01 200K context $5/M input $25/M output
17 providers
Reasoning Open weights

Safety model for policy screening, moderation, and risk-aware routing workflows

openai/gpt-oss-safeguard-20b 2025-10-29 131.072K context $0.07/M input $0.2/M output
7 providers
Reasoning Tools JSON Open weights

Safety model for policy screening, moderation, and risk-aware routing workflows

openai/gpt-oss-safeguard-120b 2025-10-29 131.072K context $0.15/M input $0.6/M output
4 providers
Open weights

Safety model for policy screening, moderation, and risk-aware routing workflows

nvidia/llama-3.1-nemotron-safety-guard-8b-v3 2025-10-28 128K context Input not listed Output not listed
1 provider
Reasoning Tools Open weights

Nemotron multimodal model for visual reasoning and agentic AI workflows

nvidia/nemotron-nano-12b-v2-vl 2025-10-28 128K context $0.2/M input $0.6/M output
3 providers
Reasoning Tools Open weights

Efficient open MiniMax model built for coding agents and tool-heavy workflows

minimax/MiniMax-M2 2025-10-27 204.8K context $0.3/M input $1.2/M output
14 providers
Tools JSON

Fast Claude model for responsive assistance, classification, and lightweight agents

anthropic/claude-haiku-4-5-20251001 2025-10-15 200K context $1/M input $5/M output
20 providers
Reasoning Tools JSON

ByteDance Seed model for long-context reasoning, instruction following, and tool-assisted tasks

bytedance-seed/seed-1-6 2025-10-15 256K context $0.119/M input $1.187/M output
4 providers
Reasoning Tools

Fast Claude lane for lightweight agents, office tasks, and responsive chat

anthropic/claude-haiku-4-5 2025-10-15 200K context $1/M input $5/M output
23 providers
Reasoning Tools JSON

Specialized Gemini 2.5 model for browser-control agents that automate UI tasks

google/gemini-2.5-computer-use-preview-10-2025 2025-10-07 128K context $1.25/M input $10/M output
2 providers
Reasoning

Higher-accuracy GPT-5 tier for tough analysis, coding reviews, and planning

openai/gpt-5-pro 2025-10-06 400K context $15/M input $120/M output
15 providers
Open weights

Compact open-weight hybrid Granite model for lightweight enterprise chat and tool calling

ibm/granite-4-h-micro 2025-10-02 131.072K context Input not listed Output not listed
Tools JSON Open weights

Open-weight hybrid model for enterprise chat, coding, retrieval-augmented generation, and tool-calling workloads

ibm/granite-4-h-small 2025-10-02 131.072K context $0.064/M input $0.265/M output
1 provider
Reasoning Tools Open weights

Late GLM-4 workhorse for coding agents, reasoning, and structured tasks

zhipuai/glm-4.6 2025-09-30 204.8K context $0.6/M input $2.2/M output
18 providers
Reasoning Tools

Balanced Claude model for coding, analysis, agent workflows, and cost control

anthropic/claude-sonnet-4-5 2025-09-29 200K context $3/M input $15/M output
20 providers
Reasoning Tools JSON

Balanced Claude model for coding, analysis, agent workflows, and cost control

anthropic/claude-sonnet-4-5-20250929 2025-09-29 200K context $3/M input $15/M output
18 providers
Reasoning Open weights

Qwen vision-language thinking model for visual reasoning, documents, and agent tasks

alibaba/qwen3-vl-235b-a22b-thinking 2025-09-23 131.072K context $0.4/M input $4/M output
8 providers

Flagship Qwen3 model for coding agents, complex reasoning, and tool use

alibaba/qwen3-max 2025-09-23 262.144K context $1.2/M input $6/M output
20 providers

Qwen vision-language model for visual reasoning, documents, and agent tasks

alibaba/qwen3-vl-plus 2025-09-23 262.144K context $0.2/M input $1.6/M output
6 providers
Open weights

Gemma 3 27B tuned by AI Singapore for Southeast Asian languages and instruction following

aisingapore/gemma-sea-lion-v4-27b-it 2025-09-23 128K context Input not listed Output not listed
1 provider
Open weights

Qwen vision-language instruct model for visual reasoning, documents, and agent tasks

alibaba/qwen3-vl-235b-a22b-instruct 2025-09-23 131.072K context $0.2/M input $0.88/M output
12 providers
Reasoning Tools JSON Open weights

Open multimodal reasoning model for transparent analysis of text and images

mistral/magistral-small-2509 2025-09-18 131.072K context $0.5/M input $1.5/M output
1 provider
Reasoning Tools JSON

Coding-optimized GPT model for repository edits, reviews, and agentic software work

openai/gpt-5-codex 2025-09-15 400K context $1.1/M input $9/M output
13 providers
Reasoning Tools Open weights

Flagship Indian-language reasoning model for enterprise multilingual applications

sarvam/sarvam-105b 2025-09-01 131.072K context $0.04/M input $0.16/M output
2 providers
Tools Open weights

Qwen instruction model for multilingual chat, reasoning, and tool use

alibaba/qwen3-next-80b-a3b-instruct 2025-09 131.072K context $0.5/M input $2/M output
13 providers
Reasoning Tools Open weights

Efficient Qwen thinking model for local reasoning, math, and coding agents

alibaba/qwen3-next-80b-a3b-thinking 2025-09 131.072K context $0.5/M input $6/M output
10 providers
Reasoning Tools JSON

Low-latency ByteDance Seed model for high-throughput chat, extraction, and lightweight tool use

bytedance-seed/seed-1-6-flash 2025-08-28 256K context $0.022/M input $0.223/M output
3 providers
Reasoning Tools Open weights

Cohere reasoning model for multilingual enterprise agents, tools, and complex workflows

cohere/command-a-reasoning-08-2025 2025-08-21 256K context $2.5/M input $10/M output
1 provider
Reasoning Tools Open weights

Hybrid-reasoning DeepSeek model with thinking and non-thinking modes

deepseek/deepseek-v3.1 2025-08-21 131.072K context $0.19/M input $0.71/M output
10 providers
Reasoning Tools Open weights

Compact Nemotron model for efficient reasoning and deployable AI agents

nvidia/nemotron-nano-9b-v2 2025-08-18 131.072K context $0.06/M input $0.23/M output
4 providers
Reasoning Tools JSON

ByteDance Seed multimodal model for image understanding, visual reasoning, and tool-assisted tasks

bytedance-seed/seed-1-6-vision 2025-08-15 256K context $0.119/M input $1.187/M output
2 providers
Reasoning

Tiny GPT-5 lane for routing, extraction, classification, and bulk jobs

openai/gpt-5-nano 2025-08-07 400K context $0.05/M input $0.4/M output
26 providers
Reasoning

Small GPT-5 for responsive agents, coding help, and everyday automation

openai/gpt-5-mini 2025-08-07 400K context $0.25/M input $2/M output
29 providers
Tools

Original GPT-5 workhorse for reasoning, coding, writing, and tool workflows

openai/gpt-5 2025-08-07 400K context $1.25/M input $10/M output
30 providers
Reasoning JSON

Chat-tuned GPT model for conversational assistance, writing, and tool workflows

openai/gpt-5-chat-latest 2025-08-07 400K context $1.25/M input $10/M output
5 providers
Reasoning Tools JSON Open weights

Open GPT reasoning model for self-hosted agents and controllable deployments

openai/gpt-oss-120b 2025-08-05 131.072K context $0.03/M input $0.17/M output
53 providers
Reasoning Tools Open weights

Open GPT reasoning model for self-hosted agents and controllable deployments

openai/gpt-oss-20b 2025-08-05 131.072K context $0.02/M input $0.1/M output
32 providers
Reasoning Tools JSON

Flagship Claude model for deep reasoning, coding, and long-horizon agents

anthropic/claude-opus-4-1-20250805 2025-08-05 200K context $15/M input $75/M output
13 providers
Reasoning Tools

Flagship Claude model for deep reasoning, coding, and long-horizon agents

anthropic/claude-opus-4-1 2025-08-05 200K context $15/M input $75/M output
9 providers
Open weights

Cohere vision model for multilingual document analysis, OCR, and image understanding

cohere/command-a-vision-07-2025 2025-07-31 128K context $2.5/M input $10/M output
1 provider
Tools

Efficient Qwen model for fast chat, extraction, and high-volume workloads

alibaba/qwen-flash 2025-07-28 1M context $0.05/M input $0.4/M output
7 providers
Reasoning Tools Open weights

Lighter GLM-4.5 variant for fast coding assistance and cheaper agents

zhipuai/glm-4.5-air 2025-07-28 131.072K context $0.2/M input $1.1/M output
11 providers

Qwen coding model for software agents, repository edits, and code reasoning

alibaba/qwen3-coder-flash 2025-07-28 1M context $0.3/M input $1.5/M output
12 providers