166 models
Reasoning Tools JSON

GPT-6 Astra is OpenAI's flagship model for demanding end-to-end work. It is suited for advanced analysis, software engineering, deep research, scientific work, and document creation, with particular strengths in long-horizon...

openai/gpt-6-astra 2026-09-04 1.05M context $10/M input $50/M output
18 providers
Reasoning Tools JSON

2026-09-02 upgraded snapshot of Qwen3.8 Max with stronger coding, collaborative agents, and multimodal document understanding

alibaba/qwen3.8-max-0902 2026-09-02 1M context $1.71/M input $5.14/M output
7 providers
Reasoning Tools JSON

Gemini 3.8 Flash is Google's most intelligent Flash model with significant gains from 3.7 Flash across software engineering, agentic tasks, and multi-step reasoning.

google/gemini-3.8-flash 2026-09-02 1.04858M context $0.75/M input $3.75/M output
17 providers
Reasoning Tools JSON

Claude model for demanding reasoning and long-horizon agentic work

anthropic/claude-fable-5-1 2026-09-01 1M context $10/M input $50/M output
24 providers
Reasoning Tools JSON Open weights

Native multimodal GLM model for efficient coding and long-horizon agent tasks

zhipuai/glm-5.3-flash 2026-08-26 1M context $0.075/M input $0.25/M output
36 providers
Reasoning Tools JSON

Qwen vision-language model for visual reasoning, documents, and agent tasks

alibaba/qwen3.8-flash 2026-08-26 1M context $0.15/M input $0.47/M output
21 providers
Reasoning Tools JSON

DeepSeek V4 Flash Vision Exp is an experimental vision-enabled version of [DeepSeek V4 Flash 0731](https://openrouter.ai/deepseek/deepseek-v4-flash-0731) from DeepSeek, adding image understanding while matching the base model on text capabilities including agents,...

deepseek/deepseek-v4-flash-vision-exp 2026-08-21 1.04858M context $0.22/M input $0.66/M output
19 providers
Reasoning Tools JSON Open weights

Dense 27B vision-language model for coding, agent tasks, and image and video understanding

alibaba/qwen3.8-27b 2026-08-14 262.144K context $0.1/M input $0.4/M output
27 providers
Reasoning Tools JSON Open weights

Flagship GLM model for long-horizon coding, agents, and complex project delivery

zhipuai/glm-5.3 2026-08-14 1M context $1.4/M input $4.4/M output
37 providers
Reasoning Tools JSON

Gemini 3.7 Flash is a multimodal model from Google for fast agentic workflows, coding, and complex multi-step reasoning. It is designed for tasks that require responsive performance and reliable multi-step...

google/gemini-3.7-flash 2026-08-13 1.04858M context $0.75/M input $3.75/M output
23 providers
Reasoning Tools JSON

High-efficiency Gemini model for agentic workflows, coding, and multimodal reasoning

google/gemini-flash-latest 2026-08-13 1.04858M context $0.75/M input $3.75/M output
7 providers
Reasoning Tools JSON

xAI's frontier model for long-running agents, coding, knowledge work, and visual projects

xai/grok-4.6 2026-08-12 500K context $2/M input $6/M output
27 providers
Reasoning Tools JSON Open weights

Open-weight sparse MoE (2.4T total, 95B active), the open-weight twin of Qwen3.8 Max for coding, research, complex reasoning, and agentic workflows

alibaba/qwen3.8-2.4t-a95b 2026-08-12 262.144K context $2/M input $6/M output
12 providers
Reasoning Tools JSON Open weights

DeepSeek V4 Pro 0813 is a large-scale mixture-of-experts model from DeepSeek. This is the GA release of DeepSeek V4 Pro.

deepseek/deepseek-v4-pro-0813 2026-08-12 1.024M context $0.579/M input $1.738/M output
30 providers
Reasoning Tools JSON Open weights

Muse Glimmer 30B is a dense, open-weight multimodal model from Meta Superintelligence Labs, distilled from Muse Spark and optimized for autonomous agents on consumer hardware. It is suited for long-horizon...

meta/muse-glimmer-30b 2026-08-10 131.072K context $0.3/M input $1.1/M output
9 providers
Tools JSON

2.4-trillion-parameter MoE flagship for coding, professional work, multimodal understanding, and long-horizon agentic workflows

alibaba/qwen3.8-max 2026-08-03 1M context $2/M input $6/M output
28 providers
Reasoning Tools JSON Open weights

DeepSeek V4 Flash 0731 is a sparse mixture-of-experts model from DeepSeek, with 13B active parameters out of 284B total. This re-post-trained revision is suited for coding, reasoning, and agent workflows....

deepseek/deepseek-v4-flash-0731 2026-07-31 1.04858M context $0.065/M input $0.18/M output
43 providers
Tools Open weights

Inkling Small is an open-weight multimodal mixture-of-experts model from Thinking Machines Lab, with 12B active parameters out of 276B total. It is positioned as the smaller, more efficient member of...

thinkingmachines/inkling-small 2026-07-30 524.288K context $0.45/M input $1.2/M output
11 providers
Reasoning Tools

Claude Opus 5 is Anthropic’s flagship model for demanding reasoning, coding, and long-horizon agentic work. It is particularly strong at end-to-end software tasks, code review and bug finding, visual analysis...

anthropic/claude-opus-5 2026-07-24 1M context $5/M input $25/M output
32 providers
Reasoning Tools JSON

Gemini 3.5 Flash Lite is a high-efficiency model from Google with upgraded agentic capabilities. It is suited for subagents that execute focused tasks within complex, multi-agent workflows.

google/gemini-3.5-flash-lite 2026-07-21 1.04858M context $0.3/M input $2.5/M output
23 providers
Reasoning Tools JSON

Gemini 3.6 Flash is a high-efficiency model from Google for coding, agentic workflows, and web and app development. It is designed to produce polished outputs with fewer unnecessary edits and...

google/gemini-3.6-flash 2026-07-21 1.04858M context $0.75/M input $3.75/M output
24 providers
Reasoning Tools JSON Open weights

Kimi K3 is a 2.8T parameter open-weight multimodal reasoning model from Moonshot AI. It is suited for complex coding, knowledge work, and long-horizon agentic workflows, and is particularly strong at...

moonshotai/kimi-k3 2026-07-16 1.04858M context $2.34/M input $11.7/M output
63 providers
Tools JSON Open weights

Inkling is an open-weight multimodal mixture-of-experts model from Thinking Machines Lab, with 41B active parameters out of 975B total. It is designed for general-purpose reasoning, coding, agentic and tool-use systems,...

thinkingmachines/inkling 2026-07-15 1.04858M context $1/M input $4.05/M output
21 providers
Reasoning Tools JSON

GPT-5.6 Sol is the flagship model in OpenAI's GPT-5.6 series. It is suited for complex reasoning, coding, and agentic workflows, and is particularly strong at command-line and multi-step coding tasks...

openai/gpt-5.6-sol 2026-07-09 1.05M context $2/M input $10/M output
36 providers
Reasoning Tools JSON

GPT-5.6 Terra is a balanced model in OpenAI's GPT-5.6 series, positioned between the flagship Sol tier and the cost-efficient Luna tier. It is suited for everyday coding, reasoning, and agentic...

openai/gpt-5.6-terra 2026-07-09 1.05M context $2/M input $12/M output
35 providers
Reasoning Tools JSON

GPT-5.6 Luna is a fast, cost-efficient model in OpenAI's GPT-5.6 series. It is suited for high-volume, latency-sensitive tasks such as chat, classification, and lightweight agentic workflows, providing capable reasoning for...

openai/gpt-5.6-luna 2026-07-09 1.05M context $0.2/M input $1.2/M output
36 providers
Reasoning Tools JSON

xAI's Grok model for chat, coding, agentic tools, and lower hallucination risk

xai/grok-4.5 2026-07-08 500K context $2/M input $6/M output
26 providers
Reasoning Tools JSON Open weights

Hy3 is a 295B-parameter Mixture-of-Experts model from Tencent (21B active, 192 experts with top-8 routing) built for reasoning, agentic workflows, and real-world production use. It supports a configurable reasoning effort:...

tencent/hy3 2026-07-06 262.144K context $0.083/M input $0.33/M output
17 providers
Reasoning Tools

Nano Banana 2 Lite (Gemini 3.1 Flash Lite Image) is Google's fastest, most cost-efficient Gemini image model, built for high-velocity developer pipelines and rapid-fire visual exploration. It delivers text-to-image generation...

google/gemini-3.1-flash-lite-image 2026-06-30 65.536K context $0.25/M input $1.5/M output
5 providers
Tools JSON

Sonnet 5 is Anthropic's most capable Sonnet-class model, with frontier performance across coding, agents, and professional work. It supports adaptive thinking with selectable reasoning effort levels (low, medium, high, max,...

anthropic/claude-sonnet-5 2026-06-30 1M context $2/M input $10/M output
35 providers
Reasoning Tools JSON Open weights

Open flagship GLM for long-horizon coding agents and million-token context work

zhipuai/glm-5.2 2026-06-13 1M context $1.4/M input $4.4/M output
72 providers
Reasoning Tools JSON Open weights

MoonshotAI: Kimi K2.7 Code is a coding-focused model in Moonshot AI's Kimi K2 family, built to complete end-to-end programming tasks reliably over long contexts. It uses a native multimodal mixture-of-experts...

moonshotai/kimi-k2.7-code 2026-06-12 262.144K context $0.71/M input $3.5/M output
61 providers
Reasoning Tools Open weights

Lower-latency Kimi Code variant for interactive edits and coding-agent loops

moonshotai/kimi-k2.7-code-highspeed 2026-06-12 262.144K context $1.9/M input $8/M output
9 providers
Reasoning Tools

Claude Fable 5 is a Mythos-class model from Anthropic, built for autonomous knowledge work and coding. It supports text, image, and file inputs with text output, with reasoning support and...

anthropic/claude-fable-5 2026-06-09 1M context $10/M input $50/M output
37 providers
Reasoning Tools Open weights

NVIDIA Nemotron 3 Ultra is an open frontier-reasoning and orchestration model from NVIDIA, with 55B active parameters out of 550B total (MoE). Built on a hybrid Transformer-Mamba mixture-of-experts architecture, it...

nvidia/nemotron-3-ultra-550b-a55b 2026-06-04 256K context $0.625/M input $3.125/M output
15 providers
Reasoning Tools Open weights

MiniMax multimodal model for long-context coding, perception, and agent planning

minimax/MiniMax-M3 2026-06-01 1.04858M context $0.3/M input $1.2/M output
37 providers
Reasoning Tools JSON Open weights

Step 3.7 Flash is StepFun's latest high-efficiency multimodal Mixture-of-Experts model. It pairs a 196B-parameter language backbone with a vision encoder for native image and video understanding, activating roughly 11B parameters...

stepfun/step-3.7-flash 2026-05-29 256K context $0.2/M input $1.15/M output
17 providers
Reasoning

Gemini 3.1 Flash Image, a.k.a. "Nano Banana 2," is Google’s latest state of the art image generation and editing model, delivering Pro-level visual quality at Flash speed. It combines advanced...

google/gemini-3.1-flash-image 2026-05-28 131.072K context $0.5/M input $3/M output
9 providers
Reasoning Tools

Top Claude Opus tier for the hardest reasoning, coding, and long-horizon agents

anthropic/claude-opus-4-8 2026-05-28 1M context $5/M input $25/M output
44 providers
Reasoning

Nano Banana Pro is Google’s most advanced image-generation and editing model, built on Gemini 3 Pro. It extends the original Nano Banana with significantly improved multimodal reasoning, real-world grounding, and...

google/gemini-3-pro-image 2026-05-28 65.536K context $2/M input $12/M output
9 providers
Reasoning Tools JSON

Gemini 3.5 Flash is Google's high-efficiency multimodal model, bringing near-Pro level coding and reasoning at Flash-tier cost and speed. It is highly optimized for coding proficiency and parallel agentic execution...

google/gemini-3.5-flash 2026-05-19 1.04858M context $1.5/M input $9/M output
31 providers
Reasoning Tools JSON

Gemini 3.1 Flash Lite is Google’s GA high-efficiency multimodal model optimized for low-latency, high-volume workloads. It supports text, image, video, audio, and PDF inputs, and is designed for lightweight agentic...

google/gemini-3.1-flash-lite 2026-05-07 1.04858M context $0.25/M input $1.5/M output
23 providers
Tools JSON Open weights

Balanced Mistral model for enterprise assistants, multilingual work, and tools

mistral/mistral-medium-latest 2026-04-29 262.144K context $1.5/M input $7.5/M output
5 providers
Reasoning Tools JSON Open weights

Balanced Mistral model for enterprise assistants, multilingual work, and tools

mistral/mistral-medium-2604 2026-04-29 262.144K context $1.5/M input $7.5/M output
9 providers
Reasoning Tools JSON Open weights

DeepSeek V4 Flash is an efficiency-optimized Mixture-of-Experts model from DeepSeek with 284B total parameters and 13B activated parameters, supporting a 1M-token context window. It is designed for fast inference and...

deepseek/deepseek-v4-flash 2026-04-24 1.024M context $0.086/M input $0.172/M output
62 providers
Reasoning Tools JSON Open weights

DeepSeek V4 Pro is a large-scale Mixture-of-Experts model from DeepSeek with 1.6T total parameters and 49B activated parameters, supporting a 1M-token context window. It is designed for advanced reasoning, coding,...

deepseek/deepseek-v4-pro 2026-04-24 1.024M context $0.86/M input $1.72/M output
62 providers
Reasoning Tools JSON

GPT-5.5 Pro is OpenAI’s high-capability model optimized for deep reasoning and accuracy on complex, high-stakes workloads. It features a 1M+ token context window (922K input, 128K output) with support for...

openai/gpt-5.5-pro 2026-04-23 1.05M context $30/M input $180/M output
18 providers
Reasoning Tools JSON

GPT-5.5 is OpenAI’s frontier model designed for complex professional workloads, building on GPT-5.4 with stronger reasoning, higher reliability, and improved token efficiency on hard tasks. It features a 1M+ token...

openai/gpt-5.5 2026-04-23 1.05M context $5/M input $30/M output
46 providers
Reasoning Tools JSON Open weights

Kimi K2.6 is Moonshot AI's next-generation multimodal model, designed for long-horizon coding, coding-driven UI/UX generation, and multi-agent orchestration. It handles complex end-to-end coding tasks across Python, Rust, and Go, and...

moonshotai/kimi-k2.6 2026-04-21 262.144K context $0.95/M input $4/M output
64 providers
Reasoning Tools JSON

xAI's default Grok for chat, coding, agentic tools, and lower hallucination risk

xai/grok-4.3 2026-04-17 1M context $1.25/M input $2.5/M output
28 providers