Reasoning Tools JSON Open weights

DeepSeek V4.1 Flash is a sparse mixture-of-experts model from DeepSeek, and the first built on the company's Causal Encoder-Decoder (CED) architecture. It activates 8B parameters on input and 16B on...

deepseek/deepseek-v4.1-flash 2026-09-10 1.04858M context $0.15/M input $0.6/M output
21 providers
Reasoning Tools JSON

GPT-6 Astra is OpenAI's flagship model for demanding end-to-end work. It is suited for advanced analysis, software engineering, deep research, scientific work, and document creation, with particular strengths in long-horizon...

openai/gpt-6-astra 2026-09-04 1.05M context $10/M input $50/M output
18 providers
Reasoning Tools JSON

Gemini 3.8 Flash is Google's most intelligent Flash model with significant gains from 3.7 Flash across software engineering, agentic tasks, and multi-step reasoning.

google/gemini-3.8-flash 2026-09-02 1.04858M context $0.75/M input $3.75/M output
17 providers
Reasoning Tools JSON

2026-09-02 upgraded snapshot of Qwen3.8 Max with stronger coding, collaborative agents, and multimodal document understanding

alibaba/qwen3.8-max-0902 2026-09-02 1M context $1.71/M input $5.14/M output
7 providers
Reasoning Tools JSON

Claude model for demanding reasoning and long-horizon agentic work

anthropic/claude-fable-5-1 2026-09-01 1M context $10/M input $50/M output
24 providers
Reasoning Tools JSON Open weights

Native multimodal GLM model for efficient coding and long-horizon agent tasks

zhipuai/glm-5.3-flash 2026-08-26 1M context $0.075/M input $0.25/M output
36 providers
Reasoning Tools JSON

Qwen vision-language model for visual reasoning, documents, and agent tasks

alibaba/qwen3.8-flash 2026-08-26 1M context $0.15/M input $0.47/M output
21 providers
Reasoning Tools JSON

DeepSeek V4 Flash Vision Exp is an experimental vision-enabled version of [DeepSeek V4 Flash 0731](https://openrouter.ai/deepseek/deepseek-v4-flash-0731) from DeepSeek, adding image understanding while matching the base model on text capabilities including agents,...

deepseek/deepseek-v4-flash-vision-exp 2026-08-21 1.04858M context $0.22/M input $0.66/M output
19 providers
Reasoning Tools JSON Open weights

Dense 27B vision-language model for coding, agent tasks, and image and video understanding

alibaba/qwen3.8-27b 2026-08-14 262.144K context $0.1/M input $0.4/M output
27 providers
Reasoning Tools JSON Open weights

Flagship GLM model for long-horizon coding, agents, and complex project delivery

zhipuai/glm-5.3 2026-08-14 1M context $1.4/M input $4.4/M output
37 providers
Reasoning Tools JSON

Gemini 3.7 Flash is a multimodal model from Google for fast agentic workflows, coding, and complex multi-step reasoning. It is designed for tasks that require responsive performance and reliable multi-step...

google/gemini-3.7-flash 2026-08-13 1.04858M context $0.75/M input $3.75/M output
23 providers
Reasoning Tools JSON Open weights

DeepSeek V4 Pro 0813 is a large-scale mixture-of-experts model from DeepSeek. This is the GA release of DeepSeek V4 Pro.

deepseek/deepseek-v4-pro-0813 2026-08-12 1.024M context $0.579/M input $1.738/M output
30 providers
Reasoning Tools JSON

xAI's frontier model for long-running agents, coding, knowledge work, and visual projects

xai/grok-4.6 2026-08-12 500K context $2/M input $6/M output
27 providers
Tools JSON

2.4-trillion-parameter MoE flagship for coding, professional work, multimodal understanding, and long-horizon agentic workflows

alibaba/qwen3.8-max 2026-08-03 1M context $2/M input $6/M output
28 providers
Reasoning Tools JSON Open weights

DeepSeek V4 Flash 0731 is a sparse mixture-of-experts model from DeepSeek, with 13B active parameters out of 284B total. This re-post-trained revision is suited for coding, reasoning, and agent workflows....

deepseek/deepseek-v4-flash-0731 2026-07-31 1.04858M context $0.065/M input $0.18/M output
43 providers
Reasoning Tools

Claude Opus 5 is Anthropic’s flagship model for demanding reasoning, coding, and long-horizon agentic work. It is particularly strong at end-to-end software tasks, code review and bug finding, visual analysis...

anthropic/claude-opus-5 2026-07-24 1M context $5/M input $25/M output
32 providers
Reasoning Tools JSON

Gemini 3.5 Flash Lite is a high-efficiency model from Google with upgraded agentic capabilities. It is suited for subagents that execute focused tasks within complex, multi-agent workflows.

google/gemini-3.5-flash-lite 2026-07-21 1.04858M context $0.3/M input $2.5/M output
23 providers
Reasoning Tools JSON

Gemini 3.6 Flash is a high-efficiency model from Google for coding, agentic workflows, and web and app development. It is designed to produce polished outputs with fewer unnecessary edits and...

google/gemini-3.6-flash 2026-07-21 1.04858M context $0.75/M input $3.75/M output
24 providers
Reasoning Tools JSON Open weights

Kimi K3 is a 2.8T parameter open-weight multimodal reasoning model from Moonshot AI. It is suited for complex coding, knowledge work, and long-horizon agentic workflows, and is particularly strong at...

moonshotai/kimi-k3 2026-07-16 1.04858M context $2.34/M input $11.7/M output
63 providers
Reasoning Tools JSON

GPT-5.6 Luna is a fast, cost-efficient model in OpenAI's GPT-5.6 series. It is suited for high-volume, latency-sensitive tasks such as chat, classification, and lightweight agentic workflows, providing capable reasoning for...

openai/gpt-5.6-luna 2026-07-09 1.05M context $0.2/M input $1.2/M output
36 providers
Reasoning Tools JSON

GPT-5.6 Sol is the flagship model in OpenAI's GPT-5.6 series. It is suited for complex reasoning, coding, and agentic workflows, and is particularly strong at command-line and multi-step coding tasks...

openai/gpt-5.6-sol 2026-07-09 1.05M context $2/M input $10/M output
36 providers
Reasoning Tools JSON

GPT-5.6 Terra is a balanced model in OpenAI's GPT-5.6 series, positioned between the flagship Sol tier and the cost-efficient Luna tier. It is suited for everyday coding, reasoning, and agentic...

openai/gpt-5.6-terra 2026-07-09 1.05M context $2/M input $12/M output
35 providers
Reasoning Tools JSON

xAI's Grok model for chat, coding, agentic tools, and lower hallucination risk

xai/grok-4.5 2026-07-08 500K context $2/M input $6/M output
26 providers
Tools JSON

Sonnet 5 is Anthropic's most capable Sonnet-class model, with frontier performance across coding, agents, and professional work. It supports adaptive thinking with selectable reasoning effort levels (low, medium, high, max,...

anthropic/claude-sonnet-5 2026-06-30 1M context $2/M input $10/M output
35 providers
Reasoning Tools JSON

Flagship ByteDance Seed 2.1 model for complex multimodal reasoning, coding, and agents

bytedance-seed/seed-2.1-pro 2026-06-23 256K context $0.707/M input $3.536/M output
2 providers
Reasoning Tools JSON

Rolling ByteDance Seed model for rapidly updated reasoning, coding, and agent capabilities

bytedance-seed/seed-evolving 2026-06-23 256K context $0.884/M input $4.42/M output
3 providers
Reasoning Tools JSON

ByteDance Seed model optimized for character-driven dialogue and consistent conversational behavior

bytedance-seed/seed-character 2026-06-23 256K context $0.119/M input $0.297/M output
2 providers
Reasoning Tools JSON

Faster ByteDance Seed 2.1 model for multimodal reasoning and latency-sensitive agent workflows

bytedance-seed/seed-2.1-turbo 2026-06-23 256K context $0.354/M input $1.77/M output
7 providers
Reasoning Tools JSON Open weights

Open flagship GLM for long-horizon coding agents and million-token context work

zhipuai/glm-5.2 2026-06-13 1M context $1.4/M input $4.4/M output
72 providers
Reasoning Tools Open weights

Lower-latency Kimi Code variant for interactive edits and coding-agent loops

moonshotai/kimi-k2.7-code-highspeed 2026-06-12 262.144K context $1.9/M input $8/M output
9 providers
Reasoning Tools JSON Open weights

MoonshotAI: Kimi K2.7 Code is a coding-focused model in Moonshot AI's Kimi K2 family, built to complete end-to-end programming tasks reliably over long contexts. It uses a native multimodal mixture-of-experts...

moonshotai/kimi-k2.7-code 2026-06-12 262.144K context $0.71/M input $3.5/M output
61 providers
Reasoning Tools

Claude Fable 5 is a Mythos-class model from Anthropic, built for autonomous knowledge work and coding. It supports text, image, and file inputs with text output, with reasoning support and...

anthropic/claude-fable-5 2026-06-09 1M context $10/M input $50/M output
37 providers
Reasoning Tools

Multimodal Qwen workhorse for long-context agents, visual inputs, and coding

alibaba/qwen3.7-plus 2026-06-02 1M context $0.5/M input $3/M output
29 providers
Reasoning Tools Open weights

MiniMax multimodal model for long-context coding, perception, and agent planning

minimax/MiniMax-M3 2026-06-01 1.04858M context $0.3/M input $1.2/M output
37 providers
Reasoning Tools

Top Claude Opus tier for the hardest reasoning, coding, and long-horizon agents

anthropic/claude-opus-4-8 2026-05-28 1M context $5/M input $25/M output
44 providers
Reasoning Tools

Qwen frontier model tuned for agent frameworks, coding assistants, and long tasks

alibaba/qwen3.7-max 2026-05-21 1M context $2.5/M input $7.5/M output
31 providers
Reasoning Tools JSON

Gemini 3.5 Flash is Google's high-efficiency multimodal model, bringing near-Pro level coding and reasoning at Flash-tier cost and speed. It is highly optimized for coding proficiency and parallel agentic execution...

google/gemini-3.5-flash 2026-05-19 1.04858M context $1.5/M input $9/M output
31 providers
Reasoning Tools JSON

Gemini 3.1 Flash Lite is Google’s GA high-efficiency multimodal model optimized for low-latency, high-volume workloads. It supports text, image, video, audio, and PDF inputs, and is designed for lightweight agentic...

google/gemini-3.1-flash-lite 2026-05-07 1.04858M context $0.25/M input $1.5/M output
23 providers

Qwen vision-language model for visual reasoning, documents, and agent tasks

alibaba/qwen3.6-flash 2026-04-27 1M context $0.188/M input $1.125/M output
22 providers
Reasoning Tools JSON Open weights

DeepSeek V4 Pro is a large-scale Mixture-of-Experts model from DeepSeek with 1.6T total parameters and 49B activated parameters, supporting a 1M-token context window. It is designed for advanced reasoning, coding,...

deepseek/deepseek-v4-pro 2026-04-24 1.024M context $0.86/M input $1.72/M output
62 providers
Reasoning Tools JSON Open weights

DeepSeek V4 Flash is an efficiency-optimized Mixture-of-Experts model from DeepSeek with 284B total parameters and 13B activated parameters, supporting a 1M-token context window. It is designed for fast inference and...

deepseek/deepseek-v4-flash 2026-04-24 1.024M context $0.086/M input $0.172/M output
62 providers
Reasoning Tools JSON Open weights

DeepSeek V4 Pro initial snapshot with million-token context and support for thinking and non-thinking modes

deepseek/deepseek-v4-pro-0423 2026-04-23 1M context $1.32/M input $3.96/M output
2 providers
Reasoning Tools JSON

GPT-5.5 is OpenAI’s frontier model designed for complex professional workloads, building on GPT-5.4 with stronger reasoning, higher reliability, and improved token efficiency on hard tasks. It features a 1M+ token...

openai/gpt-5.5 2026-04-23 1.05M context $5/M input $30/M output
46 providers
Reasoning Open weights

Initial DeepSeek V4 Flash snapshot for economical reasoning, coding, and million-token agent workloads

deepseek/deepseek-v4-flash-0423 2026-04-23 1M context $0.139/M input $0.278/M output
2 providers
Tools Open weights

Qwen vision-language model for visual reasoning, documents, and agent tasks

alibaba/qwen3.6-27b 2026-04-22 262.144K context $0.6/M input $3.6/M output
24 providers
Reasoning Tools JSON Open weights

Kimi K2.6 is Moonshot AI's next-generation multimodal model, designed for long-horizon coding, coding-driven UI/UX generation, and multi-agent orchestration. It handles complex end-to-end coding tasks across Python, Rust, and Go, and...

moonshotai/kimi-k2.6 2026-04-21 262.144K context $0.95/M input $4/M output
64 providers

Flagship Qwen model for complex reasoning, coding, and agentic workflows

alibaba/qwen3.6-max-preview 2026-04-20 262.144K context $1.3/M input $7.8/M output
12 providers
Reasoning Tools JSON

xAI's default Grok for chat, coding, agentic tools, and lower hallucination risk

xai/grok-4.3 2026-04-17 1M context $1.25/M input $2.5/M output
28 providers
Reasoning Tools

Stronger Opus tier for advanced software work and high-stakes reasoning

anthropic/claude-opus-4-7 2026-04-16 1M context $5/M input $25/M output
40 providers
Reasoning Tools JSON Open weights

Strong GLM coding model for agentic engineering, terminals, and repository generation

zhipuai/glm-5.1 2026-04-07 200K context $1.4/M input $4.4/M output
45 providers