1,112 models
Reasoning Tools JSON Open weights

DeepSeek V4.1 Flash model for reasoning and agentic coding

deepseek/deepseek-v4.1-flash 2026-09-10 1M context $0.15/M input $0.6/M output
17 providers
Reasoning Tools JSON

Fast variant of GPT-6 Astra for low-latency assistance and high-volume workloads.

openai/gpt-6-astra-fast 2026-09-04 1.05M context $20/M input $100/M output
1 provider
Reasoning Tools JSON

GPT-6 Astra is OpenAI's most capable model for complex reasoning, coding, computer use, research, and document creation.

openai/gpt-6-astra 2026-09-04 1.05M context $10/M input $50/M output
18 providers
Reasoning Tools JSON

Google's most intelligent Flash model, engineered for long-horizon software engineering, autonomous agents, and complex enterprise workflows

google/gemini-3.8-flash 2026-09-02 1.04858M context $0.75/M input $3.75/M output
17 providers
Reasoning Tools JSON

2026-09-02 upgraded snapshot of Qwen3.8 Max with stronger coding, collaborative agents, and multimodal document understanding

alibaba/qwen3.8-max-0902 2026-09-02 1M context $1.71/M input $5.14/M output
7 providers
Reasoning Tools JSON

Muse Spark 1.3 is a multimodal reasoning model from Meta for long-running agentic, multi-agent, and coding workflows. It improves long-horizon agent collaboration, instruction following, and coding efficiency relative to Muse Spark 1.2.

meta/muse-spark-1.3 2026-09-02 1.04858M context $1.25/M input $4.25/M output
9 providers
Reasoning Tools JSON

Claude model for demanding reasoning and long-horizon agentic work

anthropic/claude-fable-5-1 2026-09-01 1M context $10/M input $50/M output
24 providers
Reasoning Tools JSON Open weights

Open-weight experimental preview of the Qwen4 architecture: hybrid-attention MoE (125B total, 6B active) with vision encoder for coding, agent tasks, and image and video understanding

alibaba/qwen3.8-flash-next 2026-08-27 262.144K context $0.12/M input $0.4/M output
4 providers
Reasoning Tools JSON

Qwen vision-language model for visual reasoning, documents, and agent tasks

alibaba/qwen3.8-flash 2026-08-26 1M context $0.15/M input $0.47/M output
21 providers
Reasoning Tools JSON Open weights

Native multimodal GLM model for efficient coding and long-horizon agent tasks

zhipuai/glm-5.3-flash 2026-08-26 1M context $0.075/M input $0.25/M output
35 providers
Reasoning Tools JSON

Experimental multimodal DeepSeek V4 Flash model for image understanding, coding, and agentic work

deepseek/deepseek-v4-flash-vision-exp 2026-08-21 1M context $0.15/M input $0.6/M output
19 providers
Reasoning Tools JSON Open weights

Mixture-of-experts coding-reasoning model for agentic software tasks, tool use, and image understanding

deepreinforce/ornith-1.5-35b-a3b 2026-08-18 262.144K context $0.1/M input $0.4/M output
2 providers
Reasoning Tools JSON Open weights

Dense 27B vision-language model for coding, agent tasks, and image and video understanding

alibaba/qwen3.8-27b 2026-08-14 262.144K context $0.1/M input $0.4/M output
27 providers
Reasoning Tools JSON

High-efficiency Gemini model for agentic workflows, coding, and multimodal reasoning

google/gemini-flash-latest 2026-08-13 1.04858M context $0.75/M input $3.75/M output
7 providers
Reasoning Tools JSON

High-efficiency Gemini model for agentic workflows, coding, and multimodal reasoning

google/gemini-3.7-flash 2026-08-13 1.04858M context $0.75/M input $3.75/M output
23 providers
Reasoning Tools JSON

xAI's frontier model for long-running agents, coding, knowledge work, and visual projects

xai/grok-4.6 2026-08-12 500K context $2/M input $6/M output
27 providers
Reasoning Tools JSON

Microsoft coding model with native vision support, optimized for fast and efficient software development

microsoft/mai-code-1.1-flash 2026-08-11 256K context $0.2/M input $1.2/M output
1 provider
Reasoning Tools JSON

Muse Spark 1.2 is a coding-focused update to Muse Spark 1.1 with improvements in code generation, complex debugging, codebase understanding, and end-to-end developer workflows.

meta/muse-spark-1.2 2026-08-05 1.04858M context $1.25/M input $4.25/M output
14 providers
Reasoning Tools JSON

Japanese-specialized reasoning model based on Kimi K2.6 and tuned for Japanese language, culture, and business workflows

sakana/sakana-namazu 2026-08-03 262.144K context $0.95/M input $4/M output
4 providers
Tools JSON

2.4-trillion-parameter MoE flagship for coding, professional work, multimodal understanding, and long-horizon agentic workflows

alibaba/qwen3.8-max 2026-08-03 1M context $2/M input $6/M output
28 providers
Tools Open weights

Multimodal MoE reasoning model (276B total, 12B active) for text, image, and audio

thinkingmachines/inkling-small 2026-07-30 1.04858M context $0.45/M input $1.2/M output
10 providers
Reasoning Tools

Strongest Claude Opus model for coding, agents, and professional work

anthropic/claude-opus-5 2026-07-24 1M context $5/M input $25/M output
32 providers
Reasoning Tools JSON

Fast Gemini model balancing multimodal reasoning, tool use, and cost

google/gemini-flash-lite-latest 2026-07-21 1.04858M context $0.3/M input $2.5/M output
6 providers
Reasoning Tools JSON

Fast Gemini model balancing multimodal reasoning, tool use, and cost

google/gemini-3.5-flash-lite 2026-07-21 1.04858M context $0.3/M input $2.5/M output
23 providers
Reasoning Tools JSON

Fast Gemini model balancing multimodal reasoning, tool use, and cost

google/gemini-3.6-flash 2026-07-21 1.04858M context $0.75/M input $3.75/M output
24 providers
Reasoning Tools

Preview Qwen flagship for million-token multimodal reasoning and long-horizon agentic workflows

alibaba/qwen3.8-max-preview 2026-07-19 1M context $2/M input $6/M output
6 providers
Reasoning Tools JSON Open weights

Multimodal Kimi model with 1M context and toggleable max-effort thinking for long-horizon agent work

moonshotai/kimi-k3 2026-07-16 1.04858M context $3/M input $15/M output
62 providers
Tools JSON Open weights

Multimodal MoE reasoning model (975B total, 41B active) for text, image, and audio

thinkingmachines/inkling 2026-07-15 1.04858M context $1.87/M input $4.68/M output
21 providers
Reasoning Tools JSON

Lightweight multimodal Qwen model for high-throughput text, image, and video tasks

alibaba/qwen3.7-flash 2026-07-15 1M context $0.028/M input $0.113/M output
11 providers
Reasoning Tools JSON

Balanced GPT-5.6 model for capable, cost-efficient everyday work

openai/gpt-5.6-terra 2026-07-09 1.05M context $2/M input $12/M output
35 providers
Reasoning Tools JSON

Frontier GPT-5.6 model for complex professional work, coding, and agentic workflows

openai/gpt-5.6-sol 2026-07-09 1.05M context $4/M input $20/M output
36 providers
Reasoning Tools JSON

Cost-efficient GPT-5.6 model for fast, high-volume workloads

openai/gpt-5.6-luna 2026-07-09 1.05M context $0.2/M input $1.2/M output
36 providers
Reasoning Tools JSON

xAI's Grok model for chat, coding, agentic tools, and lower hallucination risk

xai/grok-4.5 2026-07-08 500K context $2/M input $6/M output
26 providers
Tools JSON

Everyday Claude agent model for coding, planning, browsing, and general work

anthropic/claude-sonnet-5 2026-06-30 1M context $2/M input $10/M output
35 providers
Reasoning

Video generation and editing model for fast, conversational text- and image-to-video workflows

google/gemini-omni-flash-preview 2026-06-30 1.04858M context $1.5/M input $17.5/M output
2 providers
Open weights

Large coding-reasoning model for agentic software tasks and RL search

deepreinforce/ornith-1.0-35b 2026-06-25 262.144K context Input not listed Output not listed
Open weights

Large coding-reasoning model for agentic software tasks and RL search

deepreinforce/ornith-1.0-397b 2026-06-25 262.144K context Input not listed Output not listed
Open weights

Open coding-reasoning model for repository tasks and self-improving agents

deepreinforce/ornith-1.0-9b 2026-06-25 262.144K context Input not listed Output not listed
Open weights

Open coding-reasoning model for repository tasks and self-improving agents

deepreinforce/ornith-1.0-31b 2026-06-25 262.144K context Input not listed Output not listed
Reasoning Tools JSON

Rolling ByteDance Seed model for rapidly updated reasoning, coding, and agent capabilities

bytedance-seed/seed-evolving 2026-06-23 256K context $0.884/M input $4.42/M output
3 providers
Reasoning Tools JSON

Faster ByteDance Seed 2.1 model for multimodal reasoning and latency-sensitive agent workflows

bytedance-seed/seed-2.1-turbo 2026-06-23 256K context $0.354/M input $1.77/M output
7 providers
Reasoning Tools JSON

Flagship ByteDance Seed 2.1 model for complex multimodal reasoning, coding, and agents

bytedance-seed/seed-2.1-pro 2026-06-23 256K context $0.707/M input $3.536/M output
2 providers
Reasoning Tools JSON

ByteDance Seed model optimized for character-driven dialogue and consistent conversational behavior

bytedance-seed/seed-character 2026-06-23 256K context $0.119/M input $0.297/M output
2 providers
Reasoning Tools JSON

Quality-first multi-agent model for hard research, analysis, and competitions

sakana/fugu-ultra 2026-06-15 1M context $5/M input $30/M output
11 providers
Reasoning Tools JSON

Multi-agent model for routing expert agents across complex analytical tasks

sakana/fugu 2026-06-15 1M context Input not listed Output not listed
1 provider
Reasoning Tools Open weights

Lower-latency Kimi Code variant for interactive edits and coding-agent loops

moonshotai/kimi-k2.7-code-highspeed 2026-06-12 262.144K context $1.9/M input $8/M output
9 providers
Reasoning Tools JSON Open weights

Coding-focused Kimi model, stronger on long-horizon repo work with less overthinking

moonshotai/kimi-k2.7-code 2026-06-12 262.144K context $0.95/M input $4/M output
61 providers
Reasoning Tools

Claude model for creative writing, analysis, and controlled agent workflows

anthropic/claude-fable-5 2026-06-09 1M context $10/M input $50/M output
37 providers
Reasoning Tools JSON

Restricted Claude model for advanced cybersecurity and biology research workflows

anthropic/claude-mythos-5 2026-06-09 1M context $10/M input $50/M output
2 providers
Reasoning Tools

Multimodal Qwen workhorse for long-context agents, visual inputs, and coding

alibaba/qwen3.7-plus 2026-06-02 1M context $0.5/M input $3/M output
29 providers