Claude Fable 5 is a Mythos-class model from Anthropic, built for autonomous knowledge work and coding. It supports text, image, and file inputs with text output, with reasoning support and...
Models
Every model in the catalog with source-linked pricing, context limits, provider availability, and published benchmark results.
Granite 4.2 8B is a dense reasoning model from IBM. It is suited for mathematics, code generation, multilingual dialogue, and agentic workflows that need multi-step reasoning. It supports full, low-effort,...
No provider description is available for this model yet.
No provider description is available for this model yet.
MiniMax-M2 is a compact, high-efficiency large language model optimized for end-to-end coding and agentic workflows. With 10 billion activated parameters (230 billion total), it delivers near-frontier intelligence across general reasoning,...
GPT-4o mini is OpenAI's newest model after [GPT-4 Omni](/models/openai/gpt-4o), supporting both text and image inputs with text outputs. As their most advanced small model, it is many multiples more affordable...
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
GLM-4.5V is a vision-language foundation model for multimodal agent applications. Built on a Mixture-of-Experts (MoE) architecture with 106B parameters and 12B activated parameters, it achieves state-of-the-art results in video understanding,...
No provider description is available for this model yet.
For tasks that demand low latency, GPT‑4.1 nano is the fastest and cheapest model in the GPT-4.1 series. It delivers exceptional performance at a small size with its 1 million...
Relace Apply 3 is a specialized code-patching LLM that merges AI-suggested edits straight into your source files. It can apply updates from GPT-4o, Claude, and others into your files at...
NVIDIA Nemotron 3.5 Lightning is an open mixture-of-experts model from NVIDIA, with 3B active parameters out of 30B total. It is suited for high-throughput agentic workloads and specialized tasks that...
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
Qwen2.5-VL is proficient in recognizing common objects such as flowers, birds, fish, and insects. It is also highly capable of analyzing texts, charts, icons, graphics, and layouts within images.
No provider description is available for this model yet.
Fugu Max is the cost-performance model in Sakana AI's Fugu family. Rather than a single monolithic model, Fugu is a learned multi-agent orchestration system: a language model trained to route...
A balanced model in the Ministral 3 family, Ministral 3 8B is a powerful, efficient tiny language model with vision capabilities.
No provider description is available for this model yet.
*Ling-3.0-flash* is a *124B-parameter Mixture-of-Experts (MoE) model*, with approximately *5.1B parameters activated per token*. The model is designed with *token efficiency and production-scale agentic inference* as key priorities, enabling developers...
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
A balanced model in the Ministral 3 family, Ministral 3 8B is a powerful, efficient tiny language model with vision capabilities.
Mistral's cutting-edge language model for coding released end of July 2025. Codestral specializes in low-latency, high-frequency tasks such as fill-in-the-middle (FIM), code correction and test generation. [Blog Post](https://mistral.ai/news/codestral-25-08)
OpenAI o4-mini-high is the same model as [o4-mini](/openai/o4-mini) with reasoning_effort set to high. OpenAI o4-mini is a compact reasoning model in the o-series, optimized for fast, cost-efficient performance while retaining...
No provider description is available for this model yet.
No provider description is available for this model yet.
Ministral 8B is an 8B parameter model featuring a unique interleaved sliding-window attention pattern for faster, memory-efficient inference. Designed for edge use cases, it supports up to 128k context length...
No provider description is available for this model yet.
No provider description is available for this model yet.
Opus 4.7 is the next generation of Anthropic's Opus family, built for long-running, asynchronous agents. Building on the coding and agentic strengths of Opus 4.6, it delivers stronger performance on...
No provider description is available for this model yet.
DeepSeek-V3.2-Exp is an experimental large language model released by DeepSeek as an intermediate step between V3.1 and future architectures. It introduces DeepSeek Sparse Attention (DSA), a fine-grained sparse attention mechanism...
Gemini 3.1 Pro Preview is Google’s frontier reasoning model, delivering enhanced software engineering performance, improved agentic reliability, and more efficient token usage across complex workflows. Building on the multimodal foundation...
No provider description is available for this model yet.
GPT-4.1 is a flagship large language model optimized for advanced instruction following, real-world software engineering, and long-context reasoning. It supports a 1 million token context window and outperforms GPT-4o and...
GLM-5 is Z.ai’s flagship open-source foundation model engineered for complex systems design and long-horizon agent workflows. Built for expert developers, it delivers production-grade performance on large-scale programming tasks, rivaling leading...
Claude Opus 4.1 is an updated version of Anthropic’s flagship model, offering improved performance in coding, reasoning, and agentic tasks. It achieves 74.5% on SWE-bench Verified and shows notable gains...
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
| Model | Creator | Inputs | Context | Input | Output | Released | Compare |
|---|---|---|---|---|---|---|---|
| Anthropic: Claude Fable 5 (batch)anthropic/claude-fable-5:batch | 1M | $5 | $25 | — | |||
| IBM: Granite 4.2 8Bibm-granite/granite-4.2-8b | 131.072K | $0.06 | $0.25 | — | |||
| openai/gpt-4.1-nanovercel_ai_gateway/openai/gpt-4.1-nano | 1.04758M | $0.1 | $0.4 | — | |||
| anthropic/claude-opus-4.5openrouter/anthropic/claude-opus-4.5 | 200K | $5 | $25 | — | |||
| MiniMax: MiniMax M2minimax/minimax-m2 | 204.8K | $0.255 | $1.02 | — | |||
| OpenAI: GPT-4o-mini (batch)openai/gpt-4o-mini:batch | 128K | $0.075 | $0.3 | — | |||
| openai/gpt-4.1-minivercel_ai_gateway/openai/gpt-4.1-mini | 1.04758M | $0.4 | $1.6 | — | |||
| openai/gpt-4.1vercel_ai_gateway/openai/gpt-4.1 | 1.04758M | $2 | $8 | — | |||
| anthropic/claude-sonnet-4.6openrouter/anthropic/claude-sonnet-4.6 | 1M | $3 | $15 | — | |||
| magistral-medium-2509mistral/magistral-medium-2509 | 40K | $2 | $5 | — | |||
| Z.ai: GLM 4.5Vz-ai/glm-4.5v | 65.536K | $0.6 | $1.8 | — | |||
| openai/gpt-4-turbovercel_ai_gateway/openai/gpt-4-turbo | 128K | $10 | $30 | — | |||
| OpenAI: GPT-4.1 Nano (batch)openai/gpt-4.1-nano:batch | 1.04758M | $0.05 | $0.2 | — | |||
| Relace: Relace Apply 3relace/relace-apply-3 | 256K | $0.85 | $1.25 | — | |||
| NVIDIA: Nemotron 3.5 Lightning (free)nvidia/nemotron-3.5-lightning:free | 1M | Free | Free | — | |||
| openai/gpt-3.5-turbo-instructvercel_ai_gateway/openai/gpt-3.5-turbo-instruct | 8.192K | $1.5 | $2 | — | |||
| anthropic/claude-sonnet-4openrouter/anthropic/claude-sonnet-4 | 1M | $3 | $15 | — | |||
| openai/gpt-5.6-solopenrouter/openai/gpt-5.6-sol | 1.05M | $2 | $10 | — | |||
| Qwen: Qwen2.5 VL 72B Instructqwen/qwen2.5-vl-72b-instruct | 128K | $0.8 | $1 | — | |||
| openai/gpt-3.5-turbovercel_ai_gateway/openai/gpt-3.5-turbo | 16.385K | $0.5 | $1.5 | — | |||
| Sakana: Fugu Maxsakana/fugu-max | 1M | $2 | $6 | — | |||
| Mistral: Ministral 3 8B 2512 (batch)mistralai/ministral-8b-2512:batch | 262.144K | $0.075 | $0.075 | — | |||
| morph/morph-v3-largevercel_ai_gateway/morph/morph-v3-large | 32.768K | $0.9 | $1.9 | — | |||
| inclusionAI: Ling 3.0 Flashinclusionai/ling-3.0-flash | 262.144K | $0.021 | $0.063 | — | |||
| anthropic/claude-opus-4.1openrouter/anthropic/claude-opus-4.1 | 200K | $15 | $75 | — | |||
| magistral-medium-2506mistral/magistral-medium-2506 | 40K | $2 | $5 | — | |||
| jp.anthropic.claude-sonnet-4-5-20250929-v1:0bedrock_converse/jp.anthropic.claude-sonnet-4-5-20250929-v1:0 | 200K | $3.3 | $16.5 | — | |||
| openai/gpt-oss-120bgroq/openai/gpt-oss-120b | 131.072K | $0.15 | $0.6 | — | |||
| Mistral: Ministral 3 8B 2512mistralai/ministral-8b-2512 | 262.144K | $0.15 | $0.15 | — | |||
| Mistral: Codestral 2508mistralai/codestral-2508 | 256K | $0.3 | $0.9 | — | |||
| OpenAI: o4 Mini Highopenai/o4-mini-high | 200K | $1.1 | $4.4 | — | |||
| openai-o3gradient_ai/openai-o3 | 200K | $2 | $8 | — | |||
| alibaba-qwen3-32bgradient_ai/alibaba-qwen3-32b | 131.072K | — | — | — | |||
| Mistral: Ministral 8Bmistralai/ministral-8b | 128K | $0.11 | $0.11 | — | |||
| morph/morph-v3-fastvercel_ai_gateway/morph/morph-v3-fast | 32.768K | $0.8 | $1.2 | — | |||
| moonshotai/kimi-k2vercel_ai_gateway/moonshotai/kimi-k2 | 131.072K | $0.55 | $2.2 | — | |||
| Anthropic: Claude Opus 4.7 (batch)anthropic/claude-opus-4.7:batch | 1M | $2.5 | $12.5 | — | |||
| anthropic/claude-opus-4openrouter/anthropic/claude-opus-4 | 200K | $15 | $75 | — | |||
| DeepSeek: DeepSeek V3.2 Expdeepseek/deepseek-v3.2-exp | 163.84K | $0.27 | $0.41 | — | |||
| Google: Gemini 3.1 Pro Preview (batch)google/gemini-3.1-pro-preview:batch | 1.04858M | $1 | $6 | — | |||
| mistral/pixtral-largevercel_ai_gateway/mistral/pixtral-large | 128K | $2 | $6 | — | |||
| OpenAI: GPT-4.1 (batch)openai/gpt-4.1:batch | 1.04758M | $1 | $4 | — | |||
| Z.ai: GLM 5z-ai/glm-5 | 198K | $0.6 | $1.92 | — | |||
| Anthropic: Claude Opus 4.1 (batch)anthropic/claude-opus-4.1:batch | 200K | $7.5 | $37.5 | — | |||
| mistral/pixtral-12bvercel_ai_gateway/mistral/pixtral-12b | 128K | $0.15 | $0.15 | — | |||
| gpt-chat-latestazure_ai/gpt-chat-latest | 272K | $5 | $30 | — | |||
| model-routerazure_ai/model-router | 200K | $0.14 | — | — | |||
| qwen/qwen3-coderopenrouter/qwen/qwen3-coder | 262.1K | $0.22 | $0.95 | — | |||
| cohere-command-aazure_ai/cohere-command-a | 131.072K | $2.5 | $10 | — | |||
| devstral-latestmistral/devstral-latest | 256K | $0.4 | $2 | — |