No provider description is available for this model yet.
Models
Every model in the catalog with source-linked pricing, context limits, provider availability, and published benchmark results.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
GPT-5.2 Chat (AKA Instant) is the fast, lightweight member of the 5.2 family, optimized for low-latency chat while retaining strong general intelligence. It uses adaptive reasoning to selectively “think” on...
GLM-4.7 is Z.ai’s latest flagship model, featuring upgrades in two key areas: enhanced programming capabilities and more stable multi-step reasoning/execution. It demonstrates significant improvements in executing complex agent tasks while...
No provider description is available for this model yet.
GPT-5.6 Luna Pro is the same underlying model as [GPT-5.6 Luna](https://openrouter.ai/openai/gpt-5.6-luna), served with `reasoning.mode` set to `pro` for higher-quality responses on complex tasks. Learn more in OpenAI's docs: https://developers.openai.com/api/docs/guides/reasoning#reasoning-mode
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
Ling 3.0 Tiny is a mixture-of-experts model from InclusionAI, with 1.3B active parameters out of 7.9B total. It is designed for responsive agents, instruction following, and multi-turn conversations, with switchable...
Schematron V2 Turbo is a 3B-parameter HTML-to-JSON extraction model from Inference.net. It prioritizes throughput for high-volume extraction workloads. Extraction instructions must be supplied through a JSON schema in response_format rather...
MiniMax-M2.1 is a lightweight, state-of-the-art large language model optimized for coding, agentic workflows, and modern application development. With only 10 billion activated parameters, it delivers a major jump in real-world...
The gpt-audio model is OpenAI's first generally available audio model. The new snapshot features an upgraded decoder for more natural sounding voices and maintains better voice consistency. Audio is priced...
No provider description is available for this model yet.
Gemini 3.5 Flash is Google's high-efficiency multimodal model, bringing near-Pro level coding and reasoning at Flash-tier cost and speed. It is highly optimized for coding proficiency and parallel agentic execution...
No provider description is available for this model yet.
MiniMax M2-her is a dialogue-first large language model built for immersive roleplay, character-driven chat, and expressive multi-turn conversations. Designed to stay consistent in tone and personality, it supports rich message...
No provider description is available for this model yet.
Solar Pro 3 is Upstage's powerful Mixture-of-Experts (MoE) language model. With 102B total parameters and 12B active parameters per forward pass, it delivers exceptional performance while maintaining computational efficiency. Optimized...
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
The Qwen3.5 native vision-language series Plus models are built on a hybrid architecture that integrates linear attention mechanisms with sparse mixture-of-experts models, achieving higher inference efficiency. In a variety of...
Aion-2.0 is a variant of DeepSeek V3.2 optimized for immersive roleplaying and storytelling. It is particularly strong at introducing tension, crises, and conflict into stories, making narratives feel more engaging....
The o1 series of models are trained with reinforcement learning to think before they answer and perform complex reasoning. The o1-pro model uses more compute to think harder and provide...
No provider description is available for this model yet.
No provider description is available for this model yet.
The Qwen3.5 122B-A10B native vision-language model is built on a hybrid architecture that integrates a linear attention mechanism with a sparse mixture-of-experts model, achieving higher inference efficiency. In terms of...
The Qwen3.5 27B native vision-language Dense model incorporates a linear attention mechanism, delivering fast response times while balancing inference speed and performance. Its overall capabilities are comparable to those of...
Ling 3.0 Flash VL builds on Ling 3.0 Flash (124B total / 5.5B active MoE from InclusionAI), further strengthening its language capabilities while adding native visual perception and advanced visual...
Muse Spark 1.2 contributor tier is a reasoning model from Meta designed for developers who want to start building at an even lower cost. It’s meaningfully cheaper than Muse Spark...
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
GPT-5.1 is the latest frontier-grade model in the GPT-5 series, offering stronger general-purpose reasoning, improved instruction adherence, and a more natural conversational style compared to GPT-5. It uses adaptive reasoning...
The Qwen3.5 Series 35B-A3B is a native vision-language model designed with a hybrid architecture that integrates linear attention mechanisms and a sparse mixture-of-experts model, achieving higher inference efficiency. Its overall...
Qwen3.5-9B is a multimodal foundation model from the Qwen3.5 family, designed to deliver strong reasoning, coding, and visual understanding in an efficient 9B-parameter architecture. It uses a unified vision-language design...
No provider description is available for this model yet.
| Model | Creator | Inputs | Context | Input | Output | Released | Compare |
|---|---|---|---|---|---|---|---|
| ai21.jamba-instruct-v1:0bedrock/ai21.jamba-instruct-v1:0 | 70K | $0.5 | $0.7 | — | |||
| us.writer.palmyra-x4-v1:0bedrock_converse/us.writer.palmyra-x4-v1:0 | 128K | $2.5 | $10 | — | |||
| deepseek-ai/DeepSeek-V4-Pronebius/deepseek-ai/deepseek-v4-pro | 1.04858M | $1.75 | $3.5 | — | |||
| writer.palmyra-x4-v1:0bedrock_converse/writer.palmyra-x4-v1:0 | 128K | $2.5 | $10 | — | |||
| writer.palmyra-x5-v1:0bedrock_converse/writer.palmyra-x5-v1:0 | 1M | $0.6 | $6 | — | |||
| amazon.nova-lite-v1:0bedrock_converse/amazon.nova-lite-v1:0 | 300K | $0.06 | $0.24 | — | |||
| OpenAI: GPT-5.2 Chatopenai/gpt-5.2-chat | 128K | $1.75 | $14 | — | |||
| Z.ai: GLM 4.7z-ai/glm-4.7 | 202.752K | $0.4 | $1.75 | — | |||
| amazon.nova-2-pro-preview-20251202-v1:0bedrock_converse/amazon.nova-2-pro-preview-20251202-v1:0 | 1M | $2.188 | $17.5 | — | |||
| OpenAI: GPT-5.6 Luna Pro (batch)openai/gpt-5.6-luna-pro:batch | 1.05M | $0.1 | $0.6 | — | |||
| deepseek-ai/DeepSeek-V4-Flash-0731nebius/deepseek-ai/deepseek-v4-flash-0731 | 1.024M | $0.14 | $0.28 | — | |||
| apac.amazon.nova-2-pro-preview-20251202-v1:0bedrock_converse/apac.amazon.nova-2-pro-preview-20251202-v1:0 | 1M | $2.188 | $17.5 | — | |||
| eu.amazon.nova-2-lite-v1:0bedrock_converse/eu.amazon.nova-2-lite-v1:0 | 1M | $0.33 | $2.75 | — | |||
| eu.amazon.nova-2-pro-preview-20251202-v1:0bedrock_converse/eu.amazon.nova-2-pro-preview-20251202-v1:0 | 1M | $2.188 | $17.5 | — | |||
| us.amazon.nova-2-lite-v1:0bedrock_converse/us.amazon.nova-2-lite-v1:0 | 1M | $0.33 | $2.75 | — | |||
| us.amazon.nova-2-pro-preview-20251202-v1:0bedrock_converse/us.amazon.nova-2-pro-preview-20251202-v1:0 | 1M | $2.188 | $17.5 | — | |||
| amazon.nova-micro-v1:0bedrock_converse/amazon.nova-micro-v1:0 | 128K | $0.035 | $0.14 | — | |||
| amazon.nova-pro-v1:0bedrock_converse/amazon.nova-pro-v1:0 | 300K | $0.8 | $3.2 | — | |||
| qwen3-coder-flash-2025-07-28qwen_ai_platform/qwen3-coder-flash-2025-07-28 | 997.952K | — | — | — | |||
| inclusionAI: Ling 3.0 Tiny (free)inclusionai/ling-3.0-tiny:free | 262.144K | Free | Free | — | |||
| Inference.net: Schematron V2 Turboinference-net/schematron-v2-turbo | 128K | $0.03 | $0.15 | — | |||
| MiniMax: MiniMax M2.1minimax/minimax-m2.1 | 204.8K | $0.3 | $1.2 | — | |||
| OpenAI: GPT Audioopenai/gpt-audio | 128K | $2.5 | $10 | — | |||
| openai/gpt-4o-minigmi/openai/gpt-4o-mini | 131.072K | $0.15 | $0.6 | — | |||
| Google: Gemini 3.5 Flash (batch)google/gemini-3.5-flash:batch | 1.04858M | $0.75 | $4.5 | — | |||
| amazon.titan-text-express-v1bedrock/amazon.titan-text-express-v1 | 42K | $1.3 | $1.7 | — | |||
| MiniMax: MiniMax M2-herminimax/minimax-m2-her | 65.536K | $0.3 | $1.2 | — | |||
| amazon.titan-text-premier-v1:0bedrock/amazon.titan-text-premier-v1:0 | 42K | $0.5 | $1.5 | — | |||
| Upstage: Solar Pro 3upstage/solar-pro-3 | 131.072K | $0.15 | $0.6 | — | |||
| deepseek-ai/DeepSeek-V3.2gmi/deepseek-ai/deepseek-v3.2 | 163.84K | $0.28 | $0.4 | — | |||
| deepseek-ai/DeepSeek-V4-Flashnebius/deepseek-ai/deepseek-v4-flash | 1.04858M | $0.14 | $0.28 | — | |||
| us-gov-east-1/amazon.nova-pro-v1:0bedrock/us-gov-east-1/amazon.nova-pro-v1:0 | 300K | $0.96 | $3.84 | — | |||
| qwen3-coder-flashqwen_ai_platform/qwen3-coder-flash | 997.952K | — | — | — | |||
| Qwen: Qwen3.5 Plus 2026-02-15qwen/qwen3.5-plus-02-15 | 1M | $0.26 | $1.56 | — | |||
| AionLabs: Aion-2.0aion-labs/aion-2.0 | 131.072K | $0.8 | $1.6 | — | |||
| OpenAI: o1-pro (batch)openai/o1-pro:batch | 200K | $75 | $300 | — | |||
| databricks-qwen35-122b-a10bdatabricks/databricks-qwen35-122b-a10b | 262.144K | $0.22 | $2.2 | — | |||
| anthropic.claude-haiku-4-5@20251001bedrock_converse/anthropic.claude-haiku-4-5@20251001 | 200K | $1 | $5 | — | |||
| Qwen: Qwen3.5-122B-A10Bqwen/qwen3.5-122b-a10b | 262.144K | $0.26 | $2.08 | — | |||
| Qwen: Qwen3.5-27Bqwen/qwen3.5-27b | 262.144K | $0.195 | $1.56 | — | |||
| inclusionAI: Ling 3.0 Flash VLinclusionai/ling-3.0-flash-vl | 131.072K | $0.06 | $0.18 | — | |||
| Meta: Muse Spark 1.2 Contributormeta/muse-spark-1.2-contributor | 1.04858M | $0.1 | $0.2 | — | |||
| us-gov-east-1/amazon.titan-text-lite-v1bedrock/us-gov-east-1/amazon.titan-text-lite-v1 | 42K | $0.3 | $0.4 | — | |||
| databricks-inklingdatabricks/databricks-inkling | 1M | $1 | $4.05 | — | |||
| us-gov-east-1/amazon.titan-text-premier-v1:0bedrock/us-gov-east-1/amazon.titan-text-premier-v1:0 | 42K | $0.5 | $1.5 | — | |||
| anthropic.claude-3-5-sonnet-20241022-v2:0bedrock/anthropic.claude-3-5-sonnet-20241022-v2:0 | 1M | $3 | $15 | — | |||
| OpenAI: GPT-5.1 (batch)openai/gpt-5.1:batch | 400K | $0.625 | $5 | — | |||
| Qwen: Qwen3.5-35B-A3Bqwen/qwen3.5-35b-a3b | 256K | $0.312 | $1.25 | — | |||
| Qwen: Qwen3.5-9Bqwen/qwen3.5-9b | 262.144K | $0.1 | $0.15 | — | |||
| databricks-grok-4-6databricks/databricks-grok-4-6 | 500K | $2.5 | $7.5 | — |