No provider description is available for this model yet.
Models
Every model in the catalog with source-linked pricing, context limits, provider availability, and published benchmark results.
No provider description is available for this model yet.
Euryale L3.1 70B v2.2 is a model focused on creative roleplay from [Sao10k](https://ko-fi.com/sao10k). It is the successor of [Euryale L3 70B v2.1](/models/sao10k/l3-euryale-70b).
No provider description is available for this model yet.
Qwen3.5 Plus (April 2026) is a large-scale multimodal language model from Alibaba. It accepts text, image, and video input and produces text output, with a 1M token context window. This...
This model always redirects to the latest model in the GPT Luna family.
Amazon Nova Micro 1.0 is a text-only model that delivers the lowest latency responses in the Amazon Nova family of models at a very low cost. With a context length...
No provider description is available for this model yet.
Amazon Nova Lite 1.0 is a very low-cost multimodal model from Amazon that focused on fast processing of image, video, and text inputs to generate text output. Amazon Nova Lite...
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
The Meta Llama 3.3 multilingual large language model (LLM) is a pretrained and instruction tuned generative model in 70B (text in/text out). The Llama 3.3 instruction tuned text only model...
GPT-5.4 is OpenAI’s latest frontier model, unifying the Codex and GPT lines into a single system. It features a 1M+ token context window (922K input, 128K output) with support for...
No provider description is available for this model yet.
Euryale L3.3 70B is a model focused on creative roleplay from [Sao10k](https://ko-fi.com/sao10k). It is the successor of [Euryale L3 70B v2.2](/models/sao10k/l3-euryale-70b).
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
MiniMax-01 is a combines MiniMax-Text-01 for text generation and MiniMax-VL-01 for image understanding. It has 456 billion parameters, with 45.9 billion parameters activated per inference, and can handle a context...
No provider description is available for this model yet.
No provider description is available for this model yet.
OpenAI o3-mini-high is the same model as [o3-mini](/openai/o3-mini) with reasoning_effort set to high. o3-mini is a cost-efficient language model optimized for STEM reasoning tasks, particularly excelling in science, mathematics, and...
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
Llama 4 Maverick 17B Instruct (128E) is a high-capacity multimodal language model from Meta, built on a mixture-of-experts (MoE) architecture with 128 experts and 17 billion active parameters per forward...
No provider description is available for this model yet.
The latest and strongest model family from OpenAI, o1 is designed to spend more time thinking before responding. The o1 model series is trained with large-scale reinforcement learning to reason...
Gemini 2.5 Pro is Google’s state-of-the-art AI model designed for advanced reasoning, coding, mathematics, and scientific tasks. It employs “thinking” capabilities, enabling it to reason through responses with enhanced accuracy...
No provider description is available for this model yet.
No provider description is available for this model yet.
This model always redirects to the latest model in the GPT Astra family.
Morph's high-accuracy apply model for complex code edits. ~4,500 tokens/sec with 98% accuracy for precise code transformations. The model requires the prompt to be in the following format: <instruction>{instruction}</instruction> <code>{initial_code}</code>...
No provider description is available for this model yet.
Qwen3-Next-80B-A3B-Thinking is a reasoning-first chat model in the Qwen3-Next line that outputs structured “thinking” traces by default. It’s designed for hard multi-step problems; math proofs, code synthesis/debugging, logic, and agentic...
Hunyuan-A13B is a 13B active parameter Mixture-of-Experts (MoE) language model developed by Tencent, with a total parameter count of 80B and support for reasoning via Chain-of-Thought. It offers competitive benchmark...
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
Venice Uncensored Dolphin Mistral 24B Venice Edition is a fine-tuned variant of Mistral-Small-24B-Instruct-2501, developed by dphn.ai in collaboration with Venice.ai. This model is designed as an “uncensored” instruct-tuned LLM, preserving...
Qwen2.5-VL is proficient in recognizing common objects such as flowers, birds, fish, and insects. It is also highly capable of analyzing texts, charts, icons, graphics, and layouts within images.
GLM-4.5 is our latest flagship foundation model, purpose-built for agent-based applications. It leverages a Mixture-of-Experts (MoE) architecture and supports a context length of up to 128k tokens. GLM-4.5 delivers significantly...
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
gpt-oss-20b is an open-weight 21B parameter model released by OpenAI under the Apache 2.0 license. It uses a Mixture-of-Experts (MoE) architecture with 3.6B active parameters per forward pass, optimized for...
No provider description is available for this model yet.
| Model | Creator | Inputs | Context | Input | Output | Released | Compare |
|---|---|---|---|---|---|---|---|
| qwen/qwen3.8-flashopenrouter/qwen/qwen3.8-flash | 1M | $0.15 | $0.47 | — | |||
| ai21.jamba-1-5-mini-v1:0bedrock/ai21.jamba-1-5-mini-v1:0 | 256K | $0.2 | $0.4 | — | |||
| Sao10K: Llama 3.1 Euryale 70B v2.2sao10k/l3.1-euryale-70b | 131.072K | $0.85 | $0.85 | — | |||
| ai21.jamba-1-5-large-v1:0bedrock/ai21.jamba-1-5-large-v1:0 | 256K | $2 | $8 | — | |||
| Qwen: Qwen3.5 Plus 2026-04-20qwen/qwen3.5-plus-20260420 | 1M | $0.3 | $1.8 | — | |||
| OpenAI: GPT Luna Latest~openai/gpt-luna-latest | 1.05M | $0.2 | $1.2 | — | |||
| Amazon: Nova Micro 1.0amazon/nova-micro-v1 | 128K | $0.035 | $0.14 | — | |||
| us-east-2/qwen.qwen3-coder-nextbedrock/us-east-2/qwen.qwen3-coder-next | 262.144K | $0.5 | $1.2 | — | |||
| Amazon: Nova Lite 1.0amazon/nova-lite-v1 | 300K | $0.06 | $0.24 | — | |||
| google/gemma-3-27b-itopenrouter/google/gemma-3-27b-it | 131.072K | $0.08 | $0.45 | — | |||
| qwen/qwen3.6-35b-a3bopenrouter/qwen/qwen3.6-35b-a3b | 262.144K | $0.1 | $0.9 | — | |||
| us-east-2/moonshotai.kimi-k2.5bedrock/us-east-2/moonshotai.kimi-k2.5 | 262.144K | $0.6 | $3 | — | |||
| Meta: Llama 3.3 70B Instructmeta-llama/llama-3.3-70b-instruct | 131.072K | $0.1 | $0.32 | — | |||
| OpenAI: GPT-5.4 (batch)openai/gpt-5.4:batch | 1.05M | $1.25 | $7.5 | — | |||
| us-east-2/moonshotai.kimi-k2-thinkingbedrock/us-east-2/moonshotai.kimi-k2-thinking | 262.144K | $0.6 | $2.5 | — | |||
| Sao10K: Llama 3.3 Euryale 70Bsao10k/l3.3-euryale-70b | 131.072K | $0.65 | $0.75 | — | |||
| google/gemma-3-12b-itopenrouter/google/gemma-3-12b-it | 131.072K | $0.05 | $0.15 | — | |||
| us-east-2/minimax.minimax-m2.5bedrock/us-east-2/minimax.minimax-m2.5 | 1M | $0.3 | $1.2 | — | |||
| gpt-5-minigithub_copilot/gpt-5-mini | 128K | — | — | — | |||
| MiniMax: MiniMax-01minimax/minimax-01 | 1.00019M | $0.2 | $1.1 | — | |||
| gpt-5github_copilot/gpt-5 | 128K | — | — | — | |||
| us-east-2/minimax.minimax-m2.1bedrock/us-east-2/minimax.minimax-m2.1 | 196K | $0.3 | $1.2 | — | |||
| OpenAI: o3 Mini Highopenai/o3-mini-high | 200K | $1.1 | $4.4 | — | |||
| google/gemma-3-4b-itopenrouter/google/gemma-3-4b-it | 131.072K | $0.05 | $0.1 | — | |||
| qwen/qwen3.6-flashopenrouter/qwen/qwen3.6-flash | 1M | $0.188 | $1.125 | — | |||
| openai/gpt-6-astraopenrouter/openai/gpt-6-astra | 1.05M | $10 | $50 | — | |||
| openai/gpt-5.3-codexopenrouter/openai/gpt-5.3-codex | 272K | $1.75 | $14 | — | |||
| anthropic/claude-sonnet-5openrouter/anthropic/claude-sonnet-5 | 1M | $2 | $10 | — | |||
| Meta: Llama 4 Maverickmeta-llama/llama-4-maverick | 128K | $0.2 | $0.696 | — | |||
| llama3.3-70bsnowflake/llama3.3-70b | 128K | $0.72 | $0.72 | — | |||
| OpenAI: o1 (batch)openai/o1:batch | 200K | $7.5 | $30 | — | |||
| Google: Gemini 2.5 Pro Preview 06-05google/gemini-2.5-pro-preview | 1.04858M | $1.25 | $10 | — | |||
| openai/o1-proopenrouter/openai/o1-pro | 200K | $150 | $600 | — | |||
| llama3.2-3bsnowflake/llama3.2-3b | 128K | — | — | — | |||
| OpenAI: GPT Astra Latest~openai/gpt-astra-latest | 1.05M | $10 | $50 | — | |||
| Morph: Morph V3 Largemorph/morph-v3-large | 262.144K | $0.9 | $1.9 | — | |||
| us-east-2/deepseek.v3.2bedrock/us-east-2/deepseek.v3.2 | 163.84K | $0.62 | $1.85 | — | |||
| Qwen: Qwen3 Next 80B A3B Thinkingqwen/qwen3-next-80b-a3b-thinking | 262.144K | $0.15 | $1.2 | — | |||
| Tencent: Hunyuan A13B Instructtencent/hunyuan-a13b-instruct | 131.072K | $0.14 | $0.57 | — | |||
| meta-llama/llama-4-scoutopenrouter/meta-llama/llama-4-scout | 1.31072M | $0.1 | $0.3 | — | |||
| qwen/qwen3.5-plus-20260420openrouter/qwen/qwen3.5-plus-20260420 | 1M | $0.3 | $1.8 | — | |||
| labs-leanstral-1-5mistral/labs-leanstral-1-5 | 262.144K | — | — | — | |||
| Venice: Uncensoredcognitivecomputations/dolphin-mistral-24b-venice-edition | 128K | $0.2 | $0.9 | — | |||
| Qwen: Qwen2.5 VL 72B Instructqwen/qwen2.5-vl-72b-instruct | 128K | $0.8 | $1 | — | |||
| Z.ai: GLM 4.5z-ai/glm-4.5 | 131.072K | $0.6 | $2.2 | — | |||
| meta-llama/llama-4-maverickopenrouter/meta-llama/llama-4-maverick | 1.04858M | $0.2 | $0.696 | — | |||
| claude-mythos-previewanthropic/claude-mythos-preview | 1M | $10 | $50 | — | |||
| grok-4.20-multi-agent-0309xai/grok-4.20-multi-agent-0309 | 1M | $1.25 | $2.5 | — | |||
| OpenAI: gpt-oss-20b (free)openai/gpt-oss-20b:free | 131.072K | Free | Free | — | |||
| us-east-1/qwen.qwen3-coder-nextbedrock/us-east-1/qwen.qwen3-coder-next | 262.144K | $0.5 | $1.2 | — |