Aion-RP-Llama-3.1-8B ranks the highest in the character evaluation portion of the RPBench-Auto benchmark, a roleplaying-specific variant of Arena-Hard-Auto, where LLMs evaluate each other’s responses. It is a fine-tuned base model...
Models
Every model in the catalog with source-linked pricing, context limits, provider availability, and published benchmark results.
No provider description is available for this model yet.
OpenAI o3-mini-high is the same model as [o3-mini](/openai/o3-mini) with reasoning_effort set to high. o3-mini is a cost-efficient language model optimized for STEM reasoning tasks, particularly excelling in science, mathematics, and...
No provider description is available for this model yet.
Gemini 3.1 Flash Lite is Google’s GA high-efficiency multimodal model optimized for low-latency, high-volume workloads. It supports text, image, video, audio, and PDF inputs, and is designed for lightweight agentic...
No provider description is available for this model yet.
No provider description is available for this model yet.
Qwen3-Next-80B-A3B-Instruct is an instruction-tuned chat model in the Qwen3-Next series optimized for fast, stable responses without “thinking” traces. It targets complex tasks across reasoning, code generation, knowledge QA, and multilingual...
*Ling-3.0-flash* is a *124B-parameter Mixture-of-Experts (MoE) model*, with approximately *5.1B parameters activated per token*. The model is designed with *token efficiency and production-scale agentic inference* as key priorities, enabling developers...
GPT-5.6 Sol is the flagship model in OpenAI's GPT-5.6 series. It is suited for complex reasoning, coding, and agentic workflows, and is particularly strong at command-line and multi-step coding tasks...
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
A balanced model in the Ministral 3 family, Ministral 3 8B is a powerful, efficient tiny language model with vision capabilities.
GPT-5.2 Pro is OpenAI’s most advanced model, offering major improvements in agentic coding and long context performance over GPT-5 Pro. It is optimized for complex tasks that require step-by-step reasoning,...
No provider description is available for this model yet.
No provider description is available for this model yet.
Ox Alpha is a reasoning model designed for coding, sustained agentic work, and production workloads. It is suited for long-horizon software engineering, complex reasoning, and workflows that combine text with...
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
Qwen Plus 0728, based on the Qwen3 foundation model, is a 1 million context hybrid reasoning model with a balanced performance, speed, and cost combination.
Qwen3-VL-32B-Instruct is a large-scale multimodal vision-language model designed for high-precision understanding and reasoning across text, images, and video. With 32 billion parameters, it combines deep visual perception with advanced text...
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
gpt-oss-120b is an open-weight, 117B-parameter Mixture-of-Experts (MoE) language model from OpenAI designed for high-reasoning, agentic, and general-purpose production use cases. It activates 5.1B parameters per forward pass and is optimized...
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
| Model | Creator | Inputs | Context | Input | Output | Released | Compare |
|---|---|---|---|---|---|---|---|
| AionLabs: Aion-RP 1.0 (8B)aion-labs/aion-rp-llama-3.1-8b | 32.768K | $0.8 | $1.6 | — | |||
| zai-glm-4.6cerebras/zai-glm-4.6 | 128K | $2.25 | $2.75 | — | |||
| OpenAI: o3 Mini High (batch)openai/o3-mini-high:batch | 200K | $0.55 | $2.2 | — | |||
| gemini-robotics-er-2-previewgemini/gemini-robotics-er-2-preview | 131.072K | $2 | $10 | — | |||
| Google: Gemini 3.1 Flash Lite (batch)google/gemini-3.1-flash-lite:batch | 1.04858M | $0.125 | $0.75 | — | |||
| gemini-robotics-er-1.6-previewgemini/gemini-robotics-er-1.6-preview | 131.072K | $1 | $5 | — | |||
| invoke/anthropic.claude-3-5-sonnet-20240620-v1:0bedrock/invoke/anthropic.claude-3-5-sonnet-20240620-v1:0 | 200K | $3 | $15 | — | |||
| Qwen: Qwen3 Next 80B A3B Instruct (free)qwen/qwen3-next-80b-a3b-instruct:free | 262.144K | Free | Free | — | |||
| inclusionAI: Ling 3.0 Flashinclusionai/ling-3.0-flash | 262.144K | $0.021 | $0.063 | — | |||
| OpenAI: GPT-5.6 Sol (batch)openai/gpt-5.6-sol:batch | 1.05M | $1 | $5 | — | |||
| sa-east-1/meta.llama3-70b-instruct-v1:0bedrock/sa-east-1/meta.llama3-70b-instruct-v1:0 | 8.192K | $4.45 | $5.88 | — | |||
| sa-east-1/meta.llama3-8b-instruct-v1:0bedrock/sa-east-1/meta.llama3-8b-instruct-v1:0 | 8.192K | $0.5 | $1.01 | — | |||
| sa-east-1/deepseek.v3.2bedrock/sa-east-1/deepseek.v3.2 | 163.84K | $0.74 | $2.22 | — | |||
| llama3ollama/llama3 | 8.192K | — | — | — | |||
| Mistral: Ministral 3 8B 2512mistralai/ministral-8b-2512 | 262.144K | $0.15 | $0.15 | — | |||
| OpenAI: GPT-5.2 Pro (batch)openai/gpt-5.2-pro:batch | 400K | $10.5 | $84 | — | |||
| google/gemma-4-31B-itfriendliai/google/gemma-4-31b-it | 262.144K | $0.14 | $0.4 | — | |||
| zai-org/GLM-5.2friendliai/zai-org/glm-5.2 | 1.04858M | $1.4 | $4.4 | — | |||
| Ox Alphastealth/ox-alpha | 1.04858M | — | — | — | |||
| LGAI-EXAONE/K-EXAONE-2.0-750B-A37Bfriendliai/lgai-exaone/k-exaone-2.0-750b-a37b | 262.144K | $0.6 | $2.4 | — | |||
| deepseek-ai/DeepSeek-V3.2friendliai/deepseek-ai/deepseek-v3.2 | 163.84K | $0.5 | $1.5 | — | |||
| sa-east-1/minimax.minimax-m2.1bedrock/sa-east-1/minimax.minimax-m2.1 | 196K | $0.36 | $1.44 | — | |||
| Qwen: Qwen Plus 0728 (thinking)qwen/qwen-plus-2025-07-28:thinking | 1M | $0.26 | $0.78 | — | |||
| Qwen: Qwen3 VL 32B Instructqwen/qwen3-vl-32b-instruct | 131.072K | $0.104 | $0.416 | — | |||
| MiniMaxAI/MiniMax-M2.5friendliai/minimaxai/minimax-m2.5 | 196.608K | $0.3 | $1.2 | — | |||
| gpt-6-astraazure_ai/gpt-6-astra | 922K | $10 | $50 | — | |||
| sa-east-1/minimax.minimax-m2.5bedrock/sa-east-1/minimax.minimax-m2.5 | 1M | $0.36 | $1.44 | — | |||
| sa-east-1/moonshotai.kimi-k2-thinkingbedrock/sa-east-1/moonshotai.kimi-k2-thinking | 262.144K | $0.73 | $3.03 | — | |||
| gpt-5.6-cyberopenai/gpt-5.6-cyber | 400K | $12.5 | $75 | — | |||
| sa-east-1/moonshotai.kimi-k2.5bedrock/sa-east-1/moonshotai.kimi-k2.5 | 262.144K | $0.72 | $3.6 | — | |||
| sa-east-1/qwen.qwen3-coder-nextbedrock/sa-east-1/qwen.qwen3-coder-next | 262.144K | $0.6 | $1.44 | — | |||
| OpenAI: gpt-oss-120b (free)openai/gpt-oss-120b:free | 131.072K | Free | Free | — | |||
| us-east-1/1-month-commitment/anthropic.claude-instant-v1bedrock/us-east-1/1-month-commitment/anthropic.claude-instant-v1 | 100K | — | — | — | |||
| us-east-1/1-month-commitment/anthropic.claude-v1bedrock/us-east-1/1-month-commitment/anthropic.claude-v1 | 100K | — | — | — | |||
| daybreak-red-latestopenai/daybreak-red-latest | 400K | $12.5 | $75 | — | |||
| daybreak-blue-latestopenai/daybreak-blue-latest | 1.05M | $5 | $30 | — | |||
| us-east-1/1-month-commitment/anthropic.claude-v2:1bedrock/us-east-1/1-month-commitment/anthropic.claude-v2:1 | 100K | — | — | — | |||
| us-east-1/6-month-commitment/anthropic.claude-instant-v1bedrock/us-east-1/6-month-commitment/anthropic.claude-instant-v1 | 100K | — | — | — | |||
| chat-latestopenai/chat-latest | 400K | $5 | $30 | — | |||
| zai-glm-5-2mistral/zai-glm-5-2 | 1.04858M | $1.4 | $4.4 | — | |||
| glm-5-2mistral/glm-5-2 | 1.04858M | $1.4 | $4.4 | — | |||
| anthropic/claude-opus-5openrouter/anthropic/claude-opus-5 | 1M | $5 | $25 | — | |||
| deepseek/deepseek-v4-proopenrouter/deepseek/deepseek-v4-pro | 1.04858M | $1.32 | $3.96 | — | |||
| us-east-1/6-month-commitment/anthropic.claude-v1bedrock/us-east-1/6-month-commitment/anthropic.claude-v1 | 100K | — | — | — | |||
| deepseek/deepseek-v4-pro-0813openrouter/deepseek/deepseek-v4-pro-0813 | 1.04858M | $1.32 | $3.96 | — | |||
| xai.grok-4.6bedrock_mantle/xai.grok-4.6 | 500K | $2.2 | $6.6 | — | |||
| us-east-1/6-month-commitment/anthropic.claude-v2:1bedrock/us-east-1/6-month-commitment/anthropic.claude-v2:1 | 100K | — | — | — | |||
| llama3.1ollama/llama3.1 | 8.192K | — | — | — | |||
| us-gov-west-1/amazon.nova-lite-v1:0bedrock/us-gov-west-1/amazon.nova-lite-v1:0 | 300K | $0.072 | $0.288 | — | |||
| us-gov-west-1/amazon.nova-micro-v1:0bedrock/us-gov-west-1/amazon.nova-micro-v1:0 | 128K | $0.042 | $0.168 | — |