Gemini 2.5 Flash is Google's state-of-the-art workhorse model, specifically designed for advanced reasoning, coding, mathematics, and scientific tasks. It includes built-in "thinking" capabilities, enabling it to provide responses with greater...
Models
Every model in the catalog with source-linked pricing, context limits, provider availability, and published benchmark results.
Qwen3.8 2.4T A95B is an open-weight sparse mixture-of-experts model from Qwen and the open-weight variant of [Qwen3.8 Max](/qwen/qwen3.8-max), with 95 billion active parameters out of 2.4 trillion total. It is...
Opus 4.6 is Anthropic’s strongest model for coding and long-running professional tasks. It is built for agents that operate across entire workflows rather than single prompts, making it especially effective...
Virtuoso‑Large is Arcee's top‑tier general‑purpose LLM at 72 B parameters, tuned to tackle cross‑domain reasoning, creative writing and enterprise QA. Unlike many 70 B peers, it retains the 128 k...
GPT-5.2 Pro is OpenAI’s most advanced model, offering major improvements in agentic coding and long context performance over GPT-5 Pro. It is optimized for complex tasks that require step-by-step reasoning,...
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
A balanced model in the Ministral 3 family, Ministral 3 8B is a powerful, efficient tiny language model with vision capabilities.
Claude Opus 4.5 is Anthropic’s frontier reasoning model optimized for complex software engineering, agentic workflows, and long-horizon computer use. It offers strong multimodal capabilities, competitive performance across real-world coding and...
No provider description is available for this model yet.
Grok 4.20 Multi-Agent is a variant of SpaceXAI’s Grok 4.20 designed for collaborative, agent-based workflows. Multiple agents operate in parallel to conduct deep research, coordinate tool use, and synthesize information...
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
This model always redirects to the latest model in the GPT Terra family.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
Muse Spark 1.3 Contributor is the cost-efficient contributor tier of Meta’s multimodal reasoning model for experimentation, learning, and early-stage agentic, multi-agent, and coding workflows. It is designed to track information...
DeepSeek-V3.2-Exp is an experimental large language model released by DeepSeek as an intermediate step between V3.1 and future architectures. It introduces DeepSeek Sparse Attention (DSA), a fine-grained sparse attention mechanism...
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
Qwen3-235B-A22B-Thinking-2507 is a high-performance, open-weight Mixture-of-Experts (MoE) language model optimized for complex reasoning tasks. It activates 22B of its 235B parameters per forward pass and natively supports up to 262,144...
No provider description is available for this model yet.
The o-series of models are trained with reinforcement learning to think before they answer and perform complex reasoning. The o3-pro model uses more compute to think harder and provide consistently...
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
LFM2.5-2.6B is a compact reasoning model from Liquid AI. It is suited for agent workflows, data extraction, RAG, and long-context processing. Liquid advises against using it for agentic coding or...
No provider description is available for this model yet.
| Model | Creator | Inputs | Context | Input | Output | Released | Compare |
|---|---|---|---|---|---|---|---|
| Google: Gemini 2.5 Flash (batch)google/gemini-2.5-flash:batch | 1.04858M | $0.15 | $1.25 | — | |||
| Qwen: Qwen3.8 2.4T A95Bqwen/qwen3.8-2.4t-a95b | 1M | $2 | $6 | — | |||
| Anthropic: Claude Opus 4.6 (batch)anthropic/claude-opus-4.6:batch | 1M | $2.5 | $12.5 | — | |||
| Arcee AI: Virtuoso Largearcee-ai/virtuoso-large | 131.072K | $0.75 | $1.2 | — | |||
| OpenAI: GPT-5.2 Pro (batch)openai/gpt-5.2-pro:batch | 400K | $10.5 | $84 | — | |||
| anthropic.claude-fable-5-1bedrock_converse/anthropic.claude-fable-5-1 | 1M | $10 | $50 | — | |||
| global.anthropic.claude-fable-5-1bedrock_converse/global.anthropic.claude-fable-5-1 | 1M | $10 | $50 | — | |||
| us.anthropic.claude-fable-5-1bedrock_converse/us.anthropic.claude-fable-5-1 | 1M | $11 | $55 | — | |||
| Mistral: Ministral 3 8B 2512 (batch)mistralai/ministral-8b-2512:batch | 262.144K | $0.075 | $0.075 | — | |||
| Anthropic: Claude Opus 4.5 (batch)anthropic/claude-opus-4.5:batch | 200K | $2.5 | $12.5 | — | |||
| eu.anthropic.claude-fable-5-1bedrock_converse/eu.anthropic.claude-fable-5-1 | 1M | $11 | $55 | — | |||
| SpaceXAI: Grok 4.20 Multi-Agentx-ai/grok-4.20-multi-agent | 2M | $1.25 | $2.5 | — | |||
| claude-fable-5-1azure_ai/claude-fable-5-1 | 1M | $10 | $50 | — | |||
| deepseek-v4-flash-0731azure_ai/deepseek-v4-flash-0731 | 1M | $0.19 | $0.51 | — | |||
| deepseek-v4-flashqwencloud/deepseek-v4-flash | 1M | $0.2 | $0.4 | — | |||
| deepseek-v4-flash-0731qwencloud/deepseek-v4-flash-0731 | 1M | $0.2 | $0.4 | — | |||
| kimi-k3moonshot/kimi-k3 | 1.04858M | $3 | $15 | — | |||
| deepseek-v4-proqwencloud/deepseek-v4-pro | 1M | $2.4 | $4.8 | — | |||
| glm-5.1qwencloud/glm-5.1 | 202.745K | $1.4 | $4.4 | — | |||
| glm-5.2qwencloud/glm-5.2 | 1.04858M | $1.4 | $4.4 | — | |||
| OpenAI: GPT Terra Latest~openai/gpt-terra-latest | 1.05M | $2 | $12 | — | |||
| kimi-k2.7-codeqwencloud/kimi-k2.7-code | 229.376K | $0.95 | $4 | — | |||
| qwen-coderqwencloud/qwen-coder | 1M | $0.3 | $1.5 | — | |||
| qwen-flashqwencloud/qwen-flash | 997.952K | — | — | — | |||
| Meta: Muse Spark 1.3 Contributormeta/muse-spark-1.3-contributor | 1.04858M | $0.1 | $0.2 | — | |||
| DeepSeek: DeepSeek V3.2 Expdeepseek/deepseek-v3.2-exp | 163.84K | $0.27 | $0.41 | — | |||
| qwen-flash-2025-07-28qwencloud/qwen-flash-2025-07-28 | 997.952K | — | — | — | |||
| qwen-maxqwencloud/qwen-max | 30.72K | $1.6 | $6.4 | — | |||
| qwen-plusqwencloud/qwen-plus | 129.024K | $0.4 | $1.2 | — | |||
| qwen-plus-2025-01-25qwencloud/qwen-plus-2025-01-25 | 129.024K | $0.4 | $1.2 | — | |||
| qwen-plus-2025-04-28qwencloud/qwen-plus-2025-04-28 | 129.024K | $0.4 | $1.2 | — | |||
| qwen-plus-2025-07-14qwencloud/qwen-plus-2025-07-14 | 129.024K | $0.4 | $1.2 | — | |||
| qwen-plus-2025-07-28qwencloud/qwen-plus-2025-07-28 | 997.952K | — | — | — | |||
| qwen-plus-2025-09-11qwencloud/qwen-plus-2025-09-11 | 997.952K | — | — | — | |||
| qwen-plus-latestqwencloud/qwen-plus-latest | 997.952K | — | — | — | |||
| qwen-turboqwencloud/qwen-turbo | 129.024K | $0.05 | $0.2 | — | |||
| qwen-turbo-2024-11-01qwencloud/qwen-turbo-2024-11-01 | 1M | $0.05 | $0.2 | — | |||
| qwen-turbo-2025-04-28qwencloud/qwen-turbo-2025-04-28 | 1M | $0.05 | $0.2 | — | |||
| qwen-turbo-latestqwencloud/qwen-turbo-latest | 1M | $0.05 | $0.2 | — | |||
| qwen3-30b-a3bqwencloud/qwen3-30b-a3b | 129.024K | — | — | — | |||
| qwen3-coder-flashqwencloud/qwen3-coder-flash | 997.952K | — | — | — | |||
| qwen3-coder-flash-2025-07-28qwencloud/qwen3-coder-flash-2025-07-28 | 997.952K | — | — | — | |||
| Qwen: Qwen3 235B A22B Thinking 2507qwen/qwen3-235b-a22b-thinking-2507 | 131.072K | $0.23 | $2.3 | — | |||
| qwen3-coder-plusqwencloud/qwen3-coder-plus | 997.952K | — | — | — | |||
| OpenAI: o3 Pro (batch)openai/o3-pro:batch | 200K | $10 | $40 | — | |||
| qwen3-coder-plus-2025-07-22qwencloud/qwen3-coder-plus-2025-07-22 | 997.952K | — | — | — | |||
| qwen3-max-previewqwencloud/qwen3-max-preview | 258.048K | — | — | — | |||
| qwen3-maxqwencloud/qwen3-max | 258.048K | — | — | — | |||
| LiquidAI: LFM2.5-2.6B (free)liquid/lfm-2.5-2.6b:free | 65.536K | Free | Free | — | |||
| qwen3-max-2026-01-23qwencloud/qwen3-max-2026-01-23 | 258.048K | — | — | — |