Virtuoso‑Large is Arcee's top‑tier general‑purpose LLM at 72 B parameters, tuned to tackle cross‑domain reasoning, creative writing and enterprise QA. Unlike many 70 B peers, it retains the 128 k...
Models
Every model in the catalog with source-linked pricing, context limits, provider availability, and published benchmark results.
NVIDIA Nemotron 3.5 Lightning is an open mixture-of-experts model from NVIDIA, with 3B active parameters out of 30B total. It is suited for high-throughput agentic workloads and specialized tasks that...
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
This model always redirects to the latest model in the GPT Terra family.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
Muse Spark 1.3 Contributor is the cost-efficient contributor tier of Meta’s multimodal reasoning model for experimentation, learning, and early-stage agentic, multi-agent, and coding workflows. It is designed to track information...
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
This model always redirects to the latest GLM model from Z.ai.
Qwen3.8 2.4T A95B is an open-weight sparse mixture-of-experts model from Qwen and the open-weight variant of [Qwen3.8 Max](/qwen/qwen3.8-max), with 95 billion active parameters out of 2.4 trillion total. It is...
No provider description is available for this model yet.
The o-series of models are trained with reinforcement learning to think before they answer and perform complex reasoning. The o3-pro model uses more compute to think harder and provide consistently...
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
gpt-oss-120b is an open-weight, 117B-parameter Mixture-of-Experts (MoE) language model from OpenAI designed for high-reasoning, agentic, and general-purpose production use cases. It activates 5.1B parameters per forward pass and is optimized...
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
| Model | Creator | Inputs | Context | Input | Output | Released | Compare |
|---|---|---|---|---|---|---|---|
| Arcee AI: Virtuoso Largearcee-ai/virtuoso-large | 131.072K | $0.75 | $1.2 | — | |||
| NVIDIA: Nemotron 3.5 Lightning (free)nvidia/nemotron-3.5-lightning:free | 1M | Free | Free | — | |||
| anthropic.claude-fable-5-1bedrock_converse/anthropic.claude-fable-5-1 | 1M | $10 | $50 | — | |||
| global.anthropic.claude-fable-5-1bedrock_converse/global.anthropic.claude-fable-5-1 | 1M | $10 | $50 | — | |||
| us.anthropic.claude-fable-5-1bedrock_converse/us.anthropic.claude-fable-5-1 | 1M | $11 | $55 | — | |||
| eu.anthropic.claude-fable-5-1bedrock_converse/eu.anthropic.claude-fable-5-1 | 1M | $11 | $55 | — | |||
| claude-fable-5-1azure_ai/claude-fable-5-1 | 1M | $10 | $50 | — | |||
| deepseek-v4-flash-0731azure_ai/deepseek-v4-flash-0731 | 1M | $0.19 | $0.51 | — | |||
| deepseek-v4-flashqwencloud/deepseek-v4-flash | 1M | $0.2 | $0.4 | — | |||
| deepseek-v4-flash-0731qwencloud/deepseek-v4-flash-0731 | 1M | $0.2 | $0.4 | — | |||
| kimi-k3moonshot/kimi-k3 | 1.04858M | $3 | $15 | — | |||
| deepseek-v4-proqwencloud/deepseek-v4-pro | 1M | $2.4 | $4.8 | — | |||
| OpenAI: GPT Terra Latest~openai/gpt-terra-latest | 1.05M | $2 | $12 | — | |||
| glm-5.1qwencloud/glm-5.1 | 202.745K | $1.4 | $4.4 | — | |||
| glm-5.2qwencloud/glm-5.2 | 1.04858M | $1.4 | $4.4 | — | |||
| kimi-k2.7-codeqwencloud/kimi-k2.7-code | 229.376K | $0.95 | $4 | — | |||
| qwen-coderqwencloud/qwen-coder | 1M | $0.3 | $1.5 | — | |||
| qwen-flashqwencloud/qwen-flash | 997.952K | — | — | — | |||
| Meta: Muse Spark 1.3 Contributormeta/muse-spark-1.3-contributor | 1.04858M | $0.1 | $0.2 | — | |||
| qwen-flash-2025-07-28qwencloud/qwen-flash-2025-07-28 | 997.952K | — | — | — | |||
| qwen-maxqwencloud/qwen-max | 30.72K | $1.6 | $6.4 | — | |||
| qwen-plusqwencloud/qwen-plus | 129.024K | $0.4 | $1.2 | — | |||
| qwen-plus-2025-01-25qwencloud/qwen-plus-2025-01-25 | 129.024K | $0.4 | $1.2 | — | |||
| qwen-plus-2025-04-28qwencloud/qwen-plus-2025-04-28 | 129.024K | $0.4 | $1.2 | — | |||
| qwen-plus-2025-07-14qwencloud/qwen-plus-2025-07-14 | 129.024K | $0.4 | $1.2 | — | |||
| qwen-plus-2025-07-28qwencloud/qwen-plus-2025-07-28 | 997.952K | — | — | — | |||
| qwen-plus-2025-09-11qwencloud/qwen-plus-2025-09-11 | 997.952K | — | — | — | |||
| qwen-plus-latestqwencloud/qwen-plus-latest | 997.952K | — | — | — | |||
| qwen-turboqwencloud/qwen-turbo | 129.024K | $0.05 | $0.2 | — | |||
| qwen-turbo-2024-11-01qwencloud/qwen-turbo-2024-11-01 | 1M | $0.05 | $0.2 | — | |||
| qwen-turbo-2025-04-28qwencloud/qwen-turbo-2025-04-28 | 1M | $0.05 | $0.2 | — | |||
| qwen-turbo-latestqwencloud/qwen-turbo-latest | 1M | $0.05 | $0.2 | — | |||
| qwen3-30b-a3bqwencloud/qwen3-30b-a3b | 129.024K | — | — | — | |||
| qwen3-coder-flashqwencloud/qwen3-coder-flash | 997.952K | — | — | — | |||
| qwen3-coder-flash-2025-07-28qwencloud/qwen3-coder-flash-2025-07-28 | 997.952K | — | — | — | |||
| Z.ai: GLM Latest~z-ai/glm-latest | 262.144K | $0.877 | $2.97 | — | |||
| Qwen: Qwen3.8 2.4T A95Bqwen/qwen3.8-2.4t-a95b | 1M | $2 | $6 | — | |||
| qwen3-coder-plusqwencloud/qwen3-coder-plus | 997.952K | — | — | — | |||
| OpenAI: o3 Pro (batch)openai/o3-pro:batch | 200K | $10 | $40 | — | |||
| qwen3-coder-plus-2025-07-22qwencloud/qwen3-coder-plus-2025-07-22 | 997.952K | — | — | — | |||
| qwen3-max-previewqwencloud/qwen3-max-preview | 258.048K | — | — | — | |||
| qwen3-maxqwencloud/qwen3-max | 258.048K | — | — | — | |||
| OpenAI: gpt-oss-120b (batch)openai/gpt-oss-120b:batch | 131.072K | $0.15 | $0.6 | — | |||
| qwen3-max-2026-01-23qwencloud/qwen3-max-2026-01-23 | 258.048K | — | — | — | |||
| qwen3-next-80b-a3b-instructqwencloud/qwen3-next-80b-a3b-instruct | 262.144K | $0.15 | $1.2 | — | |||
| qwen3-next-80b-a3b-thinkingqwencloud/qwen3-next-80b-a3b-thinking | 262.144K | $0.15 | $1.2 | — | |||
| qwen3-vl-235b-a22b-instructqwencloud/qwen3-vl-235b-a22b-instruct | 131.072K | $0.4 | $1.6 | — | |||
| qwen3-vl-235b-a22b-thinkingqwencloud/qwen3-vl-235b-a22b-thinking | 131.072K | $0.4 | $4 | — | |||
| qwen3-vl-32b-instructqwencloud/qwen3-vl-32b-instruct | 131.072K | $0.16 | $0.64 | — | |||
| qwen3-vl-32b-thinkingqwencloud/qwen3-vl-32b-thinking | 131.072K | $0.16 | $2.87 | — |