Qwen3-VL-235B-A22B Instruct is an open-weight multimodal model that unifies strong text generation with visual understanding across images and video. The Instruct model targets general vision-language use (VQA, document parsing, chart/table...
Models
Every model in the catalog with source-linked pricing, context limits, provider availability, and published benchmark results.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
NVIDIA Nemotron 3 Ultra is an open frontier-reasoning and orchestration model from NVIDIA, with 55B active parameters out of 550B total (MoE). Built on a hybrid Transformer-Mamba mixture-of-experts architecture, it...
Virtuoso‑Large is Arcee's top‑tier general‑purpose LLM at 72 B parameters, tuned to tackle cross‑domain reasoning, creative writing and enterprise QA. Unlike many 70 B peers, it retains the 128 k...
Claude Opus 5 is Anthropic’s flagship model for demanding reasoning, coding, and long-horizon agentic work. It is particularly strong at end-to-end software tasks, code review and bug finding, visual analysis...
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
NVIDIA Nemotron 3.5 Lightning is an open mixture-of-experts model from NVIDIA, with 3B active parameters out of 30B total. It is suited for high-throughput agentic workloads and specialized tasks that...
MiniMax-M3 is a multimodal foundation model from MiniMax. It supports text, image, and video inputs with text output, a 1M-token context window, and is suited for long-horizon agentic work, coding,...
Qwen3-235B-A22B-Thinking-2507 is a high-performance, open-weight Mixture-of-Experts (MoE) language model optimized for complex reasoning tasks. It activates 22B of its 235B parameters per forward pass and natively supports up to 262,144...
Qwen-Plus, based on the Qwen2.5 foundation model, is a 131K context model with a balanced performance, speed, and cost combination.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
This model always redirects to the latest model in the GPT Terra family.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
Muse Spark 1.3 Contributor is the cost-efficient contributor tier of Meta’s multimodal reasoning model for experimentation, learning, and early-stage agentic, multi-agent, and coding workflows. It is designed to track information...
No provider description is available for this model yet.
No provider description is available for this model yet.
Claude Opus 4.5 is Anthropic’s frontier reasoning model optimized for complex software engineering, agentic workflows, and long-horizon computer use. It offers strong multimodal capabilities, competitive performance across real-world coding and...
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
No provider description is available for this model yet.
| Model | Creator | Inputs | Context | Input | Output | Released | Compare |
|---|---|---|---|---|---|---|---|
| Qwen: Qwen3 VL 235B A22B Instructqwen/qwen3-vl-235b-a22b-instruct | 131.072K | $0.21 | $1.9 | — | |||
| deepseek-ai/DeepSeek-V4-Prodeepinfra/deepseek-ai/deepseek-v4-pro | 1.04858M | $1.3 | $2.6 | — | |||
| nvidia/NVIDIA-Nemotron-3-Super-120B-A12Bdeepinfra/nvidia/nvidia-nemotron-3-super-120b-a12b | 262.144K | $0.085 | $0.4 | — | |||
| zai-org/GLM-5.2deepinfra/zai-org/glm-5.2 | 1.04858M | $0.75 | $2.4 | — | |||
| moonshotai/Kimi-K3deepinfra/moonshotai/kimi-k3 | 1.04858M | $2.85 | $14.25 | — | |||
| anthropic/claude-opus-4-7deepinfra/anthropic/claude-opus-4-7 | 1M | $5 | $25 | — | |||
| Qwen/Qwen3.6-27Bdeepinfra/qwen/qwen3.6-27b | 262.144K | $0.32 | $3.2 | — | |||
| google/gemma-4-26B-A4B-itdeepinfra/google/gemma-4-26b-a4b-it | 262.144K | $0.07 | $0.34 | — | |||
| google/gemini-3.1-prodeepinfra/google/gemini-3.1-pro | 1M | $2 | $12 | — | |||
| XiaomiMiMo/MiMo-V2.5-Prodeepinfra/xiaomimimo/mimo-v2.5-pro | 1.04858M | $1 | $3 | — | |||
| anthropic/claude-haiku-4-5deepinfra/anthropic/claude-haiku-4-5 | 200K | $1 | $5 | — | |||
| deepseek-ai/DeepSeek-V4-Flashdeepinfra/deepseek-ai/deepseek-v4-flash | 1.04858M | $0.09 | $0.18 | — | |||
| openai/gpt-oss-120b-Ultradeepinfra/openai/gpt-oss-120b-ultra | 131.072K | $0.2 | $0.95 | — | |||
| Qwen/Qwen3.5-9Bdeepinfra/qwen/qwen3.5-9b | 262.144K | $0.1 | $0.15 | — | |||
| MiniMaxAI/MiniMax-M2.7-Turbodeepinfra/minimaxai/minimax-m2.7-turbo | 196.608K | $0.38 | $1.7 | — | |||
| zai-org/GLM-4.7deepinfra/zai-org/glm-4.7 | 202.752K | $0.4 | $1.75 | — | |||
| google/gemma-4-31B-itdeepinfra/google/gemma-4-31b-it | 262.144K | $0.13 | $0.38 | — | |||
| NVIDIA: Nemotron 3 Ultra (batch)nvidia/nemotron-3-ultra-550b-a55b:batch | 512.288K | $0.6 | $3.6 | — | |||
| Arcee AI: Virtuoso Largearcee-ai/virtuoso-large | 131.072K | $0.75 | $1.2 | — | |||
| Claude Opus 5 (batch)anthropic/claude-opus-5:batch | 1M | $2.5 | $12.5 | — | |||
| anthropic.claude-fable-5-1bedrock_converse/anthropic.claude-fable-5-1 | 1M | $10 | $50 | — | |||
| global.anthropic.claude-fable-5-1bedrock_converse/global.anthropic.claude-fable-5-1 | 1M | $10 | $50 | — | |||
| us.anthropic.claude-fable-5-1bedrock_converse/us.anthropic.claude-fable-5-1 | 1M | $11 | $55 | — | |||
| NVIDIA: Nemotron 3.5 Lightning (free)nvidia/nemotron-3.5-lightning:free | 1M | Free | Free | — | |||
| MiniMax: MiniMax M3minimax/minimax-m3 | 524.288K | $0.3 | $1.2 | — | |||
| Qwen: Qwen3 235B A22B Thinking 2507qwen/qwen3-235b-a22b-thinking-2507 | 131.072K | $0.23 | $2.3 | — | |||
| Qwen: Qwen-Plusqwen/qwen-plus | 1M | $0.26 | $0.78 | — | |||
| eu.anthropic.claude-fable-5-1bedrock_converse/eu.anthropic.claude-fable-5-1 | 1M | $11 | $55 | — | |||
| claude-fable-5-1azure_ai/claude-fable-5-1 | 1M | $10 | $50 | — | |||
| deepseek-v4-flash-0731azure_ai/deepseek-v4-flash-0731 | 1M | $0.19 | $0.51 | — | |||
| deepseek-v4-flashqwencloud/deepseek-v4-flash | 1M | $0.2 | $0.4 | — | |||
| deepseek-v4-flash-0731qwencloud/deepseek-v4-flash-0731 | 1M | $0.2 | $0.4 | — | |||
| kimi-k3moonshot/kimi-k3 | 1.04858M | $3 | $15 | — | |||
| deepseek-v4-proqwencloud/deepseek-v4-pro | 1M | $2.4 | $4.8 | — | |||
| glm-5.1qwencloud/glm-5.1 | 202.745K | $1.4 | $4.4 | — | |||
| glm-5.2qwencloud/glm-5.2 | 1.04858M | $1.4 | $4.4 | — | |||
| OpenAI: GPT Terra Latest~openai/gpt-terra-latest | 1.05M | $2 | $12 | — | |||
| kimi-k2.7-codeqwencloud/kimi-k2.7-code | 229.376K | $0.95 | $4 | — | |||
| qwen-coderqwencloud/qwen-coder | 1M | $0.3 | $1.5 | — | |||
| qwen-flashqwencloud/qwen-flash | 997.952K | — | — | — | |||
| Meta: Muse Spark 1.3 Contributormeta/muse-spark-1.3-contributor | 1.04858M | $0.1 | $0.2 | — | |||
| qwen-flash-2025-07-28qwencloud/qwen-flash-2025-07-28 | 997.952K | — | — | — | |||
| qwen-maxqwencloud/qwen-max | 30.72K | $1.6 | $6.4 | — | |||
| Anthropic: Claude Opus 4.5 (batch)anthropic/claude-opus-4.5:batch | 200K | $2.5 | $12.5 | — | |||
| qwen-plusqwencloud/qwen-plus | 129.024K | $0.4 | $1.2 | — | |||
| qwen-plus-2025-01-25qwencloud/qwen-plus-2025-01-25 | 129.024K | $0.4 | $1.2 | — | |||
| qwen-plus-2025-04-28qwencloud/qwen-plus-2025-04-28 | 129.024K | $0.4 | $1.2 | — | |||
| qwen-plus-2025-07-14qwencloud/qwen-plus-2025-07-14 | 129.024K | $0.4 | $1.2 | — | |||
| qwen-plus-2025-07-28qwencloud/qwen-plus-2025-07-28 | 997.952K | — | — | — | |||
| qwen-plus-2025-09-11qwencloud/qwen-plus-2025-09-11 | 997.952K | — | — | — |