MiniMax multimodal model for long-context coding, perception, and agent planning
Model catalog
Models
Every model in the catalog with source-linked pricing, context limits, provider availability, and published benchmark results.
37 providers
Qwen vision-language model for visual reasoning, documents, and agent tasks
24 providers
Step 3.7 Flash is StepFun's latest high-efficiency multimodal Mixture-of-Experts model. It pairs a 196B-parameter language backbone with a vision encoder for native image and video understanding, activating roughly 11B parameters...
17 providers
Large open Qwen multimodal MoE for visual agents and long technical tasks
18 providers
Open multimodal Qwen MoE for local agents that need vision, audio, and code
19 providers
Qwen vision-language model for visual reasoning, documents, and agent tasks
14 providers
Qwen vision-language model for visual reasoning, documents, and agent tasks
16 providers
| Model | Creator | Score | Inputs | Context | Input | Output | Released | Compare |
|---|---|---|---|---|---|---|---|---|
| MiniMax-M3minimax/MiniMax-M3 | 80.5 | 1.04858M | $0.3 | $1.2 | 2026-06-01 | |||
| Qwen3.6 27Balibaba/qwen3.6-27b | 77.2 | 262.144K | $0.6 | $3.6 | 2026-04-22 | |||
| StepFun: Step 3.7 Flashstepfun/step-3.7-flash | 76.5 | 256K | $0.2 | $1.15 | 2026-05-29 | |||
| Qwen3.5 397B-A17Balibaba/qwen3.5-397b-a17b | 76.4 | 262.144K | $0.6 | $3.6 | 2026-02-15 | |||
| Qwen3.6 35B-A3Balibaba/qwen3.6-35b-a3b | 73.4 | 262.144K | $0.248 | $1.485 | 2026-04-17 | |||
| Qwen3.5 27Balibaba/qwen3.5-27b | 72.4 | 262.144K | $0.3 | $2.4 | 2026-02-23 | |||
| Qwen3.5 122B-A10Balibaba/qwen3.5-122b-a10b | 72.0 | 262.144K | $0.4 | $3.2 | 2026-02-23 |