Open multimodal Qwen MoE for local agents that need vision, audio, and code
19 providers
Every model in the catalog with source-linked pricing, context limits, provider availability, and published benchmark results.
Open multimodal Qwen MoE for local agents that need vision, audio, and code
Speech transcription model for accurate audio-to-text and captioning workflows
Open Whisper checkpoint for robust multilingual transcription and captioning
| Model | Creator | Inputs | Context | Input | Output | Released | Compare |
|---|---|---|---|---|---|---|---|
| Qwen3.6 35B-A3Balibaba/qwen3.6-35b-a3b | 262.144K | $0.248 | $1.485 | 2026-04-17 | |||
| Whisper Large v3 Turboopenai/whisper-large-v3-turbo | 448 | $0.002 | $0.002 | 2024-10-01 | |||
| Whisper 3 Largeopenai/whisper-large-v3 | 448 | $0.002 | $0.002 | 2024-10-01 |