Grok 4.3 is a reasoning model from SpaceXAI. It accepts text and image inputs with text output, and is suited for agentic workflows, instruction-following tasks, and applications requiring high factual...
Models
Every model in the catalog with source-linked pricing, context limits, provider availability, and published benchmark results.
NVIDIA Nemotron 3 Ultra is an open frontier-reasoning and orchestration model from NVIDIA, with 55B active parameters out of 550B total (MoE). Built on a hybrid Transformer-Mamba mixture-of-experts architecture, it...
Largest Nemotron 3 model for maximum open-weight reasoning and agent accuracy
Low-latency Gemini model for high-volume multimodal and agent workloads
Fast Gemini workhorse for multimodal apps where latency and price matter
Gemini 2.5 Flash is Google's state-of-the-art workhorse model, specifically designed for advanced reasoning, coding, mathematics, and scientific tasks. It includes built-in "thinking" capabilities, enabling it to provide responses with greater...
DeepSeek chat model for instruction following, coding, and analysis
| Model | Creator | Score | Inputs | Context | Input | Output | Released | Compare |
|---|---|---|---|---|---|---|---|---|
| SpaceXAI: Grok 4.3 (batch)x-ai/grok-4.3:batch | 1110.0 | 1M | $1 | $2 | — | |||
| NVIDIA: Nemotron 3 Ultra (free)nvidia/nemotron-3-ultra-550b-a55b:free | 1099.0 | 1M | Free | Free | — | |||
| Nemotron 3 Ultra 550B A55Bnvidia/nemotron-3-ultra-550b-a55b | 1099.0 | 1M | $0.5 | $2.5 | 2026-06-04 | |||
| Gemini 3.1 Flash Lite Previewgoogle/gemini-3.1-flash-lite-preview | 1077.0 | 1.04858M | $0.25 | $1.5 | 2026-03-03 | |||
| Gemini 2.5 Flashgoogle/gemini-2.5-flash | 1045.0 | 1.04858M | $0.3 | $2.5 | 2025-06-17 | |||
| Google: Gemini 2.5 Flash (batch)google/gemini-2.5-flash:batch | 1045.0 | 1.04858M | $0.15 | $1.25 | — | |||
| DeepSeek Chatdeepseek/deepseek-chat | 1002.0 | 1M | $0.147 | $0.295 | 2025-12-01 |