For tasks that demand low latency, GPT‑4.1 nano is the fastest and cheapest model in the GPT-4.1 series. It delivers exceptional performance at a small size with its 1 million...
Models
Every model in the catalog with source-linked pricing, context limits, provider availability, and published benchmark results.
Llama 4 Maverick 17B Instruct (128E) is a high-capacity multimodal language model from Meta, built on a mixture-of-experts (MoE) architecture with 128 experts and 17 billion active parameters per forward...
Omni-era GPT for multimodal chat, practical coding, and general assistants
GPT-4o ("o" for "omni") is OpenAI's latest AI model, supporting both text and image inputs with text outputs. It maintains the intelligence level of [GPT-4 Turbo](/models/openai/gpt-4-turbo) while being twice as...
Fast o-series model for compact reasoning, coding, and tool use
OpenAI o4-mini is a compact reasoning model in the o-series, optimized for fast, cost-efficient performance while retaining strong multimodal and agentic capabilities. It supports tool use and demonstrates competitive reasoning...
Long-lived GPT workhorse for coding, instruction following, and production apps
GPT-4.1 is a flagship large language model optimized for advanced instruction following, real-world software engineering, and long-context reasoning. It supports a 1 million token context window and outperforms GPT-4o and...
Affordable GPT-4.1 lane for fast coding help and structured extraction
GPT-4.1 Mini is a mid-sized model delivering performance competitive with GPT-4o at substantially lower latency and cost. It retains a 1 million token context window and scores 45.1% on hard...
| Model | Creator | Score | Inputs | Context | Input | Output | Released | Compare |
|---|---|---|---|---|---|---|---|---|
| OpenAI: GPT-4.1 Nano (batch)openai/gpt-4.1-nano:batch | 953.0 | 1.04758M | $0.05 | $0.2 | — | |||
| Meta: Llama 4 Maverickmeta-llama/llama-4-maverick | 929.0 | 128K | $0.2 | $0.696 | — | |||
| GPT-4oopenai/gpt-4o | 897.0 | 128K | $2.5 | $10 | 2024-05-13 | |||
| OpenAI: GPT-4o (batch)openai/gpt-4o:batch | 897.0 | 128K | $1.25 | $5 | — | |||
| o4-miniopenai/o4-mini | 882.0 | 200K | $1.1 | $4.4 | 2025-04-16 | |||
| OpenAI: o4 Mini (batch)openai/o4-mini:batch | 882.0 | 200K | $0.55 | $2.2 | — | |||
| GPT-4.1openai/gpt-4.1 | 878.0 | 1.04758M | $2 | $8 | 2025-04-14 | |||
| OpenAI: GPT-4.1 (batch)openai/gpt-4.1:batch | 878.0 | 1.04758M | $1 | $4 | — | |||
| GPT-4.1 miniopenai/gpt-4.1-mini | 865.0 | 1.04758M | $0.4 | $1.6 | 2025-04-14 | |||
| OpenAI: GPT-4.1 Mini (batch)openai/gpt-4.1-mini:batch | 865.0 | 1.04758M | $0.2 | $0.8 | — |