110 models
Ranked by Design Arena: 3d
953.0

For tasks that demand low latency, GPT‑4.1 nano is the fastest and cheapest model in the GPT-4.1 series. It delivers exceptional performance at a small size with its 1 million...

openai/gpt-4.1-nano:batch 1.04758M context $0.05/M input $0.2/M output
929.0

Llama 4 Maverick 17B Instruct (128E) is a high-capacity multimodal language model from Meta, built on a mixture-of-experts (MoE) architecture with 128 experts and 17 billion active parameters per forward...

meta-llama/llama-4-maverick 128K context $0.2/M input $0.696/M output
Tools 897.0

Omni-era GPT for multimodal chat, practical coding, and general assistants

openai/gpt-4o 2024-05-13 128K context $2.5/M input $10/M output
23 providers
897.0

GPT-4o ("o" for "omni") is OpenAI's latest AI model, supporting both text and image inputs with text outputs. It maintains the intelligence level of [GPT-4 Turbo](/models/openai/gpt-4-turbo) while being twice as...

openai/gpt-4o:batch 128K context $1.25/M input $5/M output
Reasoning Tools JSON 882.0

Fast o-series model for compact reasoning, coding, and tool use

openai/o4-mini 2025-04-16 200K context $1.1/M input $4.4/M output
20 providers
882.0

OpenAI o4-mini is a compact reasoning model in the o-series, optimized for fast, cost-efficient performance while retaining strong multimodal and agentic capabilities. It supports tool use and demonstrates competitive reasoning...

openai/o4-mini:batch 200K context $0.55/M input $2.2/M output
Tools JSON 878.0

Long-lived GPT workhorse for coding, instruction following, and production apps

openai/gpt-4.1 2025-04-14 1.04758M context $2/M input $8/M output
28 providers
878.0

GPT-4.1 is a flagship large language model optimized for advanced instruction following, real-world software engineering, and long-context reasoning. It supports a 1 million token context window and outperforms GPT-4o and...

openai/gpt-4.1:batch 1.04758M context $1/M input $4/M output
865.0

Affordable GPT-4.1 lane for fast coding help and structured extraction

openai/gpt-4.1-mini 2025-04-14 1.04758M context $0.4/M input $1.6/M output
25 providers
865.0

GPT-4.1 Mini is a mid-sized model delivering performance competitive with GPT-4o at substantially lower latency and cost. It retains a 1 million token context window and scores 45.1% on hard...

openai/gpt-4.1-mini:batch 1.04758M context $0.2/M input $0.8/M output