OpenAI logo

OpenAI: OpenAI: o4 Mini (batch)

OpenAI o4-mini is a compact reasoning model in the o-series, optimized for fast, cost-efficient performance while retaining strong multimodal and agentic capabilities. It supports tool use and demonstrates competitive reasoning...

Source-linked Weight access not listed
API record Report
InputT
OutputT
Input price$0.55/M
Output price$2.2/M
Context200K
Max output100K
Providers0
Inference availability

Providers

Provider-specific identifiers, limits, and listed prices per million tokens. Every row links back to the provider's own documentation.

Report a price
No inference providers are listed

The model record exists, but no source-linked provider offer is available yet.

Submit a source
Specification

Capabilities

Recorded from the source catalog and provider listings.

? Reasoning Unknown
? Tool calling Unknown
? Structured output Unknown
? Attachments Unknown
Vision input Yes
? Open weights Unknown
Creator
OpenAI
Model family
Not documented
Knowledge cutoff
Not documented
License
Not documented
Release date
Not documented
Model ID
openai/o4-mini:batch
Published evaluations

Benchmarks

Every result stays attached to its source, version, metric and harness. Scores from different versions are never merged, and the profile below plots each benchmark against its own population rather than on a shared scale.

Benchmark registry
Benchmark profile
Design Arena: w…Design Arena: website Design Arena: c…Design Arena: codecategories Design Arena: g…Design Arena: gamedev Design Arena: d…Design Arena: dataviz Design Arena: u…Design Arena: uicomponent Design Arena: 3dDesign Arena: 3d This model — Design Arena: website: 998 elo (9.7th percentile) This model — Design Arena: codecategories: 991 elo (6.6th percentile) This model — Design Arena: gamedev: 1026 elo (12.7th percentile) This model — Design Arena: dataviz: 1005 elo (9.7th percentile) This model — Design Arena: uicomponent: 998 elo (11.2th percentile) This model — Design Arena: 3d: 882 elo (3.2th percentile)

Each axis is this model's percentile among the 6 benchmarks it has published results for, measured against every other model with a score on that same benchmark. Percentiles are used because benchmarks do not share a scale — a 60 on one is not a 60 on another. Hover any point for the raw score.

Design Arena: website elo Source ↗
Design Arena: website #158 of 175
Design Arena: codecategories elo Source ↗
Design Arena: codecategories #156 of 167
Design Arena: gamedev elo Source ↗
Design Arena: gamedev #145 of 166
Design Arena: dataviz elo Source ↗
Design Arena: dataviz #149 of 165
Design Arena: uicomponent elo Source ↗
Design Arena: uicomponent #143 of 161
Design Arena: 3d elo Source ↗
Design Arena: 3d #152 of 157

Each strip shows every published score for that benchmark, with this model marked. The lighter shape behind the ticks is the density of results, and the dashed line is the median.

Cost against capability

Price and performance

Listed input price plotted against Design Arena: website, the benchmark with the widest published coverage that this model appears in.

0 720 1441 $0.1 $1 $10 Qwen: Qwen3.6 Plus — $0.325/M, score 1253.0 Z.ai: GLM 5 — $0.6/M, score 1260.0 Mistral: Ministral 3 8B 2512 (batch) — $0.075/M, score 1077.0 DeepSeek: DeepSeek V3.2 Exp — $0.27/M, score 1191.0 Qwen: Qwen3 235B A22B — $0.455/M, score 1043.0 Gemini 3.5 Flash — $1.5/M, score 1275.0 Z.ai: GLM 5.3 Flash (batch) — $0.075/M, score 1291.0 GPT-4.1 mini — $0.4/M, score 1010.0 DeepSeek V4 Pro — $0.435/M, score 1247.0 Claude Sonnet 5 — $2/M, score 1289.0 Mistral: Ministral 3 8B 2512 — $0.15/M, score 1077.0 Gemini 3.1 Flash Lite Preview — $0.25/M, score 1095.0 Anthropic: Claude Fable 5.1 (batch) — $5/M, score 1322.0 Gemini 3.7 Flash — $0.75/M, score 1315.0 Muse Spark 1.3 — $1.25/M, score 1359.0 Z.ai: GLM 5.1 — $0.966/M, score 1290.0 OpenAI: GPT-5.2 (batch) — $0.875/M, score 1207.0 Anthropic: Claude Opus 4.8 (batch) — $2.5/M, score 1267.0 Mistral: Mistral Large 3 2512 — $0.5/M, score 1175.0 DeepSeek V3.2 — $0.18/M, score 1187.0 Anthropic: Claude Opus 4.7 — $5/M, score 1307.0 Google: Gemini 3.5 Flash (batch) — $0.75/M, score 1275.0 Mistral: Codestral 2508 — $0.3/M, score 1025.0 Inception: Mercury 2 — $0.25/M, score 1035.0 Z.ai: GLM 4.7 Flash — $0.061/M, score 1206.0 Gemini 3.6 Flash — $0.75/M, score 1312.0 Claude Fable 5 — $10/M, score 1310.0 Z.ai: GLM 5.3 Flash — $0.15/M, score 1291.0 MiniMax: MiniMax M3 — $0.3/M, score 1271.0 Google: Gemini 2.5 Flash (batch) — $0.15/M, score 1127.0 Anthropic: Claude Fable 5.1 — $10/M, score 1322.0 Z.ai: GLM 5V Turbo — $1.2/M, score 1243.0 MoonshotAI: Kimi K2.7 Code (batch) — $0.95/M, score 1290.0 Inkling — $1.87/M, score 1226.0 OpenAI: gpt-oss-20b (batch) — $0.05/M, score 865.0 Muse Spark 1.1 — $1.25/M, score 1282.0 MiMo-V2.5 — $0.14/M, score 1279.0 MoonshotAI: Kimi K2 0711 — $0.57/M, score 1063.0 Kimi K2 Thinking — $0.4/M, score 1125.0 Qwen: Qwen3.8 Max (0803) — $2/M, score 1295.0 Mistral: Mistral Medium 3.1 (batch) — $0.2/M, score 1145.0 GPT-5.1 — $1.25/M, score 1199.0 Google: Gemini 3.8 Flash (batch) — $0.375/M, score 1315.0 Google: Gemini 3.7 Flash (batch) — $0.375/M, score 1315.0 Z.ai: GLM 5.3 (batch) — $0.7/M, score 1319.0 DeepSeek: DeepSeek V4 Flash 0731 (batch) — $0.11/M, score 1251.0 OpenAI: GPT-5.5 (batch) — $2.5/M, score 1269.0 MoonshotAI: Kimi K3 (batch) — $3/M, score 1354.0 Anthropic: Claude Opus 4.6 (batch) — $2.5/M, score 1304.0 OpenAI: GPT-5 Nano (batch) — $0.025/M, score 1114.0 Anthropic: Claude Sonnet 5 (batch) — $1/M, score 1289.0 Anthropic: Claude Fable 5 (batch) — $5/M, score 1310.0 OpenAI: GPT-5 (batch) — $0.625/M, score 1197.0 Google: Gemini 2.5 Pro (batch) — $0.625/M, score 1179.0 Anthropic: Claude Opus 4 — $15/M, score 1177.0 GPT OSS 120B — $0.03/M, score 980.0 OpenAI: gpt-oss-120b (batch) — $0.15/M, score 980.0 Mistral: Codestral 2508 (batch) — $0.15/M, score 1025.0 OpenAI: o3 (batch) — $1/M, score 1048.0 OpenAI: GPT-4.1 Mini (batch) — $0.2/M, score 1010.0 OpenAI: GPT-4.1 Nano (batch) — $0.05/M, score 985.0 SpaceXAI: Grok 4.6 — $2/M, score 1306.0 DeepSeek: DeepSeek V3.1 Terminus — $0.27/M, score 1199.0 o3 — $2/M, score 1048.0 Gemini 3.1 Pro Preview — $2/M, score 1265.0 Qwen: Qwen3 Max — $0.78/M, score 1131.0 GPT-4.1 — $2/M, score 1051.0 Z.ai: GLM 5.2 (batch) — $0.7/M, score 1307.0 Google: Gemini 3.6 Flash (batch) — $0.375/M, score 1312.0 Z.ai: GLM 5.3 — $1.4/M, score 1319.0 Anthropic: Claude Opus 4.8 — $5/M, score 1267.0 SpaceXAI: Grok 4.3 — $1.25/M, score 1206.0 Gemini 3 Flash Preview — $0.5/M, score 1208.0 MiniMax: MiniMax M3 (batch) — $0.3/M, score 1271.0 Gemini 3.8 Flash — $0.75/M, score 1315.0 NVIDIA: Nemotron 3 Ultra (batch) — $0.6/M, score 1144.0 MiniMax: MiniMax M2.7 — $0.3/M, score 1258.0 Kimi K2.6 — $0.95/M, score 1282.0 GPT-5.2 — $1.75/M, score 1207.0 Kimi K3 — $3/M, score 1354.0 Anthropic: Claude Sonnet 4.5 (batch) — $1.5/M, score 1202.0 Claude Opus 5 — $5/M, score 1320.0 OpenAI: GPT-4.1 (batch) — $1/M, score 1051.0 Claude Opus 5 (batch) — $2.5/M, score 1320.0 SpaceXAI: Grok 4.5 — $2/M, score 1296.0 Google: Gemini 3 Flash Preview (batch) — $0.25/M, score 1208.0 Mistral: Mistral Large 3 2512 (batch) — $0.25/M, score 1175.0 OpenAI: GPT-5.1 (batch) — $0.625/M, score 1199.0 GPT-5.1 Codex — $1.07/M, score 1174.0 Anthropic: Claude Opus 4.1 (batch) — $7.5/M, score 1189.0 Anthropic: Claude Opus 4.7 (batch) — $2.5/M, score 1307.0 Mistral: Mistral Medium 3 — $0.4/M, score 1091.0 MiniMax: MiniMax M2 — $0.255/M, score 1155.0 Anthropic: Claude Sonnet 4.5 — $3/M, score 1202.0 Google: Gemini 3.1 Pro Preview (batch) — $1/M, score 1265.0 Mistral: Mistral Medium 3.1 — $0.4/M, score 1145.0 Anthropic: Claude Opus 4.1 — $15/M, score 1189.0 MiMo-V2.5-Pro — $0.435/M, score 1285.0 Thinking Machines: Inkling (batch) — $1/M, score 1226.0 GPT-5.5 — $5/M, score 1269.0 OpenAI: GPT-5.4 (batch) — $1.25/M, score 1232.0 Anthropic: Claude Sonnet 4.6 — $3/M, score 1297.0 Anthropic: Claude Sonnet 4.6 (batch) — $1.5/M, score 1297.0 DeepSeek V4 Flash 0731 — $0.05/M, score 1251.0 GPT-5 — $1.25/M, score 1197.0 Anthropic: Claude Haiku 4.5 (batch) — $0.5/M, score 1135.0 OpenAI: GPT-5 Mini (batch) — $0.125/M, score 1137.0 Gemini 2.5 Pro — $1.25/M, score 1179.0 Qwen: Qwen3 Coder 480B A35B — $0.3/M, score 1171.0 Mistral: Mistral Small 3.2 24B — $0.075/M, score 908.0 Anthropic: Claude Opus 4.5 (batch) — $2.5/M, score 1259.0 OpenAI: o4 Mini (batch) — $0.55/M, score 998.0 Qwen: Qwen3.7 Plus — $0.32/M, score 1282.0 SpaceXAI: Grok 4.3 (batch) — $1/M, score 1206.0 o4-mini — $1.1/M, score 998.0 DeepSeek Chat — $0.147/M, score 1132.0 Nemotron 3 Ultra 550B A55B — $0.5/M, score 1144.0 Anthropic: Claude Sonnet 4 — $3/M, score 1158.0 SpaceXAI: Grok 4.20 — $1.25/M, score 1242.0 Hy3 — $0.066/M, score 1194.0 GPT-5.1 Codex mini — $0.22/M, score 1123.0 GPT OSS 20B — $0.02/M, score 865.0 OpenAI: GPT-4o (batch) — $1.25/M, score 843.0 GPT-5.3 Codex — $1.75/M, score 1175.0 Z.ai: GLM 4.6 — $0.43/M, score 1187.0 GPT-5 Mini — $0.25/M, score 1137.0 GPT-4o — $2.5/M, score 843.0 Kimi K2.7 Code — $0.95/M, score 1279.0 Gemini 2.5 Flash — $0.3/M, score 1127.0 Muse Spark 1.2 — $1.25/M, score 1322.0 DeepSeek V4 Flash — $0.15/M, score 1220.0 GPT-4.1 nano — $0.1/M, score 985.0 GPT-5.4 — $2.5/M, score 1232.0 Solar Pro 4 — $0.3/M, score 1188.0 Kimi K2.5 — $0.3/M, score 1261.0 Trinity Large Thinking — $0.25/M, score 1149.0 MoonshotAI: Kimi K2 0905 — $0.6/M, score 1120.0 Qwen: Qwen3 30B A3B Thinking 2507 — $0.2/M, score 943.0 Z.ai: GLM 5.2 — $0.966/M, score 1307.0 Qwen: Qwen3.7 Max — $1.475/M, score 1285.0 Qwen: Qwen3.5 Plus 2026-02-15 — $0.26/M, score 1200.0 MiniMax: MiniMax M2.1 — $0.3/M, score 1214.0 Z.ai: GLM 4.7 — $0.4/M, score 1238.0 Mistral: Ministral 3 3B 2512 — $0.1/M, score 1040.0 Amazon: Nova Premier 1.0 — $2.5/M, score 847.0 DeepSeek: DeepSeek V3.1 — $0.25/M, score 1135.0 Z.ai: GLM 4.5 — $0.6/M, score 1182.0 Meta: Llama 4 Maverick — $0.2/M, score 883.0 Nex AGI: Nex-N2-Pro — $0.25/M, score 1236.0 Z.ai: GLM 5 Turbo — $1.2/M, score 1278.0 Qwen: Qwen3.5 397B A17B — $0.55/M, score 1203.0 MiniMax: MiniMax M2.5 — $0.3/M, score 1235.0 Anthropic: Claude Opus 4.6 — $5/M, score 1304.0 Mistral: Ministral 3 14B 2512 — $0.2/M, score 1093.0 Anthropic: Claude Opus 4.5 — $5/M, score 1259.0 Anthropic: Claude Haiku 4.5 — $1/M, score 1135.0 Qwen: Qwen3 Coder 30B A3B Instruct — $0.07/M, score 1100.0 Z.ai: GLM 4.5 Air — $0.13/M, score 1159.0 Qwen: Qwen3 235B A22B Thinking 2507 — $0.23/M, score 1065.0 Qwen: Qwen3 235B A22B Instruct 2507 — $0.22/M, score 1070.0 DeepSeek: R1 0528 — $0.5/M, score 1161.0 Qwen: Qwen3 30B A3B — $0.12/M, score 967.0 Meta: Llama 4 Scout — $0.1/M, score 762.0 Amazon: Nova Pro 1.0 — $0.8/M, score 808.0 GPT-5 Nano — $0.05/M, score 1114.0 Step 3.7 Flash — $0.185/M, score 1206.0 OpenAI: o4 Mini (batch) Input price per million tokens (log scale) Elo

The stepped line is the efficient frontier: at each price, the best score available for that money or less. A model sitting on it is not being beaten by anything cheaper. Price is log-scaled because listed rates span four orders of magnitude. Only models with both a listed price and a score on this benchmark can appear.

Catalog activity

Change log

Field-level changes detected between successful source imports.

Full change log
No changes recorded

This record has not changed within the retained import history.

Public API

Use this record

Fetch the complete source-linked model record. No key, no account, no rate-limited tier.

API documentation
Endpoint
GET https://model.kyssta.lol/api/v1/models/openai/o4-mini:batch
curl
curl "https://model.kyssta.lol/api/v1/models/openai/o4-mini:batch"
Common questions

Frequently asked questions

Answered directly from the stored record — nothing here is generated beyond the catalog's own fields.

What is OpenAI: o4 Mini (batch)?

OpenAI o4-mini is a compact reasoning model in the o-series, optimized for fast, cost-efficient performance while retaining strong multimodal and agentic capabilities. It supports tool use and demonstrates competitive reasoning. It is published by OpenAI and catalogued here from OpenRouter.

What is the context length of OpenAI: o4 Mini (batch)?

OpenAI: o4 Mini (batch) accepts up to 200K tokens of context and returns up to 100K output tokens.

Does OpenAI: o4 Mini (batch) support tool calling and structured output?

Provider catalogs list support for image input.

More models from OpenAI

View all →