OpenAI logo

OpenAI: OpenAI: GPT-4.1 (batch)

GPT-4.1 is a flagship large language model optimized for advanced instruction following, real-world software engineering, and long-context reasoning. It supports a 1 million token context window and outperforms GPT-4o and...

Source-linked Weight access not listed
API record Report
InputT
OutputT
Input price$1/M
Output price$4/M
Context1.04758M
Max output32.768K
Providers0
Inference availability

Providers

Provider-specific identifiers, limits, and listed prices per million tokens. Every row links back to the provider's own documentation.

Report a price
No inference providers are listed

The model record exists, but no source-linked provider offer is available yet.

Submit a source
Specification

Capabilities

Recorded from the source catalog and provider listings.

? Reasoning Unknown
? Tool calling Unknown
? Structured output Unknown
? Attachments Unknown
Vision input Yes
? Open weights Unknown
Creator
OpenAI
Model family
Not documented
Knowledge cutoff
Not documented
License
Not documented
Release date
Not documented
Model ID
openai/gpt-4.1:batch
Published evaluations

Benchmarks

Every result stays attached to its source, version, metric and harness. Scores from different versions are never merged, and the profile below plots each benchmark against its own population rather than on a shared scale.

Benchmark registry
Benchmark profile
Design Arena: w…Design Arena: website Design Arena: c…Design Arena: codecategories Design Arena: g…Design Arena: gamedev Design Arena: d…Design Arena: dataviz Design Arena: u…Design Arena: uicomponent Design Arena: 3dDesign Arena: 3d This model — Design Arena: website: 1051 elo (16.6th percentile) This model — Design Arena: codecategories: 1040 elo (13.8th percentile) This model — Design Arena: gamedev: 1102 elo (23.2th percentile) This model — Design Arena: dataviz: 1118 elo (23.6th percentile) This model — Design Arena: uicomponent: 1015 elo (12.4th percentile) This model — Design Arena: 3d: 878 elo (1.9th percentile)

Each axis is this model's percentile among the 6 benchmarks it has published results for, measured against every other model with a score on that same benchmark. Percentiles are used because benchmarks do not share a scale — a 60 on one is not a 60 on another. Hover any point for the raw score.

Design Arena: website elo Source ↗
Design Arena: website #146 of 175
Design Arena: codecategories elo Source ↗
Design Arena: codecategories #144 of 167
Design Arena: gamedev elo Source ↗
Design Arena: gamedev #127 of 166
Design Arena: dataviz elo Source ↗
Design Arena: dataviz #126 of 165
Design Arena: uicomponent elo Source ↗
Design Arena: uicomponent #141 of 161
Design Arena: 3d elo Source ↗
Design Arena: 3d #154 of 157

Each strip shows every published score for that benchmark, with this model marked. The lighter shape behind the ticks is the density of results, and the dashed line is the median.

Cost against capability

Price and performance

Listed input price plotted against Design Arena: website, the benchmark with the widest published coverage that this model appears in.

0 720 1441 $0.1 $1 $10 OpenAI: GPT-5.1 (batch) — $0.625/M, score 1199.0 Anthropic: Claude Haiku 4.5 (batch) — $0.5/M, score 1135.0 OpenAI: GPT-5 (batch) — $0.625/M, score 1197.0 OpenAI: GPT-4.1 Mini (batch) — $0.2/M, score 1010.0 Mistral: Ministral 3 8B 2512 (batch) — $0.075/M, score 1077.0 Mistral: Mistral Small 3.2 24B — $0.075/M, score 908.0 Mistral: Mistral Large 3 2512 (batch) — $0.25/M, score 1175.0 DeepSeek: DeepSeek V3.2 Exp — $0.27/M, score 1191.0 Anthropic: Claude Opus 4 — $15/M, score 1177.0 Qwen: Qwen3 235B A22B — $0.455/M, score 1043.0 Gemini 3.5 Flash — $1.5/M, score 1275.0 Z.ai: GLM 5.3 Flash (batch) — $0.075/M, score 1291.0 GPT-4.1 mini — $0.4/M, score 1010.0 DeepSeek V4 Pro — $0.435/M, score 1247.0 Claude Sonnet 5 — $2/M, score 1289.0 Mistral: Ministral 3 8B 2512 — $0.15/M, score 1077.0 Z.ai: GLM 5.1 — $0.966/M, score 1290.0 Gemini 3.7 Flash — $0.75/M, score 1315.0 Muse Spark 1.3 — $1.25/M, score 1359.0 Google: Gemini 3 Flash Preview (batch) — $0.25/M, score 1208.0 Anthropic: Claude Opus 4.8 (batch) — $2.5/M, score 1267.0 Mistral: Mistral Large 3 2512 — $0.5/M, score 1175.0 DeepSeek V3.2 — $0.18/M, score 1187.0 OpenAI: GPT-5.5 (batch) — $2.5/M, score 1269.0 Anthropic: Claude Opus 4.7 — $5/M, score 1307.0 Google: Gemini 3.5 Flash (batch) — $0.75/M, score 1275.0 Google: Gemini 2.5 Pro (batch) — $0.625/M, score 1179.0 Inception: Mercury 2 — $0.25/M, score 1035.0 OpenAI: GPT-4.1 Nano (batch) — $0.05/M, score 985.0 Gemini 3.6 Flash — $0.75/M, score 1312.0 Anthropic: Claude Sonnet 4.5 — $3/M, score 1202.0 Z.ai: GLM 4.7 Flash — $0.061/M, score 1206.0 Claude Fable 5 — $10/M, score 1310.0 Gemini 2.5 Pro — $1.25/M, score 1179.0 Z.ai: GLM 5.3 Flash — $0.15/M, score 1291.0 MiniMax: MiniMax M3 — $0.3/M, score 1271.0 Google: Gemini 2.5 Flash (batch) — $0.15/M, score 1127.0 Anthropic: Claude Fable 5.1 — $10/M, score 1322.0 Z.ai: GLM 5V Turbo — $1.2/M, score 1243.0 MoonshotAI: Kimi K2.7 Code (batch) — $0.95/M, score 1290.0 OpenAI: o3 (batch) — $1/M, score 1048.0 Inkling — $1.87/M, score 1226.0 OpenAI: gpt-oss-20b (batch) — $0.05/M, score 865.0 Muse Spark 1.1 — $1.25/M, score 1282.0 MiMo-V2.5 — $0.14/M, score 1279.0 MoonshotAI: Kimi K2 0711 — $0.57/M, score 1063.0 Kimi K2 Thinking — $0.4/M, score 1125.0 Qwen: Qwen3.8 Max (0803) — $2/M, score 1295.0 Mistral: Mistral Medium 3.1 (batch) — $0.2/M, score 1145.0 GPT-5.1 — $1.25/M, score 1199.0 GPT-5.4 — $2.5/M, score 1232.0 Anthropic: Claude Fable 5.1 (batch) — $5/M, score 1322.0 Google: Gemini 3.8 Flash (batch) — $0.375/M, score 1315.0 Google: Gemini 3.7 Flash (batch) — $0.375/M, score 1315.0 GPT-5.3 Codex — $1.75/M, score 1175.0 Z.ai: GLM 5.3 (batch) — $0.7/M, score 1319.0 DeepSeek: DeepSeek V4 Flash 0731 (batch) — $0.11/M, score 1251.0 Anthropic: Claude Opus 4.6 (batch) — $2.5/M, score 1304.0 MoonshotAI: Kimi K3 (batch) — $3/M, score 1354.0 GPT-5.5 — $5/M, score 1269.0 Anthropic: Claude Sonnet 4.5 (batch) — $1.5/M, score 1202.0 OpenAI: GPT-5 Nano (batch) — $0.025/M, score 1114.0 Mistral: Codestral 2508 (batch) — $0.15/M, score 1025.0 Anthropic: Claude Sonnet 5 (batch) — $1/M, score 1289.0 OpenAI: GPT-4.1 (batch) — $1/M, score 1051.0 OpenAI: GPT-5.2 (batch) — $0.875/M, score 1207.0 Anthropic: Claude Fable 5 (batch) — $5/M, score 1310.0 GPT-5.1 Codex mini — $0.22/M, score 1123.0 Thinking Machines: Inkling (batch) — $1/M, score 1226.0 GPT OSS 120B — $0.03/M, score 980.0 OpenAI: GPT-4o (batch) — $1.25/M, score 843.0 OpenAI: gpt-oss-120b (batch) — $0.15/M, score 980.0 Mistral: Codestral 2508 — $0.3/M, score 1025.0 SpaceXAI: Grok 4.6 — $2/M, score 1306.0 Claude Opus 5 — $5/M, score 1320.0 DeepSeek: DeepSeek V3.1 Terminus — $0.27/M, score 1199.0 GPT-5 Mini — $0.25/M, score 1137.0 Claude Opus 5 (batch) — $2.5/M, score 1320.0 Gemini 3.1 Flash Lite Preview — $0.25/M, score 1095.0 o3 — $2/M, score 1048.0 OpenAI: GPT-5.4 (batch) — $1.25/M, score 1232.0 DeepSeek V4 Flash — $0.15/M, score 1220.0 DeepSeek Chat — $0.147/M, score 1132.0 Qwen: Qwen3 Max — $0.78/M, score 1131.0 GPT-5.1 Codex — $1.07/M, score 1174.0 Kimi K2.7 Code — $0.95/M, score 1279.0 GPT-4.1 — $2/M, score 1051.0 Z.ai: GLM 5.2 (batch) — $0.7/M, score 1307.0 Google: Gemini 3.6 Flash (batch) — $0.375/M, score 1312.0 Z.ai: GLM 5.3 — $1.4/M, score 1319.0 Anthropic: Claude Opus 4.8 — $5/M, score 1267.0 SpaceXAI: Grok 4.3 — $1.25/M, score 1206.0 Gemini 3 Flash Preview — $0.5/M, score 1208.0 MiniMax: MiniMax M3 (batch) — $0.3/M, score 1271.0 Gemini 3.8 Flash — $0.75/M, score 1315.0 NVIDIA: Nemotron 3 Ultra (batch) — $0.6/M, score 1144.0 OpenAI: o4 Mini (batch) — $0.55/M, score 998.0 MiniMax: MiniMax M2.7 — $0.3/M, score 1258.0 SpaceXAI: Grok 4.5 — $2/M, score 1296.0 Mistral: Mistral Medium 3.1 — $0.4/M, score 1145.0 Kimi K2.6 — $0.95/M, score 1282.0 GPT-5.2 — $1.75/M, score 1207.0 Kimi K3 — $3/M, score 1354.0 MiniMax: MiniMax M2 — $0.255/M, score 1155.0 Solar Pro 4 — $0.3/M, score 1188.0 Gemini 3.1 Pro Preview — $2/M, score 1265.0 Anthropic: Claude Opus 4.1 (batch) — $7.5/M, score 1189.0 Qwen: Qwen3 Coder 480B A35B — $0.3/M, score 1171.0 Anthropic: Claude Opus 4.7 (batch) — $2.5/M, score 1307.0 Mistral: Mistral Medium 3 — $0.4/M, score 1091.0 MiMo-V2.5-Pro — $0.435/M, score 1285.0 Google: Gemini 3.1 Pro Preview (batch) — $1/M, score 1265.0 Anthropic: Claude Opus 4.1 — $15/M, score 1189.0 Z.ai: GLM 5 — $0.6/M, score 1260.0 Anthropic: Claude Sonnet 4.6 — $3/M, score 1297.0 Anthropic: Claude Sonnet 4.6 (batch) — $1.5/M, score 1297.0 DeepSeek V4 Flash 0731 — $0.05/M, score 1251.0 GPT-5 — $1.25/M, score 1197.0 OpenAI: GPT-5 Mini (batch) — $0.125/M, score 1137.0 Anthropic: Claude Opus 4.5 (batch) — $2.5/M, score 1259.0 Qwen: Qwen3.7 Plus — $0.32/M, score 1282.0 SpaceXAI: Grok 4.3 (batch) — $1/M, score 1206.0 o4-mini — $1.1/M, score 998.0 Nemotron 3 Ultra 550B A55B — $0.5/M, score 1144.0 Anthropic: Claude Sonnet 4 — $3/M, score 1158.0 SpaceXAI: Grok 4.20 — $1.25/M, score 1242.0 Hy3 — $0.066/M, score 1194.0 Qwen: Qwen3.6 Plus — $0.325/M, score 1253.0 Z.ai: GLM 4.6 — $0.43/M, score 1187.0 GPT-4o — $2.5/M, score 843.0 Gemini 2.5 Flash — $0.3/M, score 1127.0 Muse Spark 1.2 — $1.25/M, score 1322.0 GPT-4.1 nano — $0.1/M, score 985.0 GPT OSS 20B — $0.02/M, score 865.0 Kimi K2.5 — $0.3/M, score 1261.0 Trinity Large Thinking — $0.25/M, score 1149.0 MoonshotAI: Kimi K2 0905 — $0.6/M, score 1120.0 Qwen: Qwen3 30B A3B Thinking 2507 — $0.2/M, score 943.0 Z.ai: GLM 5.2 — $0.966/M, score 1307.0 Qwen: Qwen3.7 Max — $1.475/M, score 1285.0 Qwen: Qwen3.5 Plus 2026-02-15 — $0.26/M, score 1200.0 MiniMax: MiniMax M2.1 — $0.3/M, score 1214.0 Z.ai: GLM 4.7 — $0.4/M, score 1238.0 Mistral: Ministral 3 3B 2512 — $0.1/M, score 1040.0 Amazon: Nova Premier 1.0 — $2.5/M, score 847.0 DeepSeek: DeepSeek V3.1 — $0.25/M, score 1135.0 Z.ai: GLM 4.5 — $0.6/M, score 1182.0 Meta: Llama 4 Maverick — $0.2/M, score 883.0 Nex AGI: Nex-N2-Pro — $0.25/M, score 1236.0 Z.ai: GLM 5 Turbo — $1.2/M, score 1278.0 Qwen: Qwen3.5 397B A17B — $0.55/M, score 1203.0 MiniMax: MiniMax M2.5 — $0.3/M, score 1235.0 Anthropic: Claude Opus 4.6 — $5/M, score 1304.0 Mistral: Ministral 3 14B 2512 — $0.2/M, score 1093.0 Anthropic: Claude Opus 4.5 — $5/M, score 1259.0 Anthropic: Claude Haiku 4.5 — $1/M, score 1135.0 Qwen: Qwen3 Coder 30B A3B Instruct — $0.07/M, score 1100.0 Z.ai: GLM 4.5 Air — $0.13/M, score 1159.0 Qwen: Qwen3 235B A22B Thinking 2507 — $0.23/M, score 1065.0 Qwen: Qwen3 235B A22B Instruct 2507 — $0.22/M, score 1070.0 DeepSeek: R1 0528 — $0.5/M, score 1161.0 Qwen: Qwen3 30B A3B — $0.12/M, score 967.0 Meta: Llama 4 Scout — $0.1/M, score 762.0 Amazon: Nova Pro 1.0 — $0.8/M, score 808.0 GPT-5 Nano — $0.05/M, score 1114.0 Step 3.7 Flash — $0.185/M, score 1206.0 OpenAI: GPT-4.1 (batch) Input price per million tokens (log scale) Elo

The stepped line is the efficient frontier: at each price, the best score available for that money or less. A model sitting on it is not being beaten by anything cheaper. Price is log-scaled because listed rates span four orders of magnitude. Only models with both a listed price and a score on this benchmark can appear.

Catalog activity

Change log

Field-level changes detected between successful source imports.

Full change log
No changes recorded

This record has not changed within the retained import history.

Public API

Use this record

Fetch the complete source-linked model record. No key, no account, no rate-limited tier.

API documentation
Endpoint
GET https://model.kyssta.lol/api/v1/models/openai/gpt-4.1:batch
curl
curl "https://model.kyssta.lol/api/v1/models/openai/gpt-4.1:batch"
Common questions

Frequently asked questions

Answered directly from the stored record — nothing here is generated beyond the catalog's own fields.

What is OpenAI: GPT-4.1 (batch)?

GPT-4.1 is a flagship large language model optimized for advanced instruction following, real-world software engineering, and long-context reasoning. It supports a 1 million token context window and outperforms GPT-4o and. It is published by OpenAI and catalogued here from OpenRouter.

What is the context length of OpenAI: GPT-4.1 (batch)?

OpenAI: GPT-4.1 (batch) accepts up to 1.04758M tokens of context and returns up to 32.768K output tokens.

Does OpenAI: GPT-4.1 (batch) support tool calling and structured output?

Provider catalogs list support for image input.

More models from OpenAI

View all →