Meta logo

Meta: Llama 4 Maverick 17B Instruct

Open multimodal Llama for strong reasoning with efficient everyday serving

Source-linked llama Open weights Released 2025-04-05
API record Report
InputT
OutputT
Input price$0.14/M
Output price$0.59/M
Context1M
Max output16.384K
Providers6
Inference availability

Providers

Provider-specific identifiers, limits, and listed prices per million tokens. Every row links back to the provider's own documentation.

Report a price
Providers offering Llama 4 Maverick 17B Instruct
ProviderProvider model IDContextMax outputInputOutputCache readCapabilitiesDocs
Cortecs llama-4-maverick 1M 16.384K $0.124 $0.603 $0.03 Tools Docs ↗
Abacus meta-llama/Llama-4-Maverick-17B-128E-Instruct-FP8 1.04858M 8.192K $0.14 $0.59 Tools Docs ↗
Amazon Bedrock meta.llama4-maverick-17b-instruct-v1:0 1M 16.384K $0.24 $0.97 Tools Docs ↗
DevPass (LLM Gateway) llama-4-maverick-17b-instruct 1.04858M 2.048K $0.27 $0.85 Not listed Docs ↗
Charm Hyper llama-4-maverick-17b-128e-instruct-fp8 430K 43K $0.274 $0.899 $0.137 Tools Docs ↗
Neon llama-4-maverick 1M 8.192K $0.5 $1.5 ToolsJSON Docs ↗

Capability badges appear only where the provider catalog explicitly lists support. A blank cell means the source is silent, not that the feature is absent.

Listed rates

Price across providers

Input price per million tokens as published by each provider. Bars are drawn from listed rates only — no traffic weighting, since the catalog observes no requests.

Lowest input $0.124/M

Across 6 priced providers

Median input $0.255/M

Midpoint of listed rates

Highest input $0.5/M

4.0× the lowest listed rate

Output range $0.59 – $1.5

Per million output tokens

Cortecs $0.124/MLowest
Abacus $0.14/M
Amazon Bedrock $0.24/M
DevPass (LLM Gateway) $0.27/M
Charm Hyper $0.274/M
Neon $0.5/M
Cost calculator

Estimate a workload

$0.00
Excludes taxes, non-token charges, and tiered discounts.

Context limits also differ by provider, from 430K to 1.04858M tokens. Compare the provider table above before choosing on price alone.

Specification

Capabilities

Recorded from the source catalog and provider listings.

× Reasoning No
Tool calling Yes
? Structured output Unknown
Attachments Yes
Vision input Yes
Open weights Yes
Creator
Meta
Model family
llama
Knowledge cutoff
2024-08
License
Not documented
Release date
2025-04-05
Model ID
meta/llama-4-maverick-17b-instruct

Weights: Hugging Face ↗

Published evaluations

Benchmarks

Every result stays attached to its source, version, metric and harness. Scores from different versions are never merged, and the profile below plots each benchmark against its own population rather than on a shared scale.

Benchmark registry
SWE-Bench Pro resolve rate Source ↗
SWE-Bench Pro #38 of 38
Aider Polyglot percent correct · 2025-04-06 Source ↗
Aider Polyglot #27 of 31

Each strip shows every published score for that benchmark, with this model marked. The lighter shape behind the ticks is the density of results, and the dashed line is the median.

Cost against capability

Price and performance

Listed input price plotted against SWE-Bench Pro, the benchmark with the widest published coverage that this model appears in.

0 43 85 $1 $10 Qwen3 235B-A22B — $0.7/M, score 21.41 Qwen3-Coder 480B-A35B Instruct — $1.5/M, score 38.7 Qwen3.7 Max — $2.5/M, score 60.6 Claude Fable 5 — $10/M, score 80.3 Claude Haiku 4.5 (latest) — $1/M, score 39.45 Claude Opus 4.5 — $5/M, score 45.89 Claude Opus 4.6 — $5/M, score 51.9 Claude Opus 4.7 — $5/M, score 64.3 Claude Opus 4.8 — $5/M, score 69.2 Claude Opus 5 — $5/M, score 79.2 Claude Sonnet 4 (latest) — $2.898/M, score 42.7 Claude Sonnet 4.5 (latest) — $3/M, score 43.6 Claude Sonnet 5 — $2/M, score 63.2 Gemini 3 Flash Preview — $0.5/M, score 34.63 Gemini 3 Pro Preview — $0.57/M, score 43.3 Gemini 3.1 Pro Preview — $2/M, score 54.2 Gemini 3.5 Flash — $1.5/M, score 55.1 Gemini 3.5 Flash Lite — $0.3/M, score 54.2 LongCat-2.0 — $0.3/M, score 59.5 Llama 4 Maverick 17B Instruct — $0.14/M, score 5.24 Muse Glimmer 30B — $0.2/M, score 51.2 Muse Spark 1.1 — $1.25/M, score 61.5 MiniMax-M2.1 — $0.3/M, score 36.81 GPT-5 — $1.25/M, score 41.78 GPT-5.2 — $1.75/M, score 29.94 GPT-5.2 Codex — $0.14/M, score 41.04 GPT-5.4 — $2.5/M, score 59.1 GPT-5.4 mini — $0.75/M, score 54.4 GPT-5.4 nano — $0.2/M, score 52.4 GPT-5.5 — $5/M, score 58.6 GPT-5.6 Luna — $0.2/M, score 62.7 GPT-5.6 Sol — $4/M, score 64.6 GPT-5.6 Terra — $2/M, score 63.4 Step 3.7 Flash — $0.185/M, score 56.3 Grok 4.5 — $2/M, score 64.7 MiMo-V2.5-Pro — $0.435/M, score 57.2 GLM-4.6 — $0.6/M, score 9.67 GLM-5.2 — $1.4/M, score 62.1 Llama 4 Maverick 17B Instruct Input price per million tokens (log scale) Resolve rate

The stepped line is the efficient frontier: at each price, the best score available for that money or less. A model sitting on it is not being beaten by anything cheaper. Price is log-scaled because listed rates span four orders of magnitude. Only models with both a listed price and a score on this benchmark can appear.

Catalog activity

Change log

Field-level changes detected between successful source imports.

Full change log
Price Completion0.603 → 0.59
Price Prompt0.124 → 0.14
Provenance

Sources & verification

Every figure on this page traces back to one of these records.

Methodology
Public API

Use this record

Fetch the complete source-linked model record. No key, no account, no rate-limited tier.

API documentation
Endpoint
GET https://model.kyssta.lol/api/v1/models/meta/llama-4-maverick-17b-instruct
curl
curl "https://model.kyssta.lol/api/v1/models/meta/llama-4-maverick-17b-instruct"
Common questions

Frequently asked questions

Answered directly from the stored record — nothing here is generated beyond the catalog's own fields.

What is Llama 4 Maverick 17B Instruct?

Open multimodal Llama for strong reasoning with efficient everyday serving. It is published by Meta and catalogued here from Models.dev.

How much does Llama 4 Maverick 17B Instruct cost?

Listed input pricing starts at $0.124 per million tokens from Cortecs, rising to $0.5 across 6 listed providers.

What is the context length of Llama 4 Maverick 17B Instruct?

Llama 4 Maverick 17B Instruct accepts up to 1M tokens of context and returns up to 16.384K output tokens.

Does Llama 4 Maverick 17B Instruct support tool calling and structured output?

Provider catalogs list support for tool calling, and image input.

Which providers serve Llama 4 Maverick 17B Instruct?

6 providers list this model: Cortecs, Abacus, Amazon Bedrock, DevPass (LLM Gateway), Charm Hyper, Neon.

Are the weights for Llama 4 Maverick 17B Instruct open?

Yes. The weights are published and downloadable from Hugging Face.

When was Llama 4 Maverick 17B Instruct released?

The catalog records a release date of 2025-04-05, last verified Sep 11, 2026.

More models from Meta

View all →