Meta logo

Meta: Llama-3.3-70B-Instruct

Popular open Llama workhorse for multilingual chat, coding, and self-hosting

Source-linked llama Open weights Released 2024-12-06
API record Report
InputT
OutputT
Input price$0.1/M
Output price$0.32/M
Context128K
Max output4.096K
Providers28
Inference availability

Providers

Provider-specific identifiers, limits, and listed prices per million tokens. Every row links back to the provider's own documentation.

Report a price
Providers offering Llama-3.3-70B-Instruct
ProviderProvider model IDContextMax outputInputOutputCache readCapabilitiesDocs
Nvidia meta/llama-3.3-70b-instruct 128K 4.096K ToolsJSON Docs ↗
Snowflake Cortex snowflake-llama3.3-70b 128K 4.096K Tools Docs ↗
Vercel AI Gateway meta/llama-3.3-70b 128K 4.096K Tools Docs ↗
Pendra llama3.3:70b 128K 4.096K Tools Docs ↗
GitHub Models meta/llama-3.3-70b-instruct 128K 32.768K ReasoningTools Docs ↗
Llama llama-3.3-70b-instruct 128K 4.096K Tools Docs ↗
Kilo Gateway meta-llama/llama-3.3-70b-instruct 131.072K 16.384K $0.1 $0.32 ToolsJSON Docs ↗
Eden AI deepinfra/meta-llama/Llama-3.3-70B-Instruct 131.072K 4.096K $0.1 $0.32 ToolsJSON Docs ↗
OpenRouter meta-llama/llama-3.3-70b-instruct 131.072K 16.384K $0.1 $0.32 ToolsJSON Docs ↗
Cortecs llama-3.3-70b-instruct 131K 131K $0.129 $0.399 ToolsJSON Docs ↗
Eden AI nebius/meta-llama/Llama-3.3-70B-Instruct 131.072K 4.096K $0.13 $0.4 $0.13 ToolsJSON Docs ↗
DevPass (LLM Gateway) llama-3.3-70b-instruct 131.072K 4.096K $0.135 $0.4 Tools Docs ↗
Merge Gateway meta/llama-3.3-70b-instruct 131.072K 32.768K $0.22 $0.5 $0.11 Tools Docs ↗
Crusoe meta-llama/Llama-3.3-70B-Instruct 128K 4.096K $0.25 $0.75 $0.13 Tools Docs ↗
Neon meta-llama-3-3-70b-instruct 128K 8.192K $0.5 $1.5 ToolsJSON Docs ↗
Abacus meta-llama/Meta-Llama-3.3-70B-Instruct 131.072K 8.192K $0.59 $0.79 Tools Docs ↗
Hugging Face meta-llama/Llama-3.3-70B-Instruct 131.072K 4.096K $0.59 $0.79 ToolsJSON Docs ↗
Charm Hyper llama-3.3-70b-instruct 128K 12.8K $0.607 $1.039 $0.303 Tools Docs ↗
Azure llama-3.3-70b-instruct 128K 32.768K $0.71 $0.71 Tools Docs ↗
Azure Cognitive Services llama-3.3-70b-instruct 128K 32.768K $0.71 $0.71 Tools Docs ↗
watsonx.ai meta-llama/llama-3-3-70b-instruct 131.072K 4.096K $0.753 $0.753 Tools Docs ↗
Eden AI ionos/meta-llama/Llama-3.3-70B-Instruct 128K 4.096K $0.755 $0.755 Not listed Docs ↗
Pioneer meta-llama/Llama-3.3-70B-Instruct 16.384K 16.384K $0.9 $0.9 $0.9 Tools Docs ↗
Scaleway llama-3.3-70b-instruct 100K 16.384K $0.9 $0.9 Tools Docs ↗
Eden AI scaleway/llama-3.3-70b-instruct 128K 4.096K $1.046 $1.046 Tools Docs ↗
evroc nvidia/Llama-3.3-70B-Instruct-FP8 128K 4.096K $1.15 $1.15 Tools Docs ↗
GreenPT llama-3.3-70b-instruct 100K 16.384K $1.254 $1.254 Tools Docs ↗
Tinfoil llama3-3-70b 131.072K 4.096K $1.75 $2.75 Tools Docs ↗

Capability badges appear only where the provider catalog explicitly lists support. A blank cell means the source is silent, not that the feature is absent.

Listed rates

Price across providers

Input price per million tokens as published by each provider. Bars are drawn from listed rates only — no traffic weighting, since the catalog observes no requests.

Lowest input $0.1/M

Across 22 priced providers

Median input $0.598/M

Midpoint of listed rates

Highest input $1.75/M

17.5× the lowest listed rate

Output range $0.32 – $2.75

Per million output tokens

Kilo Gateway $0.1/M
Eden AI $0.1/M
OpenRouter $0.1/M
Cortecs $0.129/M
Eden AI $0.13/M
DevPass (LLM Gateway) $0.135/M
Merge Gateway $0.22/M
Crusoe $0.25/M
Neon $0.5/M
Abacus $0.59/M
Hugging Face $0.59/M
Charm Hyper $0.607/M
Azure $0.71/M
Azure Cognitive Services $0.71/M
watsonx.ai $0.753/M
Eden AI $0.755/M
Pioneer $0.9/M
Scaleway $0.9/M
Eden AI $1.046/M
evroc $1.15/M
GreenPT $1.254/M
Tinfoil $1.75/M
Cost calculator

Estimate a workload

$0.00
Excludes taxes, non-token charges, and tiered discounts.

3 providers list the identical $0.1 input rate, so price alone will not separate them — compare context limits, max output, and capabilities above.

Context limits also differ by provider, from 16.384K to 131.072K tokens. Compare the provider table above before choosing on price alone.

Specification

Capabilities

Recorded from the source catalog and provider listings.

× Reasoning No
Tool calling Yes
? Structured output Unknown
Attachments Yes
× Vision input No
Open weights Yes
Creator
Meta
Model family
llama
Knowledge cutoff
2023-12
License
Not documented
Release date
2024-12-06
Model ID
meta/llama-3.3-70b-instruct

Weights: Hugging Face ↗

Published evaluations

Benchmarks

Every result stays attached to its source, version, metric and harness. Scores from different versions are never merged, and the profile below plots each benchmark against its own population rather than on a shared scale.

Benchmark registry
Benchmark profile
SciCodeSciCode AA Coding IndexArtificial Analysis Coding Index Terminal-Bench…Terminal-Bench Hard This model — SciCode: 26 percent correct (20.4th percentile) This model — Artificial Analysis Coding Index: 10.7 index (7.7th percentile) This model — Terminal-Bench Hard: 3 success rate (2.3th percentile)

Each axis is this model's percentile among the 3 benchmarks it has published results for, measured against every other model with a score on that same benchmark. Percentiles are used because benchmarks do not share a scale — a 60 on one is not a 60 on another. Hover any point for the raw score.

SciCode percent correct · 2026-03-11 Source ↗
SciCode #22 of 27
Artificial Analysis Coding Index index · 2026-03-11 Source ↗
Artificial Analysis Coding Index #24 of 26
Terminal-Bench Hard success rate · 2026-03-11 Source ↗
Terminal-Bench Hard #22 of 22

Each strip shows every published score for that benchmark, with this model marked. The lighter shape behind the ticks is the density of results, and the dashed line is the median.

Cost against capability

Price and performance

Listed input price plotted against SciCode, the benchmark with the widest published coverage that this model appears in.

0 23 45 $0.1 $1 $10 Qwen3-Coder 30B-A3B Instruct — $0.45/M, score 27.8 Qwen3 Max — $1.2/M, score 38.3 DeepSeek-R1 — $0.7/M, score 35.7 Gemini 2.5 Flash — $0.3/M, score 39.4 Gemini 2.5 Flash-Lite — $0.1/M, score 19.3 Gemini 2.5 Pro — $1.25/M, score 42.8 Llama-3.3-70B-Instruct — $0.1/M, score 26.0 Devstral 2 — $0.4/M, score 33.1 Mistral Large 2.1 — $2/M, score 29.2 Mistral Large 3 — $0.5/M, score 36.2 Mistral Medium 3 — $0.4/M, score 33.1 Mistral Small 4 — $0.15/M, score 38.0 GPT-4 Turbo — $10/M, score 31.9 GPT-4o (2024-05-13) — $5/M, score 30.9 GPT-4o (2024-08-06) — $2.5/M, score 33.1 GPT-4o (2024-11-20) — $2.5/M, score 33.3 GPT-4o mini — $0.15/M, score 22.9 GPT-5-Codex — $1.1/M, score 40.9 Sonar — $1/M, score 22.9 Sonar Pro — $3/M, score 22.6 Step 3.5 Flash — $0.1/M, score 40.4 Step 3.5 Flash 2603 — $0.1/M, score 38.5 Step 3.7 Flash — $0.185/M, score 40.0 GLM-4.5 — $0.6/M, score 34.8 GLM-4.5-Air — $0.2/M, score 30.6 GLM-4.5V — $0.6/M, score 22.1 GLM-4.6 — $0.6/M, score 38.4 Llama-3.3-70B-Instruct Input price per million tokens (log scale) Percent correct

The stepped line is the efficient frontier: at each price, the best score available for that money or less. A model sitting on it is not being beaten by anything cheaper. Price is log-scaled because listed rates span four orders of magnitude. Only models with both a listed price and a score on this benchmark can appear.

Catalog activity

Change log

Field-level changes detected between successful source imports.

Full change log
Price Completion0.4 → 0.32
Price Prompt0.13 → 0.1
Price Completion0.32 → 0.4
Price Prompt0.1 → 0.13
Provenance

Sources & verification

Every figure on this page traces back to one of these records.

Methodology
Public API

Use this record

Fetch the complete source-linked model record. No key, no account, no rate-limited tier.

API documentation
Endpoint
GET https://model.kyssta.lol/api/v1/models/meta/llama-3.3-70b-instruct
curl
curl "https://model.kyssta.lol/api/v1/models/meta/llama-3.3-70b-instruct"
Common questions

Frequently asked questions

Answered directly from the stored record — nothing here is generated beyond the catalog's own fields.

What is Llama-3.3-70B-Instruct?

Popular open Llama workhorse for multilingual chat, coding, and self-hosting. It is published by Meta and catalogued here from Models.dev.

How much does Llama-3.3-70B-Instruct cost?

Listed input pricing starts at $0.1 per million tokens from Kilo Gateway, rising to $1.75 across 22 listed providers.

What is the context length of Llama-3.3-70B-Instruct?

Llama-3.3-70B-Instruct accepts up to 128K tokens of context and returns up to 4.096K output tokens.

Does Llama-3.3-70B-Instruct support tool calling and structured output?

Provider catalogs list support for tool calling.

Which providers serve Llama-3.3-70B-Instruct?

25 providers list this model: Nvidia, Snowflake Cortex, Vercel AI Gateway, Pendra, GitHub Models, Llama and 19 more.

Are the weights for Llama-3.3-70B-Instruct open?

Yes. The weights are published and downloadable from Hugging Face.

When was Llama-3.3-70B-Instruct released?

The catalog records a release date of 2024-12-06, last verified Sep 11, 2026.

More models from Meta

View all →