Skip to main content
FLAGSHIP · FREEVRAM Calculator, which GPU runs which model? →
Featured AI Models & Fashion Pack →
Vantaige

Free Tool

Meta (Llama) API Cost Calculator

Estimate your monthly and annual Meta (Llama) Llama API cost from your request volume and average token counts.

Not sure how many tokens your prompts use? Paste your text into the Token Calculator for exact counts first.

Formula: monthly cost = monthly requests × ((avg input tokens × input price) + (avg output tokens × output price)) / 1,000,000

Cheapest

Meta-Llama-3-8B

$0.1/M in · $0.1/M out

Monthly cost$1.20
Annual cost$14.40
Cost per request$0.00012

Llama-3-8b-chat-hf

$0.2/M in · $0.2/M out

Monthly cost$2.40
Annual cost$28.80
Cost per request$0.00024

Llama-4-Scout-17B-16E

$0.18/M in · $0.59/M out

Monthly cost$3.80
Annual cost$45.60
Cost per request$0.00038

Llama-4-Maverick-17B-128E

$0.27/M in · $0.85/M out

Monthly cost$5.56
Annual cost$66.72
Cost per request$0.000556

Meta (Llama) Llama pricing

Published list price per million tokens for Meta (Llama)'s current Llama models.

ModelInput $/M tokensOutput $/M tokens
Llama-4-Maverick-17B-128E$0.27$0.85
Llama-4-Scout-17B-16E$0.18$0.59
Llama-3-8b-chat-hf$0.2$0.2
Meta-Llama-3-8B$0.1$0.1
Llama-3-70b-chat-hf$0.9$0.9
Llama-3.3-70B$0.88$0.88
Llama-3.2-11B-Vision$0.18$0.18
Llama-3.2-3B$0.06$0.06

Frequently asked questions

How much does Meta (Llama)'s Llama-4-Maverick-17B-128E cost?

Llama-4-Maverick-17B-128E costs $0.27 per million input tokens and $0.85 per million output tokens. Use the calculator above to estimate your own monthly cost from your expected request volume and average token counts.

What is the cheapest Meta (Llama) model?

Llama-3.2-3B is Meta (Llama)'s cheapest current Llama model, priced at $0.06 per million input tokens and $0.06 per million output tokens, versus $0.27 / $0.85 for the flagship Llama-4-Maverick-17B-128E.

How does Meta (Llama) API pricing work?

Meta (Llama) bills per token, separately for input tokens (what you send) and output tokens (what the model generates), priced per million tokens. Output tokens typically cost more than input tokens because generation is more compute-intensive. Total request cost is (input tokens × input price + output tokens × output price) / 1,000,000, multiplied by your request volume.

How do I estimate my monthly Meta (Llama) API cost?

Enter your expected monthly request volume and average input/output tokens per request into the calculator above, pick one or more Llama models to compare, and it computes monthly cost, annual cost, cost per request, and cost per user instantly. Add an optional monthly growth rate to see a 12-month cost projection.

Compare across providers

Want to compare Meta (Llama) against other providers side by side? Use the full AI Cost Calculator.