تخطَّ إلى المحتوى الرئيسي
رائد · مجانيحاسبة VRAM، أي GPU يشغّل أي نموذج؟ →
مميز حزمة نماذج الذكاء الاصطناعي والأزياء →
Vantaige

Free Tool

Meta (Llama) API Cost Calculator

Estimate your monthly and annual Meta (Llama) Llama API cost from your request volume and average token counts.

لست متأكدا من عدد الرموز التي تستخدمها موجهاتك؟ الصق نصك في حاسبة الرموز للحصول على العدد الدقيق أولا.

المعادلة: التكلفة الشهرية = الطلبات الشهرية × ((متوسط توكنات الإدخال × سعر الإدخال) + (متوسط توكنات الإخراج × سعر الإخراج)) / 1,000,000

الأرخص

Meta-Llama-3-8B

$0.1/M إدخال · $0.1/M إخراج

التكلفة الشهرية$1.20
التكلفة السنوية$14.40
التكلفة لكل طلب$0.00012

Llama-3-8b-chat-hf

$0.2/M إدخال · $0.2/M إخراج

التكلفة الشهرية$2.40
التكلفة السنوية$28.80
التكلفة لكل طلب$0.00024

Llama-4-Scout-17B-16E

$0.18/M إدخال · $0.59/M إخراج

التكلفة الشهرية$3.80
التكلفة السنوية$45.60
التكلفة لكل طلب$0.00038

Llama-4-Maverick-17B-128E

$0.27/M إدخال · $0.85/M إخراج

التكلفة الشهرية$5.56
التكلفة السنوية$66.72
التكلفة لكل طلب$0.000556

Meta (Llama) Llama pricing

Published list price per million tokens for Meta (Llama)'s current Llama models.

ModelInput $/M tokensOutput $/M tokens
Llama-4-Maverick-17B-128E$0.27$0.85
Llama-4-Scout-17B-16E$0.18$0.59
Llama-3-8b-chat-hf$0.2$0.2
Meta-Llama-3-8B$0.1$0.1
Llama-3-70b-chat-hf$0.9$0.9
Llama-3.3-70B$0.88$0.88
Llama-3.2-11B-Vision$0.18$0.18
Llama-3.2-3B$0.06$0.06

Frequently asked questions

How much does Meta (Llama)'s Llama-4-Maverick-17B-128E cost?

Llama-4-Maverick-17B-128E costs $0.27 per million input tokens and $0.85 per million output tokens. Use the calculator above to estimate your own monthly cost from your expected request volume and average token counts.

What is the cheapest Meta (Llama) model?

Llama-3.2-3B is Meta (Llama)'s cheapest current Llama model, priced at $0.06 per million input tokens and $0.06 per million output tokens, versus $0.27 / $0.85 for the flagship Llama-4-Maverick-17B-128E.

How does Meta (Llama) API pricing work?

Meta (Llama) bills per token, separately for input tokens (what you send) and output tokens (what the model generates), priced per million tokens. Output tokens typically cost more than input tokens because generation is more compute-intensive. Total request cost is (input tokens × input price + output tokens × output price) / 1,000,000, multiplied by your request volume.

How do I estimate my monthly Meta (Llama) API cost?

Enter your expected monthly request volume and average input/output tokens per request into the calculator above, pick one or more Llama models to compare, and it computes monthly cost, annual cost, cost per request, and cost per user instantly. Add an optional monthly growth rate to see a 12-month cost projection.

Compare across providers

Want to compare Meta (Llama) against other providers side by side? Use the full AI Cost Calculator.