Meta · live price
Llama 3.1 70B Instruct API pricing
Meta's latest class of model (Llama 3.1) launched with a variety of sizes & flavors. This 70B instruct-tuned version is optimized for high quality dialogue usecases. It has demonstrated strong...
- Input
- $0.4
- per 1M tokens
- Output
- $0.4
- per 1M tokens
- Cached input
- –
- per 1M tokens
- Context
- 131K
- 16K max output
List price from OpenRouter, which passes Meta's price through without markup. Refreshed hourly. Released July 23, 2024. Model id meta-llama/llama-3.1-70b-instruct.
What Llama 3.1 70B Instruct costs in practice
1,000 chat messages
$0.60
1,000 input + 500 output tokens each
1,000 RAG questions
$3.40
8,000 input + 500 output tokens each
100 coding-agent steps
$1.28
30,000 input + 2,000 output tokens each, no caching
Llama 3.1 70B Instruct price history
Llama 3.1 70B Instruct price by host
- DeepInfra$0.4 in · $0.4 out · 131K ctx · 95.3% uptime 24h
- Amazon Bedrock$0.72 in · $0.72 out · 131K ctx · 100.0% uptime 24h
Where to buy Meta credits for less
Disclosure: this site is run by the team behind smaaart, one of the options compared. Tables are ranked by price only, from live quotes and published fees, including when smaaart isn't first. How we compare
Llama 3.1 70B Instruct cost calculator
Cheaper alternatives to Llama 3.1 70B Instruct
| Model | Input / 1M | Output / 1M |
|---|---|---|
| OpenAI: GPT-6 Luna Pro | $0.1 | $0.5 |
| OpenAI: GPT-6 Luna | $0.1 | $0.5 |
| Qwen: Qwen3.8 Omni Flash | $0.15 | $0.47 |
| DeepSeek: DeepSeek V4.1 Flash | $0.03 | $0.5 |
| Qwen: Qwen3.8 Flash | $0.15 | $0.47 |
| Z.ai: GLM 5.3 Flash | $0.15 | $0.5 |
Other Meta models
| Model | Input / 1M | Output / 1M |
|---|---|---|
| Llama 3.1 8B Instruct | $0.05 | $0.08 |
| Llama 3.2 1B Instruct | $0.027 | $0.201 |
| Llama 3.2 3B Instruct | $0.05 | $0.33 |
| Llama 3.3 70B Instruct | $0.1 | $0.32 |
| Llama 4 Scout | $0.1 | $0.3 |
| Llama 4 Maverick | $0.188 | $0.652 |