Meta: Llama 3.1 8B

meta-llama/llama-3.1-8b

Meta's Llama 3.1 is a major upgrade, introducing new model sizes. The flagship 405B model rivals top closed-source AI like GPT-4o, while a new, highly efficient 5B model runs on consumer hardware. Key improvements include a massive 128K context window for all models, enabling analysis of long documents, and significantly enhanced coding capabilities. By keeping the weights open, Llama 3.1 makes state-of-the-art AI more powerful and accessible than ever, from large-scale deployment to personal devices.

Model specifications

Input
text
Output
text
Context
32,800 tokens
Max output
8,200 tokens
Input price
$0.05 / 1M tokens
Output price
$0.08 / 1M tokens
Released
2026-02-10

Capabilities

  • Streaming
  • Function calling
  • JSON mode
  • Playground

Provider pricing, discounts and data privacy

Compare effective provider prices, published discounts, regions, retention policies, training use, compliance, and privacy links by service tier.

Standard service tier

2 available providers · tier input average $0.075 / 1M tokens · tier output average $0.09 / 1M tokens

Groq

Tier: Standard · Region: US

Pricing
Input
$0.05 / 1M tokens
Output
$0.08 / 1M tokens

No provider discount is currently published.

Data privacy and compliance
Region
US
Zero data retention
No
Data retention
30-day retention
Used for training
No
Data collection
Moderated
No
GDPR compliant
No
HIPAA compliant
No
SOC 2 certified
No
BYOK supported
No

Cerebras

Tier: Standard · Region: US

Pricing
Input
$0.1 / 1M tokens
Output
$0.1 / 1M tokens
Cache read
$0.1 / 1M tokens

No provider discount is currently published.

Data privacy and compliance
Region
US
Zero data retention
Yes
Data retention
Zero retention
Used for training
No
Data collection
Moderated
No
GDPR compliant
No
HIPAA compliant
No
SOC 2 certified
No
BYOK supported
No

Frequently asked questions

What is Meta: Llama 3.1 8B?
Meta's Llama 3.1 is a major upgrade, introducing new model sizes. The flagship 405B model rivals top closed-source AI like GPT-4o, while a new, highly efficient 5B model runs on consumer hardware. Key improvements include a massive 128K context window for all models, enabling analysis of long documents, and significantly enhanced coding capabilities. By keeping the weights open, Llama 3.1 makes state-of-the-art AI more powerful and accessible than ever, from large-scale deployment to personal devices.
How much does Meta: Llama 3.1 8B cost?
Input costs start at $0.05 / 1M tokens and output costs start at $0.08 / 1M tokens. Provider-level prices vary by service tier.
What is the context length of Meta: Llama 3.1 8B?
Meta: Llama 3.1 8B supports a 32,800 token context window and up to 8,200 output tokens.
What capabilities does Meta: Llama 3.1 8B support?
Meta: Llama 3.1 8B supports Streaming, Function calling, JSON mode, Playground.
Which providers offer Meta: Llama 3.1 8B?
Meta: Llama 3.1 8B is available from Groq, Cerebras.
How do providers handle data privacy for Meta: Llama 3.1 8B?
2 of 2 providers report that customer data is not used for training, and 1 offer zero-data-retention routing. Retention, compliance, and privacy-policy links are listed per provider.

Browse all AI models