Meta: Llama 3.1 8B
meta-llama/llama-3.1-8b
Meta's Llama 3.1 is a major upgrade, introducing new model sizes. The flagship 405B model rivals top closed-source AI like GPT-4o, while a new, highly efficient 5B model runs on consumer hardware. Key improvements include a massive 128K context window for all models, enabling analysis of long documents, and significantly enhanced coding capabilities. By keeping the weights open, Llama 3.1 makes state-of-the-art AI more powerful and accessible than ever, from large-scale deployment to personal devices.
Model specifications
- Input
- text
- Output
- text
- Context
- 32,800 tokens
- Max output
- 8,200 tokens
- Input price
- $0.05 / 1M tokens
- Output price
- $0.08 / 1M tokens
- Released
- 2026-02-10
Capabilities
- Streaming
- Function calling
- JSON mode
- Playground
Provider pricing, discounts and data privacy
Compare effective provider prices, published discounts, regions, retention policies, training use, compliance, and privacy links by service tier.
Standard service tier
2 available providers · tier input average $0.075 / 1M tokens · tier output average $0.09 / 1M tokens
Groq
Tier: Standard · Region: US
Pricing
- Input
- $0.05 / 1M tokens
- Output
- $0.08 / 1M tokens
No provider discount is currently published.
Data privacy and compliance
- Region
- US
- Zero data retention
- No
- Data retention
- 30-day retention
- Used for training
- No
- Data collection
- Moderated
- No
- GDPR compliant
- No
- HIPAA compliant
- No
- SOC 2 certified
- No
- BYOK supported
- No
Privacy policy · Terms · Official website · Documentation · Status · Support
Cerebras
Tier: Standard · Region: US
Pricing
- Input
- $0.1 / 1M tokens
- Output
- $0.1 / 1M tokens
- Cache read
- $0.1 / 1M tokens
No provider discount is currently published.
Data privacy and compliance
- Region
- US
- Zero data retention
- Yes
- Data retention
- Zero retention
- Used for training
- No
- Data collection
- Moderated
- No
- GDPR compliant
- No
- HIPAA compliant
- No
- SOC 2 certified
- No
- BYOK supported
- No
Privacy policy · Terms · Official website · Documentation · Support
Frequently asked questions
- What is Meta: Llama 3.1 8B?
- Meta's Llama 3.1 is a major upgrade, introducing new model sizes. The flagship 405B model rivals top closed-source AI like GPT-4o, while a new, highly efficient 5B model runs on consumer hardware. Key improvements include a massive 128K context window for all models, enabling analysis of long documents, and significantly enhanced coding capabilities. By keeping the weights open, Llama 3.1 makes state-of-the-art AI more powerful and accessible than ever, from large-scale deployment to personal devices.
- How much does Meta: Llama 3.1 8B cost?
- Input costs start at $0.05 / 1M tokens and output costs start at $0.08 / 1M tokens. Provider-level prices vary by service tier.
- What is the context length of Meta: Llama 3.1 8B?
- Meta: Llama 3.1 8B supports a 32,800 token context window and up to 8,200 output tokens.
- What capabilities does Meta: Llama 3.1 8B support?
- Meta: Llama 3.1 8B supports Streaming, Function calling, JSON mode, Playground.
- Which providers offer Meta: Llama 3.1 8B?
- Meta: Llama 3.1 8B is available from Groq, Cerebras.
- How do providers handle data privacy for Meta: Llama 3.1 8B?
- 2 of 2 providers report that customer data is not used for training, and 1 offer zero-data-retention routing. Retention, compliance, and privacy-policy links are listed per provider.