Deepseek: Deepseek Chat V3.1
deepseek/deepseek-chat-v3.1
DeepSeek-V3.1 is a large hybrid reasoning model (671B parameters, 37B active) that supports both thinking and non-thinking modes via prompt templates. It extends the DeepSeek-V3 base with a two-phase long-context training process, reaching up to 128K tokens, and uses FP8 microscaling for efficient inference. Users can control the reasoning behaviour with the reasoning enabled boolean. Learn more in our docs The model improves tool use, code generation, and reasoning efficiency, achieving performance comparable to DeepSeek-R1 on difficult benchmarks while responding more quickly. It supports structured tool calling, code agents, and search agents, making it suitable for research, coding, and agentic workflows. It succeeds the DeepSeek V3-0324 model and performs well on a variety of tasks.
Model specifications
- Input
- text
- Output
- text
- Context
- 131,072 tokens
- Max output
- 131,072 tokens
- Input price
- $0.21 / 1M tokens
- Output price
- $0.79 / 1M tokens
- Released
- 2026-02-18
Capabilities
- Streaming
- Function calling
- JSON mode
- Playground
Provider pricing, discounts and data privacy
Compare effective provider prices, published discounts, regions, retention policies, training use, compliance, and privacy links by service tier.
Standard service tier
1 available provider · tier input average $0.21 / 1M tokens · tier output average $0.79 / 1M tokens
DeepInfra
Tier: Standard · Region: US
Pricing
- Input
- $0.21 / 1M tokens
- Output
- $0.79 / 1M tokens
- Cache read
- $0.13 / 1M tokens
No provider discount is currently published.
Data privacy and compliance
- Region
- US
- Zero data retention
- Yes
- Data retention
- Zero retention
- Used for training
- No
- Data collection
- Moderated
- No
- GDPR compliant
- Yes
- HIPAA compliant
- No
- SOC 2 certified
- Yes
- BYOK supported
- Yes
Privacy policy · Terms · Official website · Documentation · Status · Support
Frequently asked questions
- What is Deepseek: Deepseek Chat V3.1?
- DeepSeek-V3.1 is a large hybrid reasoning model (671B parameters, 37B active) that supports both thinking and non-thinking modes via prompt templates. It extends the DeepSeek-V3 base with a two-phase long-context training process, reaching up to 128K tokens, and uses FP8 microscaling for efficient inference. Users can control the reasoning behaviour with the reasoning enabled boolean. Learn more in our docs The model improves tool use, code generation, and reasoning efficiency, achieving performance comparable to DeepSeek-R1 on difficult benchmarks while responding more quickly. It supports structured tool calling, code agents, and search agents, making it suitable for research, coding, and agentic workflows. It succeeds the DeepSeek V3-0324 model and performs well on a variety of tasks.
- How much does Deepseek: Deepseek Chat V3.1 cost?
- Input costs start at $0.21 / 1M tokens and output costs start at $0.79 / 1M tokens. Provider-level prices vary by service tier.
- What is the context length of Deepseek: Deepseek Chat V3.1?
- Deepseek: Deepseek Chat V3.1 supports a 131,072 token context window and up to 131,072 output tokens.
- What capabilities does Deepseek: Deepseek Chat V3.1 support?
- Deepseek: Deepseek Chat V3.1 supports Streaming, Function calling, JSON mode, Playground.
- Which providers offer Deepseek: Deepseek Chat V3.1?
- Deepseek: Deepseek Chat V3.1 is available from DeepInfra.
- How do providers handle data privacy for Deepseek: Deepseek Chat V3.1?
- 1 of 1 providers report that customer data is not used for training, and 1 offer zero-data-retention routing. Retention, compliance, and privacy-policy links are listed per provider.