DeepSeek: DeepSeek V3.1 Terminus
deepseek/deepseek-v3.1-terminus
Description
DeepSeek-V3.1 Terminus is an update to DeepSeek V3.1 that maintains the model's original capabilities while addressing issues reported by users, including language consistency and agent capabilities, further optimizing the model's performance in coding and search agents. It is a large hybrid reasoning model (671B parameters, 37B active) that supports both thinking and non-thinking modes. It extends the DeepSeek-V3 base with a two-phase long-context training process, reaching up to 128K tokens, and uses FP8 microscaling for efficient inference. Users can control the reasoning behaviour with the reasoning enabled boolean.
The model improves tool use, code generation, and reasoning efficiency, achieving performance comparable to DeepSeek-R1 on difficult benchmarks while responding more quickly. It supports structured tool calling, code agents, and search agents, making it suitable for research, coding, and agentic workflows.
How this model compares
Overall covers the full catalog. By plan covers only models available on that tier (same rules as available models in your list). Position on min–average–max. Prices use the higher of prompt or completion per token, shown per 1M tokens.
Price (per 1M tokens)
Min
Max
This model
339 models in this groupPrice (per 1M tokens)
- Min
- $0.04
- Avg
- $12.395447
- Max
- $750.00
This model: $0.95 / 1M tokens
Context length (tokens)
Min
Max
This model
339 models in this groupContext length (tokens)
- Min
- 4,095 tokens
- Avg
- 379,884.782 tokens
- Max
- 10,000,000 tokens
This model: 163,840 tokens
Capabilities
Text → TextContext: 163,840 tokens
Input:
Text
Output:
Text