DeepSeek: DeepSeek V3.1 Terminus (exacto)✓ Catalog verified
DeepSeek-V3.1 Terminus is an update to DeepSeek V3.1 that maintains the model's original capabilities while addressing issues reported by users, including language consistency and agent capabilities, further optimizing the model's performance in coding and search agents. It is a large hybrid reasoning model (671B parameters, 37B active) that supports both thinking and non-thinking modes. It extends the DeepSeek-V3 base with a two-phase long-context training process, reaching up to 128K tokens, and uses FP8 microscaling for efficient inference. Users can control the reasoning behaviour with the `reasoning` `enabled` boolean. Learn more in our docs The model improves tool use, code generation, and reasoning efficiency, achieving performance comparable to DeepSeek-R1 on difficult benchmarks while responding more quickly. It supports structured tool calling, code agents, and search agents, making it suitable for research, coding, and agentic workflows.
Overview
DeepSeek-V3.1 Terminus is an update to DeepSeek V3.1 that maintains the model's original capabilities while addressing issues reported by users, including language consistency and agent capabilities, further optimizing the model's performance in coding and search agents. It is a large hybrid reasoning model (671B parameters, 37B active) that supports both thinking and non-thinking modes. It extends the DeepSeek-V3 base with a two-phase long-context training process, reaching up to 128K tokens, and uses FP8 microscaling for efficient inference. Users can control the reasoning behaviour with the `reasoning` `enabled` boolean. Learn more in our docs The model improves tool use, code generation, and reasoning efficiency, achieving performance comparable to DeepSeek-R1 on difficult benchmarks while responding more quickly. It supports structured tool calling, code agents, and search agents, making it suitable for research, coding, and agentic workflows.
Access: available through the official DeepSeek API. Context window: 164K.
Specifications
| API identifier | deepseek/deepseek-v3.1-terminus:exacto |
| Provider | DeepSeek |
| Model type | Text LLM |
| Context window | 164K catalog |
| Input modalities | text |
| Output modalities | text |
| Released | Sep 22, 2025 |
| Tokenizer | DeepSeek |
| Instruction format | deepseek-v3.1 |
| Moderated | No |
| Architecture modality | text->text |
API defaults
{
"top_p": null,
"temperature": null,
"frequency_penalty": null
}Pricing
Live pricing components for deepseek/deepseek-v3.1-terminus:exacto as published in the catalog.
| Pricing component | Raw unit price | Normalized | Unit |
|---|---|---|---|
| Prompt tokens | $2.1e-7 | $0.21 / 1M | Per input token |
| Completion tokens | $7.9e-7 | $0.79 / 1M | Per output token |
| Request fee | N/A | N/A | Per request |
| Image fee | N/A | N/A | Per image unit |
| Web search fee | N/A | N/A | Per search request |
Cost calculator
Estimate monthly spend from your own token volumes.
Estimated Total Cost
$0.26Calculated from current list pricing in the ModelsAtlas catalog.Capabilities
| Capability | Status | What it means |
|---|---|---|
| Visual Understanding | Not advertised | Image and document analysis support |
| Audio Processing | Not advertised | Speech and voice aligned flows |
| Tool Calling | ✓ Supported | Supports tools / function calling |
| Self-Hosting | ✓ Supported | Deploy outside managed APIs |
Capability signals
Directional signals derived from model metadata and capability tags — not official benchmark submissions.
Quick start
Call DeepSeek: DeepSeek V3.1 Terminus (exacto) through an OpenAI-compatible client.
from openai import OpenAI
client = OpenAI(
base_url="https://openrouter.ai/api/v1",
api_key="YOUR_API_KEY"
)
response = client.chat.completions.create(
model="deepseek/deepseek-v3.1-terminus:exacto",
messages=[{"role": "user", "content": "Explain quantum physics."}]
)
print(response.choices[0].message.content)Alternatives
Closest models by context window from a different provider.
| Model | Provider | Context | Input /1M | Actions |
|---|---|---|---|---|
| Meta: Llama Guard 4 12B | Meta | 164K | $0.18 | View → |
| TNG: DeepSeek R1T2 Chimera | Tngtech | 164K | $0.30 | View → |
Sources & attribution
Pricing and metadata are maintained in the ModelsAtlas catalog.
Frequently asked questions
How much does DeepSeek: DeepSeek V3.1 Terminus (exacto) cost?
$0.21 per 1M input tokens and $0.79 per 1M output tokens on the official DeepSeek API.
What is the context window of DeepSeek: DeepSeek V3.1 Terminus (exacto)?
DeepSeek: DeepSeek V3.1 Terminus (exacto) supports up to 164K tokens of context.
Does DeepSeek: DeepSeek V3.1 Terminus (exacto) support tool / function calling?
Yes, DeepSeek: DeepSeek V3.1 Terminus (exacto) supports tool / function calling — you can register tools and the model will emit structured tool calls.
How do I access DeepSeek: DeepSeek V3.1 Terminus (exacto)?
Use the official DeepSeek API with the model id `deepseek/deepseek-v3.1-terminus:exacto`. See the Quick start section above for code examples.