DeepSeek V3.2 Thinking
DeepSeek
DeepSeek‑V3.2 from DeepSeek harmonizes high computational efficiency with superior reasoning and agent performance. It builds on three main techniques: DeepSeek Sparse Attention for long‑context efficiency, a scalable reinforcement learning framework, and a large‑scale agentic task synthesis pipeline. This model excels at long-context reasoning and agentic tasks, efficiently handling extended inputs while maintaining strong accuracy. Its sparse attention design enables it to process complex, multi-step workflows without excessive compute costs. Overall, DeepSeek‑V3.2 targets long‑context reasoning, tool‑using agents, and efficient deployment in production environments.
- Language
- Tool use
- Implicit caching
- Context
- 128K
- Max output
- 8K
- Input / 1M
- $0.620
- Output / 1M
- $1.85
Quickstart
Point any OpenAI-compatible client at Maxynetic Systems and pass deepseek/deepseek-v3.2-thinking as the model.
curl https://api.maxynetic.com/v1/chat/completions \
-H "Authorization: Bearer $MAXYNETIC_API_KEY" \
-H "Content-Type: application/json" \
-d '{
"model": "deepseek/deepseek-v3.2-thinking",
"messages": [{ "role": "user", "content": "Explain routing in one line." }]
}'Pricing
Billed per token with no markup on the underlying provider rate.
Input
Per 1M prompt tokens
$0.620
Output
Per 1M completion tokens
$1.85
A 1M-token prompt with 250K tokens of output costs about $1.08.
Rates for this model vary by upstream provider; Maxynetic Systems routes to the cheapest healthy one by default.
