We introduce
DeepSeek-V3.2
, a model that harmonizes high computational efficiency with superior reasoning and agent performance. Our approach is built upon three key technical breakthroughs:
DeepSeek Sparse Attention (DSA):
We introduce DSA, an efficient attention mechanism that substantially reduces computational complexity while preserving model performance, specifically optimized for long-context scenarios.
Scalable Reinforcement Learning Framework:
By implementing a robust RL protocol and scaling post-training compute,
DeepSeek-V3.2
performs comparably to GPT-5. Notably, our high-compute variant,
DeepSeek-V3.2-Speciale
,
surpasses GPT-5
and exhibits reasoning proficiency on par with Gemini-3.0-Pro.
Achievement:
🥇
Gold-medal performance
in the 2025 International Mathematical Olympiad (IMO) and International Olympiad in Informatics (IOI).
Large-Scale Agentic Task Synthesis Pipeline:
To integrate
reasoning into tool-use
scenarios, we developed a novel synthesis pipeline that systematically generates training data at scale. This facilitates scalable agentic post-training, improving compliance and generalization in complex interactive environments.
We have also released the final submissions for IOI 2025, ICPC World Finals, IMO 2025 and CMO 2025, which were selected based on our designed pipeline. These materials are provided for the community to conduct secondary verification. The files can be accessed at
assets/olympiad_cases
.
Chat Template
DeepSeek-V3.2 introduces significant updates to its chat template compared to prior versions. The primary changes involve a revised format for tool calling and the introduction of a "thinking with tools" capability.
To assist the community in understanding and adapting to this new template, we have provided a dedicated
encoding
folder, which contains Python scripts and test cases demonstrating how to encode messages in OpenAI-compatible format into input strings for the model and how to parse the model's text output.
This release does not include a Jinja-format chat template. Please refer to the Python code mentioned above.
The output parsing function included in the code is designed to handle well-formatted strings only. It does not attempt to correct or recover from malformed output that the model might occasionally generate. It is not suitable for production use without robust error handling.
A new role named
developer
has been introduced in the chat template. This role is dedicated exclusively to search agent scenarios and is designated for no other tasks. The official API does not accept messages assigned to
developer
.
How to Run Locally
The model structure of DeepSeek-V3.2 and DeepSeek-V3.2-Speciale are the same as DeepSeek-V3.2-Exp. Please visit
DeepSeek-V3.2-Exp
repo for more information about running this model locally.
Usage Recommendations:
For local deployment, we recommend setting the sampling parameters to
temperature = 1.0, top_p = 0.95
.
Please note that the DeepSeek-V3.2-Speciale variant is designed exclusively for deep reasoning tasks and does not support the tool-calling functionality.
License
This repository and the model weights are licensed under the
MIT License
.
Citation
@misc{deepseekai2025deepseekv32,
title={DeepSeek-V3.2: Pushing the Frontier of Open Large Language Models},
author={DeepSeek-AI},
year={2025},
}
Contact
If you have any questions, please raise an issue or contact us at
[email protected]
.
Runs of deepseek-ai DeepSeek-V3.2-Speciale on huggingface.co
5.2K
Total runs
0
24-hour runs
50
3-day runs
65
7-day runs
3.6K
30-day runs
More Information About DeepSeek-V3.2-Speciale huggingface.co Model
DeepSeek-V3.2-Speciale huggingface.co is an AI model on huggingface.co that provides DeepSeek-V3.2-Speciale's model effect (), which can be used instantly with this deepseek-ai DeepSeek-V3.2-Speciale model. huggingface.co supports a free trial of the DeepSeek-V3.2-Speciale model, and also provides paid use of the DeepSeek-V3.2-Speciale. Support call DeepSeek-V3.2-Speciale model through api, including Node.js, Python, http.
DeepSeek-V3.2-Speciale huggingface.co is an online trial and call api platform, which integrates DeepSeek-V3.2-Speciale's modeling effects, including api services, and provides a free online trial of DeepSeek-V3.2-Speciale, you can try DeepSeek-V3.2-Speciale online for free by clicking the link below.
deepseek-ai DeepSeek-V3.2-Speciale online free url in huggingface.co:
DeepSeek-V3.2-Speciale is an open source model from GitHub that offers a free installation service, and any user can find DeepSeek-V3.2-Speciale on GitHub to install. At the same time, huggingface.co provides the effect of DeepSeek-V3.2-Speciale install, users can directly use DeepSeek-V3.2-Speciale installed effect in huggingface.co for debugging and trial. It also supports api for free installation.
DeepSeek-V3.2-Speciale install url in huggingface.co: