Keep pulling the thread on Liang Wenfeng.
The high-compute variant, DeepSeek-V3.2-Speciale, surpasses GPT-5 in performance.
The DeepSeek-V3.2-Speciale model exhibits reasoning proficiency on par with Gemini-3.0-Pro.
The DeepSeek-V3.2-Speciale model achieved a performance level equivalent to a gold medal in the 2025 International Mathematical Olympiad (IMO).
The DeepSeek-V3.2-Speciale model achieved a performance level equivalent to a gold medal in the 2025 International Olympiad in Informatics (IOI).
The DeepSeek-V3.2 model incorporates DeepSeek Sparse Attention (DSA), an efficient attention mechanism that reduces computational complexity while preserving performance in long-context scenarios.
DeepSeek-AI developed a novel synthesis pipeline to systematically generate training data at scale for agentic tasks.
The agentic task synthesis pipeline developed by DeepSeek-AI yields substantial improvements in generalization and instruction-following robustness within complex, interactive environments.