金色财经
金色财经|12月 01, 2025 11:14
[DeepSeek V3.2 Official Release: Enhanced Agent Capabilities, Integrated Thinking and Reasoning] Reported by Jinse Finance, today we are releasing two official model versions simultaneously: DeepSeek-V3.2 and DeepSeek-V3.2-Speciale. DeepSeek-V3.2 is our first model that integrates thinking into tool usage, supporting both thinking mode and non-thinking mode for tool invocation. We have proposed a large-scale Agent training data synthesis method, constructing numerous 'hard-to-answer, easy-to-verify' reinforcement learning tasks (1800+ environments, 85,000+ complex instructions), significantly improving the model's generalization capabilities. (DeepSeek)
+4
Mentioned
Share To

Timeline

HotFlash

APP

X

Telegram

Facebook

Reddit

CopyLink

Hot Reads