当前位置:首页 >英文主页 >中英对照 > 报告详情

DeepSeek:2025 DeepSeek-V3.2技术报告(英文版)(23页).pdf

上传人: 1****1 编号:981090 2025-12-03 23页 885.83KB

下载:

1、DeepSeek-V3.2:Pushing the Frontier of OpenLarge Language ModelsDeepSeek-AIAbstractWe introduce DeepSeek-V3.2,a model that harmonizes high computational efficiency with supe-rior reasoning and agent performance.The key technical breakthroughs of DeepSeek-V3.2 are asfollows:(1)DeepSeek Sparse Attentio

2、n(DSA):We introduce DSA,an efficient attention mecha-nism that substantially reduces computational complexity while preserving model performancein long-context scenarios.(2)Scalable Reinforcement Learning Framework:By implementinga robust reinforcement learning protocol and scaling post-training com

3、pute,DeepSeek-V3.2performs comparably to GPT-5.Notably,our high-compute variant,DeepSeek-V3.2-Speciale,surpasses GPT-5 and exhibits reasoning proficiency on par with Gemini-3.0-Pro,achievinggold-medal performance in both the 2025 International Mathematical Olympiad(IMO)and theInternational Olympiad

4、in Informatics(IOI).(3)Large-Scale Agentic Task Synthesis Pipeline:To integrate reasoning into tool-use scenarios,we developed a novel synthesis pipeline thatsystematically generates training data at scale.This methodology facilitates scalable agenticpost-training,yielding substantial improvements i

5、n generalization and instruction-followingrobustness within complex,interactive environments.AIME 2025(Pass1)HMMT 2025(Pass1)HLE(Pass1)Codeforces(Rating)SWEVerified(Resolved)TerminalBench 2.0(Acc)2Bench(Pass1)ToolDecathlon(Pass1)020406080100Accuracy/Pass1(%)96.093.194.687.095.099.290.288.379.297.530

6、.625.126.313.737.72701238625371480270873.174.977.276.246.435.242.854.280.380.284.785.435.229.038.636.4Reasoning CapabilitiesAgentic Capabilities050010001500200025003000Codeforces RatingDeepSeek-V3.2-SpecialeDeepSeek-V3.2-ThinkingGPT-5-HighClaude-4.5-SonnetGemini-3.0-ProFigure 1|Benchmark of DeepSeek

word格式文档无特别注明外均可编辑修改,预览文件经过压缩,下载原文更清晰!
三个皮匠报告文库所有资源均是客户上传分享,仅供网友学习交流,未经上传用户书面授权,请勿作商用。
DeepSeek-V3.2 是一种高效且强大的大型语言模型,它在推理和代理性能方面取得了显著突破。以下是全文关键点: 1. **DeepSeek Sparse Attention (DSA)**: 引入DSA,显著降低计算复杂度,同时保持长上下文场景下的模型性能。 2. **可扩展强化学习框架**:通过实施稳健的强化学习协议和扩展后训练计算,DeepSeek-V3.2 在推理任务上与GPT-5相当,其高性能变体DeepSeek-V3.2-Speciale甚至超越了GPT-5。 3. **大规模代理任务合成管道**:开发了一种新的合成管道,系统地生成大量训练数据,提高模型在复杂交互环境中的泛化能力和指令遵循鲁棒性。 4. **性能表现**:DeepSeek-V3.2 在多个基准测试中表现出色,包括MMLU-Pro、GPQA Diamond、HLE、LiveCodeBench等,与GPT-5和Gemini-3.0-Pro相当。 5. **DeepSeek-V3.2-Speciale**:在2025年国际数学奥林匹克(IMO)和国际信息学奥林匹克(IOI)中取得金牌成绩,证明了其在数学和编码领域的强大能力。
效率与推理的完美结合" "超越GPT-5,开源模型新突破!" 如何提升AI推理能力?"
客服
商务合作
小程序
服务号
折叠