Back
AI summary
Written by AI from the official notes. Check them for exact details.DeepSeek has upgraded its models to DeepSeek V2.5, merging chat and coding capabilities.
- DeepSeek V2.5 merges deepseek-coder and deepseek-chat models.
- New model shows improved performance in writing and instruction tasks.
- Code generation capabilities optimized for common programming scenarios.
- HumanEval score reaches 89% in code generation tests.
Why it matters: Developers and users of AI coding tools should consider upgrading for enhanced performance.
Full release notes6 changes
The DeepSeek V2 Chat and DeepSeek Coder V2 models have been merged and upgraded into the new model, DeepSeek V2.5.
For backward compatibility, API users can access the new model through either deepseek-coder or deepseek-chat.
The new model significantly surpasses the previous versions in both general capabilities and code abilities.
The new model better aligns with human preferences and has been optimized in various areas such as writing tasks and instruction following:
- ArenaHard win rate improved from 68.3% to 76.3%
- AlpacaEval 2.0 LC win rate increased from 46.61% to 50.52%
- MT-Bench score rose from 8.84 to 9.02
- AlignBench score increased from 7.88 to 8.04
The new model has further enhanced its code generation capabilities based on the original Coder model, optimized for common programming application scenarios, and achieved the following results on the standard test set:
- HumanEval: 89%
- LiveCodeBench (January-September): 41%