deepseek-chat upgraded to DeepSeek-V3-0324, stronger reasoning and code
deepseek-chat
deepseek-chat has been upgraded to DeepSeek-V3-0324, with gains in reasoning, front-end development, Chinese writing and Function Calling. MMLU-Pro rose from 75.9 to 81.2, GPQA from 59.1 to 68.4, AIME from 39.6 to 59.4 and LiveCodeBench from 39.2 to 49.2.
Why it matters: DeepSeek-V3-0324 improved on all four benchmarks, and its AIME gain of 19.8 is larger than the gains on the knowledge and code benchmarks.