DeepSeek V3.2 Livebench Test Rankings Revealed

DeepSeek V3.2 has released its latest results in the Livebench benchmark, providing a comprehensive comparison with leading AI models in the industry such as Claude 4.5 Opus Thinking, Gemini 3 Pro Preview, and GPT-5. The test results show that V3.2 ranked ninth in reasoning tasks, sixteenth in programming ability, fourteenth in agent programming capability, tenth in mathematical ability, and demonstrated outstanding performance in data analysis, ranking third. These data points reflect the rapid iteration of current AI technology and intense competition among models, offering valuable reference for AI professionals, researchers, and developers to evaluate the performance advantages of different models and drive the advancement of artificial intelligence technology. The test results also highlight DeepSeek’s competitiveness in specific domains, particularly its strong performance in data analysis.

Original Link:Linux.do

C code80.ai · AI 编码 API 聚合 Claude / GPT 多模型统一接入,稳定不限速,按量计费,几行配置接入 Claude Code。 了解一下 ›

抢沙发

评论前必须登录!

立即登录   注册