DeepSeek V4 Pro 跑分曝光:力压 Kimi 2.6 夺得 Livebench 开源模型冠军
据 Livebench 最新基准测试数据显示,DeepSeek V4 Pro 模型表现抢眼,成功超越 Kimi 2.6 荣登现阶段开源模型榜首。虽然在编程(Coding)单项指标上略显不足,但综合性能已处于行业领先地位。这一成绩标志着国产开...
标签索引
这个标签下有 3 篇文章。按时间回看相关判断与实践记录。
标签精选
据 Livebench 最新基准测试数据显示,DeepSeek V4 Pro 模型表现抢眼,成功超越 Kimi 2.6 荣登现阶段开源模型榜首。虽然在编程(Coding)单项指标上略显不足,但综合性能已处于行业领先地位。这一成绩标志着国产开...
DeepSeek V3.2 ranks 3rd in data analysis in Livebench tests, showing strong performance against leading AI models like C...
DeepSeek V3.2 ranks 3rd in data analysis on Livebench benchmark, competing with Claude, Gemini, and GPT models.