待验证50% 置信事实精确时间
DeepSeek-Coder performs excellently on mainstream code evaluation benchmarks like HumanEval and MBPP.
1
来源数
50%
置信度
中期 (~90 天)
时效性
2026/7/2
首次发现
有效期至:2026/9/30
来源
5 Steps to Connect Codex with DeepSeek — No GPT Account Required
bilibili港大张磊AI2026/6/15
相关事实
已验证DeepSeek在多项代码基准测试(如HumanEval、MBPP、LiveCodeBench)中表现优异,部分指标接近甚至超越GPT-4级别模型77% 相似待验证Community benchmarks showed DeepSeek V3-0324 achieved significant improvements on mainstream code evaluation benchmarks including HumanEval and SWE-bench77% 相似待验证DeepSeek-Coder-V2 在 HumanEval 等代码评测榜上长期位居前列76% 相似待验证DeepSeek Coder在HumanEval、MBPP等代码基准测试上表现接近同参数规模的商业模型,且对中文注释场景理解能力更优76% 相似待验证DeepSeek-Coder在多项编程基准测试中达到国际领先水平75% 相似
引用此条事实
Stable URI
https://kongchang.com/claim/45671API
curl https://kongchang.com/api/v1/knowledge/claims/45671MCP
get_claim(id=45671)