待验证50% 置信事实精确时间
SWBench is an AI coding evaluation benchmark developed by a Princeton University team that extracts issues and pull requests from real GitHub open-source projects.
1
来源数
50%
置信度
中期 (~90 天)
时效性
2026/7/2
首次发现
有效期至:2026/9/30
来源
相关事实
待验证SWE-bench(Software Engineering Benchmark)是目前衡量智能体编码能力的核心基准,从真实的GitHub Issue出发要求模型自主定位代码库中的问题并提交修复补丁77% 相似部分验证SWE-Bench was developed by a Princeton University team and extracts tasks from real GitHub issues, requiring models to locate problems within complete codebases and generate fix patches75% 相似部分验证SweetBench Pro (SWE-bench) is a coding benchmark developed by a Princeton University research team that uses real GitHub Issues and Pull Requests75% 相似
引用此条事实
Stable URI
https://kongchang.com/claim/51096API
curl https://kongchang.com/api/v1/knowledge/claims/51096MCP
get_claim(id=51096)