Unverified50% confidenceFactExact time
SWBench is an AI coding evaluation benchmark developed by a Princeton University team that extracts issues and pull requests from real GitHub open-source projects.
1
Sources
50%
Confidence
Medium-term (~90 days)
Relevance
7/2/2026
First Seen
Valid until: 9/30/2026
Sources
Related Claims
UnverifiedSWE-bench(Software Engineering Benchmark)是目前衡量智能体编码能力的核心基准,从真实的GitHub Issue出发要求模型自主定位代码库中的问题并提交修复补丁77% similarPartially VerifiedSWE-Bench was developed by a Princeton University team and extracts tasks from real GitHub issues, requiring models to locate problems within complete codebases and generate fix patches75% similarPartially VerifiedSweetBench Pro (SWE-bench) is a coding benchmark developed by a Princeton University research team that uses real GitHub Issues and Pull Requests75% similar
Cite This Claim
Stable URI
https://kongchang.com/claim/51096API
curl https://kongchang.com/api/v1/knowledge/claims/51096MCP
get_claim(id=51096)