[KongchangAI]
Unverified50% confidenceBenchmarkExact time

2024-2025年间,Claude 3.5 Sonnet、GPT-4o、Gemini 2.0等模型在HumanEval、SWE-bench等编程基准测试的通过率从不足50%提升至90%以上

1
Sources
50%
Confidence
Long-term
Relevance
8/23/2026
First Seen

Sources

Related Entities

Related Claims

Cite This Claim

Stable URI
https://kongchang.com/claim/789977
API
curl https://kongchang.com/api/v1/knowledge/claims/789977
MCP
get_claim(id=789977)