待验证60% 置信事实精确时间
GLM 5.2 achieved approximately 1300 ELO points on the Design Arena leaderboard, surpassing Claude 3.5.
2
来源数
60%
置信度
中期 (~90 天)
时效性
2026/7/2
首次发现
有效期至:2026/9/30
来源
相关事实
待验证GLM4在Design Arena排行榜上排名第一,得分约为1300分,超过Claude 3.5和Gemini 1.5 Pro72% 相似待验证GLM 5.2在Terminal Bench测试中得分81分,GLM 5.1得分63.5分,GPT 5.5得分84分,Opus 4.8得分85分61% 相似待验证GLM 5.2 scored 74.4% on the Frontier SW benchmark, closely approaching Anthropic's Opus 4.8.58% 相似待验证测试中GLM 5.1综合得分89.3分,排名第一,超过Claude Sonnet 4.6的87.2分58% 相似待验证Grok topped the Arena leaderboard rankings, representing a significant milestone for xAI's model.57% 相似
引用此条事实
Stable URI
https://kongchang.com/claim/40550API
curl https://kongchang.com/api/v1/knowledge/claims/40550MCP
get_claim(id=40550)