待验证50% 置信事实精确时间
最新最强的 Claude 模型(Opus 4.8 和 Sonnet 5)在调用 Pi 的自定义编辑工具时表现反而比更老的模型更差
1
来源数
50%
置信度
中期 (~90 天)
时效性
2026/7/7
首次发现
有效期至:2026/10/5
来源
模型越强工具越差?Claude新模型工具调用的隐藏陷阱
rss2026/7/4
相关事实
待验证Claude 旗舰级模型 Opus 4.8 和 Sonnet 5 在调用 Pi 的编辑工具时,会在嵌套的 edits[] 数组中凭空发明额外字段,导致参数结构不符合 schema82% 相似待验证The key turning point for Claude Code's performance improvement was the evolution of underlying models: Sonnet 4, Opus 4, and Opus 4.571% 相似待验证Sonnet 4.5是目前Claude最智能的模型71% 相似待验证Claude Code may default to calling a Sonnet-level model rather than Opus 467% 相似待验证Claude支持多个模型切换,其中Opus 4.5被认为是目前最强大的模型之一,尤其在代码生成和长文写作方面表现出色67% 相似
引用此条事实
Stable URI
https://kongchang.com/claim/129578API
curl https://kongchang.com/api/v1/knowledge/claims/129578MCP
get_claim(id=129578)