已过期50% 置信事件精确时间
All four AI models (ChatGPT 5.4, Gemini 3.1, DeepSeek V4 Pro, and Kimi 5.1) failed in the first round of a dynamic web scraping task targeting Baidu
1
来源数
50%
置信度
中期 (~90 天)
时效性
2026/7/2
首次发现
有效期至:2026/9/30(已过期)
来源
相关事实
待验证DeepSeek V4 Pro and Kimi 5.1 operated in Solid mode (autonomous completion of the entire workflow without human intervention), while ChatGPT 5.4 and Gemini 3.1 operated in standard interactive mode67% 相似待验证此前的PokeAgent挑战证明,即使是GPT-4V、Gemini等顶级多模态模型,在没有专门设计的辅助框架时,几乎无法在宝可梦RPG中取得任何进展67% 相似待验证前端生成测试中还对比了疑似Gemini 3的refer模型、Kimi KL和Haiku 4.5,GPT 5.1在指令遵循和视觉质感上处于领先63% 相似待验证该AI大模型调用网站已接入Gemini、通义千问、DeepSeek、豆包四款模型62% 相似待验证AI行业中模型版本号的命名缺乏统一标准,Google的Gemini中间的.5版本实际上是重大架构升级61% 相似
引用此条事实
Stable URI
https://kongchang.com/claim/48432API
curl https://kongchang.com/api/v1/knowledge/claims/48432MCP
get_claim(id=48432)