待验证50% 置信事实精确时间
The comparison article evaluates Gemini 2.5 Pro 0605 against OpenAI o3, Claude Opus 4, Qwen3 235B, and Claude 3.5 Haiku across coding, reasoning, and writing dimensions.
1
来源数
50%
置信度
中期 (~90 天)
时效性
2026/7/2
首次发现
有效期至:2026/9/30
来源
相关事实
待验证Qwen3-235B在基准测试中超过了Grok 3 Beta、Gemini 2.5 Pro、OpenAI的o3-mini和o1模型72% 相似待验证OpenAI O3, O4 Mini, O3 Mini, Gemini 2.5 Pro, and Claude 3.7 were compared in a coding benchmark using identical prompts and conditions.72% 相似待验证Qwen 3.6 35B A3B在长上下文记忆基准测试中相比Gemma 4和Qwen 3.5具有明显优势69% 相似待验证Qwen3(千问3)在Ollama上提供0.6B/1.7B/4B/8B等多种规格67% 相似待验证Qwen2.5-3B-Instruct、Phi-3.5-mini(3.8B)和Llama-3.2-3B-Instruct是2024年末在小参数规模上综合表现最优的三款模型66% 相似
引用此条事实
Stable URI
https://kongchang.com/claim/58526API
curl https://kongchang.com/api/v1/knowledge/claims/58526MCP
get_claim(id=58526)