已验证70% 置信事实精确时间
DeepSeek V4 Flash uses a Mixture of Experts (MOE) architecture with a total parameter count exceeding 600 billion (600B+).
4
来源数
70%
置信度
中期 (~90 天)
时效性
2026/7/2
首次发现
有效期至:2026/9/30
来源
相关事实
待验证In DeepSeek V4 Flash's MOE architecture, each token is routed to only a few experts for computation despite the model containing hundreds of expert modules.74% 相似待验证在DeepSeek V4 Flash部署中DiSpark使单用户生成速度提升60%到85%,Pro系列提升57%到78%,并保持同等推理质量72% 相似待验证DeepSeek在4月底开源了V4 Flash版本,参数量为2840亿,采用MIT协议72% 相似待验证盘古2.0 Flash采用MoE(混合专家)架构,参数规模为92B A6B(总参数92B,激活参数6B)71% 相似待验证V4-Pro采用混合专家(MoE)架构,总参数量达到1.6万亿(1.6T),活跃参数为490亿(49B)71% 相似
引用此条事实
Stable URI
https://kongchang.com/claim/51718API
curl https://kongchang.com/api/v1/knowledge/claims/51718MCP
get_claim(id=51718)