待验证85% 置信事实时间未知
Anthros UD Q5 GGUF方案运行Qwen3.6-27B的生成速度约为18 tok/s
1
来源数
85%
置信度
中期 (~90 天)
时效性
2026/5/31
首次发现
有效期至:2026/8/29
来源
Mac本地跑Qwen3.6-27B:4种方案实测对比
bilibilikate人不错
涉及实体
相关事实
待验证Anthros MLX 6bit + Diflash方案运行Qwen3.6-27B的生成速度约为22 tok/s83% 相似待验证Anthros UD Q5 GGUF quantization achieves approximately 18 tokens per second generation speed on Mac74% 相似待验证实测中Qwen3 27B在未开启NVLink的4张魔改2080Ti上生成速度约为40多token/秒72% 相似待验证Qwen3.6 MTP-GGUF版本在单GPU上将35B-A3B模型的推理速度推到220 token/s,比原版GGUF快超过1.4倍,且精度零损失70% 相似待验证在24GB显存的显卡上运行Qwen3.6 27B模型,未经优化可达每秒约40个Token的生成速度,优化后可达50-60 Token/s68% 相似
引用此条事实
Stable URI
https://kongchang.com/claim/4332API
curl https://kongchang.com/api/v1/knowledge/claims/4332MCP
get_claim(id=4332)