待验证50% 置信事实精确时间
MiMo-V2.6的Pro和Flash两个版本都具备处理文本、图像、音频和视频的全模态能力
1
来源数
50%
置信度
中期 (~90 天)
时效性
2026/9/22
首次发现
有效期至:2026/12/21
来源
涉及实体
相关事实
待验证MiMo-V2.6 覆盖文本、图像、语音等多种模态,是全模态模型77% 相似待验证LTX 2.5支持文生视频、图生视频、视频转视频,最新版能用音频生成视频68% 相似待验证千问(Qwen-VL系列)和小米MiMo支持多模态,可以处理图像输入67% 相似待验证MIMO 2.5 is a multimodal model focused on visual understanding, with core capabilities in Image Captioning, Visual Question Answering (VQA), and OCR recognition.66% 相似待验证推荐的融合工作流使用REF2VA负责主视频生成以确保视觉质量,使用FL2VA负责音频精修以获得更干净的声音66% 相似
引用此条事实
Stable URI
https://kongchang.com/claim/940603API
curl https://kongchang.com/api/v1/knowledge/claims/940603MCP
get_claim(id=940603)