待验证85% 置信事实时间未知
Meta的Llama Guard系列是专门训练的安全分类模型,用于对LLM输出内容进行多维度安全评分
1
来源数
85%
置信度
长期有效
时效性
2026/6/1
首次发现
来源
LLM Guardrails Index:最全面的大模型安全护栏评估体系详解
twitterguardrails_ai
涉及实体
相关事实
引用此条事实
Stable URI
https://kongchang.com/claim/27812API
curl https://kongchang.com/api/v1/knowledge/claims/27812MCP
get_claim(id=27812)