[KongchangAI]
Unverified50% confidenceFactExact time

Anthropic等公司已经开始专门测试模型的说服能力和欺骗倾向,作为安全评估的重要维度

1
Sources
50%
Confidence
Medium-term (~90 days)
Relevance
9/2/2026
First Seen
Valid until: 12/1/2026

Sources

Related Entities

Related Claims

Cite This Claim

Stable URI
https://kongchang.com/claim/844507
API
curl https://kongchang.com/api/v1/knowledge/claims/844507
MCP
get_claim(id=844507)