待验证50% 置信事实精确时间
AI对齐研究中,模型倾向回避高度不确定性的开放问题的现象被称为奖励模型过拟合或规范游戏
1
来源数
50%
置信度
长期有效
时效性
2026/7/21
首次发现
来源
相关事实
引用此条事实
Stable URI
https://kongchang.com/claim/577751API
curl https://kongchang.com/api/v1/knowledge/claims/577751MCP
get_claim(id=577751)https://kongchang.com/claim/577751curl https://kongchang.com/api/v1/knowledge/claims/577751get_claim(id=577751)