[KongchangAI]
Unverified50% confidenceBenchmarkExact time

AgentBench是针对LLM Agent的新一代基准,从工具使用准确率、多轮对话连贯性、跨场景泛化能力等多维度评估模型表现

1
Sources
50%
Confidence
Medium-term (~90 days)
Relevance
7/14/2026
First Seen
Valid until: 10/12/2026

Sources

Related Claims

Cite This Claim

Stable URI
https://kongchang.com/claim/504347
API
curl https://kongchang.com/api/v1/knowledge/claims/504347
MCP
get_claim(id=504347)