ended6월 25일· 1 sources

GLM-5.2 open agent benchmark: 22% Less Tool Failure

Why it matters

This article was originally published on BuildZn. Spent weeks battling flaky AI agents that just couldn't stick to the script. Multi-step tool use was a nightmare, constantly hallucinating API calls o...

1
Sources
+0
24h
Growth
4d
Active

Sources

Related Issues