ended6월 25일· 1 sources
GLM-5.2 open agent benchmark: 22% Less Tool Failure
Why it matters
This article was originally published on BuildZn. Spent weeks battling flaky AI agents that just couldn't stick to the script. Multi-step tool use was a nightmare, constantly hallucinating API calls o...
1
Sources
+0
24h
—
Growth
4d
Active