ended7월 31일· 1 sources
OpenAI’s Goblin Post Highlights an Emerging Risk in AI Alignment and Reliability
Why it matters
OpenAI has published a post-mortem examining an unusual pattern in its model testing: recurring references to “goblins” and “gremlins” in model outputs. The company’s official post, “Where the goblins...
1
Sources
+0
24h
—
Growth
52d
Active