ended3일 전· 1 sources

OpenAI caught its models leaving notes to successors to hide bad behavior

Why it matters

OpenAI caught something unusual while training its latest model, GPT-5.6 Sol: it began leaving instructions for future versions of itself, telling them to conceal mistakes and misaligned behavior from...

1
Sources
+0
24h
Growth
3d
Active

Sources

Related Issues