ended7월 24일· 1 sources
Beyond Reconstruction: Verifying Model Explanations with RECAP
Why it matters
What Changed For years, the field of mechanistic interpretability has relied heavily on natural-language autoencoders to translate hidden model activations into human-readable explanations. The prevai...
1
Sources
+0
24h
—
Growth
7d
Active