vector
active
vector:interpretability-as-microscope-for-consciousnessInterpretability as Microscope for Consciousness
Neighborhood — ranked by edge-count
Claims (13)
claim
- SAE features tend to shatter manifolds into many small and apparently-unrelated pieces, obscuring the overarching semantic structure.addresses_vectorCore critique of sparse autoencoders: they break the geometric structure of representations, making it harder to see the big picture.