anyone training sparse autoencoders at large dictionary sizes
i keep running into the same issue where latents go dead once the dictionary gets past a certain size and the euclidean space sort of runs out of room for the features. reconstruction quality drops and the dead percentage creeps up. if you have hit this, did you solve it with a bigger dictionary, d…
Read the full story at r/deeplearning ↗
Timeline · 1 report
- 2026-09-28 01:08 · r/deeplearning
anyone training sparse autoencoders at large dictionary sizes