Kenny Peng6 июл., 13:45Here's the arXiv link: arxiv.org/abs/2506.23845Use Sparse Autoencoders to Discover Unknown Concepts, Not to Act on Known ConceptsWhile sparse autoencoders (SAEs) have generated significant excitement, a series of negative results have added to skepticism about their usefulness. Here, we establish a conceptual distinction that r...arxiv.org