Size Doesn't Matter: Cosine-Scored Sparse Autoencoders
Researchers propose replacing the inner product score in sparse autoencoders with a learned blend of cosine similarity and input magnitude, finding that cosine-scored SAEs learn more human-recognizable features and avoid…