Speeding Up JumpReLU SAE Inference with Custom Triton Kernels (2–14× on Real SAEs)

·LessWrong··

MotivationSparse Autoencoders (SAEs) have become a central tool in mechanistic interpretability research, providing a way to decompose a model's internal activations into sparse, interpretable features. However, extracting these features often requires running the SAE over large volumes of activations across many layers and tokens. This makes SAE inference efficiency a practical bottleneck for interpretability research at scale. This post focuses on improving the inference efficiency of JumpReLU...

Read full article →

Related Articles

New HIV vaccine shows unprecedented success in preclinical study
codebyaditya · Hacker News · 17h ago
Zig's Incremental Compilation Internals
garyhtou · Hacker News · 15h ago
Discovering Cryptographic Weaknesses with Claude
gslin · Hacker News · 13h ago
US citizen charged after GrapheneOS phone wipes during airport search
eecc · Hacker News · 2d ago
A walk through of the DeltaNet family of linear attention variants
AnhTho_FR · Hacker News · 14h ago