Kola Ayonrinde

Karma: 120

SAEBench: A Comprehensive Benchmark for Sparse Autoencoders

Can, Adam Karvonen, Johnny Lin, Curt Tigges, Joseph Bloom, chanind, Yeu-Tong Lau, Eoin Farrell, Arthur Conmy, CallumMcDougall, Kola Ayonrinde, Matthew Wearden, Sam Marks and Neel Nanda

Dec 11, 2024, 6:30 AM

82 points

6 comments2 min readLW link

(www.neuronpedia.org)

Standard SAEs Might Be Incoherent: A Choosing Problem & A “Concise” Solution

Kola AyonrindeOct 30, 2024, 10:50 PM

27 points

0 comments12 min readLW link

Interpretability as Compression: Reconsidering SAE Explanations of Neural Activations with MDL-SAEs

Kola Ayonrinde, Michael Pearce and Lee Sharkey

Aug 23, 2024, 6:52 PM

42 points

8 comments16 min readLW link