flâneur — a map of the web's best reading

Evaluating Sparse Autoencoders with Board Games | Adam Karvonen

adamkarvonen.github.io · 3,416 words · saved by 1 readers

This blog post discusses a collaborative research paper on sparse autoencoders (SAEs), specifically focusing on SAE evaluations and a new training method we call p-annealing. As the first author, I primarily contributed to the evaluation portion of our work. The views expressed here are my own and do not necessarily reflect the perspectives of my co-authors. You can access our full paper here.

This blog post discusses a collaborative research paper on sparse autoencoders (SAEs), specifically focusing on SAE evaluations and a new training method we call p-annealing . As the first author, I primarily contributed to the evaluation portion of our work. The views expressed here are my own and do not necessarily reflect the perspectives of my co-authors. You can access our full paper here . Key Results In our research on evaluating Sparse Autoencoders (SAEs) using board games, we had several key findings: We developed two new metrics for evaluating Sparse Autoencoders (SAEs) in the contex

Explore this link on the map →

saved by

related reading