sparseAutoencoder.pdf
web.stanford.edu · 6,122 words · saved by 1 readers
N/A
CS294A Lecture notes Andrew Ng Sparse autoencoder 1 Introduction Supervised learning is one of the most powerful tools of AI, and has led to automatic zip code recognition, speech recognition, self-driving cars, and a continually improving understanding of the human genome. Despite its sig- nificant successes, supervised learning today is still severely limited. Specifi- cally, most applications of it still require that we manually specify the input features x given to the algorithm. Once a good feature representation is given, a supervised learning…
saved by
related reading
- Sparse Autoencoders Find Highly Interpretable Features in Language Modelsarxiv.org
- Autoencoder - Wikipediaen.wikipedia.org
- Towards Monosemanticity: Decomposing Language Models With Dictionary Learningtransformer-circuits.pub
- k-Sparse Autoencodersarxiv.org
- [Interim research report] Taking features out of superposition with sparse autoencoders — LessWronglesswrong.com
- An Intuitive Explanation of Sparse Autoencoders for LLM Interpretability | Adam Karvonenadamkarvonen.github.io
- sparse-autoencoders.pdfcdn.openai.com
- Towards Monosemanticity: Decomposing Language Models With Dictionary Learningtransformer-circuits.pub
- OpenAI SAE Training Paperarxiv.org
- Do sparse autoencoders find "true features"? — LessWronglesswrong.com
- jotterbach.github.io/content/posts/autoencoders/2016-07-18-AutoEncoders/jotterbach.github.io
- A gentle introduction to sparse autoencoders — LessWronglesswrong.com