flâneur — a map of the web's best reading

Linearity of Relation Decoding in Transformer Language Models

lre.baulab.info · 621 words · saved by 1 readers

Approximating relation decoding in transormer LMs as simple linear transformations on the subject token.

Linearity of Relation Decoding in Transformer Language Models Linearity of Relation Decoding in Transformer LMs Evan Hernandez 1* , Arnab Sen Sharma 2* , Tal Haklay 3 , Kevin Meng 1 , Martin Wattenberg 4 , Jacob Andreas 1 , Yonatan Belinkov 3 , David Bau 2 1 MIT CSAIL , 2 Northeastern University , 3 Technion - IIT ; * Equal contribution New! Try interacting with a lre-edited GPT to see the effect of inserting hundreds of memories. --> ArXiv Preprint Source Code Dataset How do Transformer LMs Decode Relations? Much of the knowledge contained in neural language models may be expressed in terms o

Explore this link on the map →

related reading