flâneur

A socratic dialogue over the utility of DNA language models (Part 1 of 2)

substack.com · 3,642 words · saved by 1 readers

3.7k words, 17 minute reading time

Introduction The dialogue Part 1 is focused on variant pathogenicity prediction using these models. Part 2 is focused on genome generation using these models. I am aware that DNA language models are useful for things other than those two (like protein fitness), but variant prediction and genome generation are the two bits that I find most interesting. I think I, alongside many other people in this field, live in this seemingly parallel universe where we don’t really understand why anyone is working on DNA language models. I say ‘parallel’, because there is obviously a world in which…

related reading