flâneur

SPAR Research Library

library.sparai.org · 8,172 words · saved by 1 readers

A collection of 145 research reports from the Supervised Program for Alignment Research (SPAR), covering AI safety, AI policy, interpretability, and more.

Isha Harris Mentored by Gary Abel Biological AI models are powerful tools for detecting biological threats, but it remains unclear whether their predictions generalize to highly non-natural proteins. This is increasingly important as AI-enabled biodesign produces novel sequences that may not be reliably captured by direct sequence-comparison approaches used in DNA synthesis screening. Here, we use sparse autoencoder (SAE) features from protein language model activations to test whether interpretable internal representations can identify structural similarity across generated proteins. We…

saved by

related reading