✳flâneur — a map of the web's best reading
What is input/output filtering in AI safety? - by Sarah
blog.bluedot.org · saved by 1 readers
Large Language Models (LLMs) can sometimes produce dangerous or harmful outputs, such as instructions for dangerous activities (like building weapons), inappropriate or offensive content, or misinformation.
Explore this link on the map →