flâneur — a map of the web's best reading

What is input/output filtering in AI safety? - by Sarah

blog.bluedot.org · saved by 1 readers

Large Language Models (LLMs) can sometimes produce dangerous or harmful outputs, such as instructions for dangerous activities (like building weapons), inappropriate or offensive content, or misinformation.

Explore this link on the map →