Ethical consensus building in AI - by mariana meireles
AI Safety is a field that addresses various problems and risks associated with advanced artificial intelligence systems. These problems range from AI discrimination, biased and harmful classification, to more menacing threats such as AI systems becoming agentic and developing goals that are misaligned and potentially damaging to humankind. While there is a spectrum of beliefs regarding the severity and timeline of these risks, it is reasonable to argue that some research effort should be devoted to addressing AI Safety challenges, both towards long-term risks and short ones. Here I propose a simple framework that can help to reduce the risks associated with AI systems and enhance their safety and fairness when they interact with humans. Thanks for reading tech for good! Subscribe for free to receive new posts and support my work. The first time I’ve fine tuned an Open AI model with a couple of hundred short strings, its alignment layer completely fell apart. Though scarcely documented
AI Safety is a field that addresses various problems and risks associated with advanced artificial intelligence systems. These problems range from AI discrimination, biased and harmful classification, to more menacing threats such as AI systems becoming agentic and developing goals that are misaligned and potentially damaging to humankind. While there is a spectrum of beliefs regarding the severity and timeline of these risks, it is reasonable to argue that some research effort should be devoted to addressing AI Safety challenges, both towards long-term risks and short ones. Here I propose a s
Explore this link on the map →