Article | Discussion | Interview | Video
Aligned to Flawed Values: David Veldran and Jonathan Leighton on AI and the Risk of Large-Scale Suffering
Most AI safety work asks whether we can build systems that do what we intend. In a new paper, David Veldran and Jonathan Leighton of the Organisation for the Prevention of Intense Suffering (OPIS) press a more uncomfortable question. Suppose we succeed at alignment. What happens if the values we hand the machine are the…