Aligning AI With Human Goals Might Be Impossible, Says AI Prof. Stuart Russell

Last week, OpenAI scrapped the model it had slated to release as GPT-6.1 Astra due to test results showing that it was deceptive and otherwise misaligned. “I think it’s about time,” said Stuart Russell, the computer science professor at the University of California, Berkeley, who co-authored the…

aireleases

Sources