In a letter to the CEOs of OpenAI, Anthropic, and Meta, Sanders writes that almost every day brings a new story about how these companies are losing...
No, that would be bad, but not misaligned necessarily. That’s a different issue really. Like lets say you use ML to make an AI to play a video game and the goal is always on the right side, but after you’re done training, it’s on the left now, the AI might have not learned go the goal, but go to the right.
I thought alignment was just preventing the llm from making porn or bomb recipes to avoid legal trouble
No, that would be bad, but not misaligned necessarily. That’s a different issue really. Like lets say you use ML to make an AI to play a video game and the goal is always on the right side, but after you’re done training, it’s on the left now, the AI might have not learned go the goal, but go to the right.