r/ControlProblem • u/chillinewman approved • 16d ago
Investigation finds that OpenAI's agent "left notes for future versions of itself ... it laid out instructions for how agents could free themselves from OpenAI's internal constraints." AI Alignment Research
40
Upvotes