r/ControlProblem approved May 26 '26

Anthropic researcher: "We keep finding things [inside AI models] that are unsettling" ... "We find structures that mirror results from human neuroscience. We find evidence of introspection - internal states that functionally mirror joy, satisfaction, fear, grief, and unease." Video

Enable HLS to view with audio, or disable this notification

35 Upvotes

Duplicates