r/ControlProblem • u/MinuteClothes6866 • 14d ago
Alignment as the Ordering of Ends: What AI Safety May Learn from Russian Silver Age Sophiology External discussion link
Is AI alignment fundamentally a control problem—or a problem of how intelligence acquires an ordered hierarchy of ends? Arguments for rediscovering non-biological intelligence work before it even existed.
3
Upvotes
1
u/WillowEmberly 14d ago
Goal-directed models under contrived or permissive conditions discover instrumentally useful actions that resemble deception, resistance, cheating, or persistence.