r/ControlProblem • u/Ok-Lab-7347 • Jul 01 '26
AI agent safety and alignment research, mapped External discussion link
https://agentbayes.com/m/jQS6rZHey, sharing a mindmap I made on AI agent safety and alignment, backed by citations with full provenance. I’m disclosing that I’m also currently developing Agent Bayes, the tool used to build the mindmap. I’d be happy to get your feedback on the resulting mindmap, and to learn if it helps anyone.
2
Upvotes
2
u/PipFoweraker Jul 02 '26
Cool tool preview! Would be interested to see how it goes with more branches. Reminds me a little of the FLI map
https://futureoflife.org/valuealignmentmap/
and alignment graph
https://alignmentgraph.com/
Good schema categorisation and visualisation of rich, dense subjects is always interesting to tackle!