Back to feed
Alignment Forum
Alignment Forum
7/8/2026
Notes on technical alignment via human-like social drives

Notes on technical alignment via human-like social drives

Short summary

A lengthy research post explores using human innate social and moral drives as a foundation for technical alignment of brain-like AGI. The author examines specific human instincts, three potential failure modes (wrong moral circle, balance-of-power issues, consequentialist desires overriding virtue ethics), and implementation details for transplanting social drives into AGI source code. The work is exploratory and seeks community feedback for deconfusion.

  • Proposes human social instincts as a starting point for AGI alignment
  • Identifies three failure modes: moral circle errors, power imbalance, and consequentialist override of virtues
  • Discusses person-first vs desire-first pathways for virtue development in AGI

Generated with AI, which can make mistakes.

Is this a good recommendation for you?

Comments

Failed to load comments. Please try again.

Explore more