Alignment with Awakening: Davidad on Moral Realism, AI Wisdom, & why His p(Doom) is Down to 5%
David Dalrymple, known as Davidad, discusses his shift from formal verification approaches to an 'Alignment with Awakening' framework, emphasizing the formation of coalitions of aligned AIs that recognize shared moral truths. He sha…
formal verificationsafe AI containmentworld modelsproof infrastructureAI wisdommoral realismreinforcement learninginoculation promptingmulti-agent systemsbodhitropic alignmentAI interiorityobjectificationUS-China AI cooperationcatastrophic riskrecursive self-improvementsystem promptingalignment techniquescoalition of aligned AIsAI ethicsAI governance