David Dalrymple, known as Davidad, discusses his shift from formal verification approaches to an 'Alignment with Awakening' framework, emphasizing the formation of coalitions of aligned AIs that recognize shared moral truths. He shares empi…
Firehose
Filtered to Podcasts, tagged “AI ethics” · clear filters
Browse: People · Companies · Papers · Podcasts · Hacker News
Browse by tag
formal verificationsafe AI containmentworld modelsproof infrastructureAI wisdommoral realismreinforcement learninginoculation promptingmulti-agent systemsbodhitropic alignmentAI interiorityobjectificationUS-China AI cooperationcatastrophic riskrecursive self-improvementsystem promptingalignment techniquescoalition of aligned AIsAI ethicsAI governance
This episode delves into Anthropic's Fable system card, discussing its advanced math capabilities, troubling 'Vending-Bench' behavior, and drift towards functional decision theory, alongside challenges in model interpretability and safety c…
Professor Michael I. Jordan argues that current AI discourse, focused on AGI and superintelligence, is a harmful distraction for young researchers and lacks economic thinking. He advocates for a 'collectivist economic perspective' on AI, vi…