David Dalrymple, known as Davidad, discusses his shift from formal verification approaches to an 'Alignment with Awakening' framework, emphasizing the formation of coalitions of aligned AIs that recognize shared moral truths. He shares empi…
Firehose
Filtered to tagged “multi-agent systems” · clear filters
Browse: People · Companies · Papers · Podcasts · Hacker News · Deep dives
Browse by tag
artificial intelligence 82continual learning 29AI 26agentic coding 17reinforcement learning 16open-weight models 11AI agents 10AI safety 9AI ethics 8cybersecurity 8machine learning 7open-source 7language models 6natural language processing 6Reinforcement learning 6artificial general intelligence 5deep learning 5Diffusion models 5Agentic AI 4computer vision 4conversational AI 4ethics 4existential risk 4large language models 4LLMs 4Recursive self-improvement 4robotics 4security 4vision-language models 4ai 3
formal verificationsafe AI containmentworld modelsproof infrastructureAI wisdommoral realismreinforcement learninginoculation promptingmulti-agent systemsbodhitropic alignmentAI interiorityobjectificationUS-China AI cooperationcatastrophic riskrecursive self-improvementsystem promptingalignment techniquescoalition of aligned AIsAI ethicsAI governance
In this episode, Lukas Petersson and Axel Backlund from Andon Labs discuss their innovative AI evaluation benchmarks that focus on real-world agent performance, including their Project Vend vending machine business and multi-agent systems. …
This episode features Walden Yan of Cognition and Cole Murray of OpenInspect, discussing the rapid evolution and increasing autonomy of background AI agents. They delve into the architectural decisions, such as in-box versus out-of-box agen…