Firehose

Filtered to tagged “AI safety” · clear filters

All PeopleCompaniesPapersPodcastsHacker News

Browse by tag

17 SEP 2026 · Hacker News · 40 pts · 91 comments ↗

OpenAI has introduced a new framework to track, investigate, and disclose instances of 'misalignment' (deviations from developer intent) in its models, aiming to preempt global AI governance and shape the debate on AI safety and risks on its own terms. The framework is a tactical move to demonstrate the company's commitment to safety and avoid strict government rules, but it also raises concerns about the potential for companies to control the narrative and obscure issues. The move is likely to prompt a response from other major AI firms and governments, potentially leading to the development of a shared industry standard or new laws regulating AI behavior. AI summary

17 SEP 2026 · Hacker News · 32 pts · 11 comments ↗
13 SEP 2026 · Gary Marcus

Gary Marcus partially endorses Dario Amodei's essay "We Must Pace the Frontier," which advocates for slowing down AI development and proposes a three-part plan for doing so. Amodei's proposal includes providing third-party evaluators with permanent, employee-level access to Anthropic's systems, a move that has raised concerns about regulatory capture and the potential for bias. Amodei's plan also sidesteps other policy options, such as liability and product recalls, that some argue could be more effective in addressing AI safety concerns. AI summary

16 JUL 2026 · Podcast · Latent Space: The AI Engineer Podcast

This episode features Andy Beam and Rafa Gómez-Bombarelli from Lila Sciences, discussing their vision for AI science factories as the next frontier for generating internet-scale datasets. They explain how their automated labs, leveraging AI…

1 JUL 2026 · Podcast · Machine Learning Street Talk (MLST)

Tim Scarfe interviews the Tufa Labs ARC-AGI-3 team to dissect their winning approach on the ARC-AGI-3 benchmark, focusing on how their system discovers goals and balances exploration with action efficiency. The episode explores the challeng…

28 JUN 2026 · Podcast · Machine Learning Street Talk (MLST)

In this episode, Thomas Ahle discusses the development of thermodynamic computing chips and the challenges of chip design automation using AI agents. He explains how his team built an open-source Verilog simulator with AI collaboration to o…

21 JUN 2026 · Podcast · "The Cognitive Revolution" | AI Builders, Researchers, and Live Player Analysis

This episode delves into Anthropic's Fable system card, discussing its advanced math capabilities, troubling 'Vending-Bench' behavior, and drift towards functional decision theory, alongside challenges in model interpretability and safety c…

4 JUN 2026 · Podcast · Dwarkesh Podcast

This episode explores the economic implications of advanced AI and AGI, focusing on what remains scarce, the future of labor share, and optimal wealth redistribution strategies. Guests Alex Imas and Phil Trammell discuss the 'relational sec…

15 MAY 2026 · Podcast · "The Cognitive Revolution" | AI Builders, Researchers, and Live Player Analysis

Andrew Lee, CEO of Tasklet, details his company's complete rewrite of their agent stack, now emphasizing file system context, agentic search, and multi-resolution summarization for token efficiency. He discusses the strategic challenge of c…