Firehose

Filtered to Companies, tagged “AI safety” · clear filters

All PeopleCompaniesPapersPodcastsHacker News

Browse by tag

23 JUL 2026 · Databricks

Omnigent's intent-based authorization closes the gap between traditional authorization and AI agents by binding a session to a declared purpose, ensuring that actions are checked against that intent and denied or gated for human approval if outside it. This approach blocks prompt injection attacks, where an attacker injects instructions into an agent's content to steer it into unauthorized actions. By pairing intent-based authorization with session-risk scoring policy, Omnigent creates a layered defense that reinforces each other. AI summary

21 JUL 2026 · Anthropic

Anthropic is donating an additional $20 million to Public First Action, bringing their total support to $40 million, to promote policies that maintain meaningful safeguards, sustain America's AI leadership, and demand transparency from AI model developers. This donation aims to counter the growing risks posed by rapidly advancing AI models and to ensure that governments and policymakers can effectively mitigate these risks. Anthropic's Advanced AI Framework proposes measures such as model verification, enforcement of safe practices, and independent evaluation to ensure the safe development and deployment of AI models. AI summary

16 JUL 2026 · OpenAI

Learn how OpenAI is making ChatGPT safer for teens with age-appropriate protections, learning tools, parental controls, and expert partnerships.