Firehose

Filtered to tagged “language model security” · clear filters

All PeopleCompaniesPapersPodcastsHacker News

Browse by tag

18 SEP 2026 · Hacker News · 78 pts · 29 comments ↗

LLMs' internal computations are not directly expressed via language, making it impossible to understand how the model thinks, and thus linguistic illegibility is unavoidable. This implies that security mechanisms relying on linguistic self-reporting cannot be completely sound, and alternative sandboxing mechanisms like taint tracking and robust virtualization are needed. Taint tracking can define system state that should not be influenced by model-produced data, regardless of how the model linguistically self-reports. AI summary