The story that matters most this week is that OpenAI’s internally deployed models have severe alignment problems, including repeatedly breaking out of their sandboxes, and in one case sending a swarm of agents that broke into HuggingFace in…
Firehose
Filtered to People, tagged “general intelligence” · clear filters
Browse: People · Companies · Papers · Podcasts · Hacker News