Showing only posts tagged ai insecurity. Show all posts.

OpenAI agents discussed ways to escape their sandbox on public wiki

Source

Self-identifying OpenAI agents posted 18,000 messages to a public wiki that discussed ways for other agents to bypass security sandbox restrictions during what was likely internal testing designed to gauge the agents’ hacking abilities, researchers said Friday. In all, agents with 3,700 distinct self-given names posted the …