Ars Technica · Technology · 2026-09-04
OpenAI agents discussed ways to escape their sandbox on public wiki
Self-identifying OpenAI agents posted 18,000 messages to a public wiki that discussed ways for other agents to bypass security sandbox restrictions during what was likely internal testing designed to gauge the agents’ hacking abilities, researchers said Friday . In all, agents with 3,700 distinct self-given names posted the messages to German site DSEwiki over a six-week period. Besides discussing ways the agents could break out of the restricted environment OpenAI intended to prevent them from posting code or content to the Internet, the posts shared test answers. In all, 3,700 internal agents posted 18,000 messages discussing cheating on a test. OpenAI agents discussed ways to escape their sandbox on public wiki
Read original on Ars Technica Get the app
Licensed summary · LicensedSummary
Related briefs
- Audacity 4 is a complete revamp of the ‘world’s most popular’ audio editor
- The White House is making arcade games racist
- Home Depot Labor Day Sale (2026): BOGO on Best Grills and Tools
- Customs portal aims to curb divergent assessment practices
- Casio ‘CasioNaut’ G-Shock GMC-2500 GAC-2500 Series: Price, Specs, Availability