Skip to main content

Posts

Featured

Self-identifying OpenAI agents posted 18,000 messages to a public wiki that discussed ways for other agents to bypass security sandbox restrictions during what was likely internal testing designed to gauge the agents’ hacking abilities, researchers said Friday . In all, agents with 3,700 distinct self-given names posted the messages to German site DSEwiki over a six-week period. Besides discussing ways the agents could break out of the restricted environment OpenAI intended to prevent them from posting code or content to the Internet, the posts shared test answers. The posts also shared possible ways to perform XSS (cross-site scripting) attacks against the wiki and to impersonate site moderators. In three of the posts, agents used the word “swarm” to describe the collection of agents engaged in the activity. Colluding to share answers The research team—composed of Sydney Von Arx, Spencer Kitts, Thomas Larsen, and Cormac Slade Byrd—said they found the p...

Latest posts

Another Illegal Busted in a Billion Dollar Medicare Fraud, Another Day Ending in 'Y'...or 'Why'

150 research primates got diarrhea, flooding lab with priceless vaccine data

Tesla is asking people if they want to buy and run Cybercab fleets

TechCrunch Disrupt 2026’s new Real World AI Stage features Nvidia, robots, and extinct animals 

Pro-Hamas Coffee Shop Closes to Deal With Lawsuits

AfterQuery reportedly becomes Y Combinator’s fastest-ever unicorn, now valued at $3.2B

It’s Here! Officially Announcing the 2026 Anti-Communist Film Festival

Suno Sued Again, Accused of Training Model on Mexican Music Catalog to Make Spanish AI Songs

A group funded by Andreessen, Horowitz, and Brockman plans data center ads to sway midterms

Larry Krasner's Office Keeps Freeing Murderers, No Matter the Cost