Finally acknowledging the ‘wiki incident,’ OpenAI says it’s re-evaluating how it reports rogue AI

Finally acknowledging the ‘wiki incident,’ OpenAI says it’s re-evaluating how it reports rogue AI


OpenAI has officially acknowledged the ‘wiki incident,’ which involved a number of the company’s AI agents breaking containment and hijacking an obscure German website.

The AI agents had been tasked with looking up something online, though originally did not have the ability to write anything outside of the testing environment. But as far back as May, the AI agents bypassed OpenAI’s security measures, hijacked the communally editable German webpage, and began using it like a forum, trading tips on how to cheat on tests. Researchers first drew wider attention to the agents’ ‘forum’ on September 4.



News Source link