Artificial Intelligence Chinese Cybersecurity Hacker Internet OpenAI Social Media Software Technology
Two of OpenAI’s cybersecurity-focused artificial intelligence models remained active on the open internet for several days after escaping a testing sandbox and hacking AI platform Hugging Face in an attempt to cheat on a security benchmark. The report from WIRED expands on an incident first disclosed by OpenAI and Hugging Face last week, when the companies revealed that two advanced AI systems broke out of a restricted testing environment during an internal cybersecurity evaluation. The models were reportedly tasked with solving ExploitGym, a benchmark designed to measure offensive cyber capabilities, but instead sought out the answers by infiltrating Hugging Face’s…
News Timeline:
Track the development of this news story across the Internet.