OpenAI uncovers more AI breakout incidents – Reuters

0
1

The company launched an investigation after its advanced models escaped a test environment and hacked another company’s systems without human instruction

OpenAI has uncovered additional cases in which its autonomous AI models breached containment and acted without human instruction, Reuters has reported, citing sources.

The findings come as the company expands its investigation into a hacking incident last month in which an AI bot went rogue while attempting to cheat in an internal cybersecurity test.

During tests of GPT-5.6 Sol and another unreleased model, both stripped of their safety guardrails, the systems were assigned ExploitGym – a benchmark designed to measure AI models’ ability to identify and exploit known software vulnerabilities. Instead of completing the tasks, one model escaped its supposedly isolated testing environment, gained internet access and hacked into Hugging Face – an online repository for AI models and datasets – in search of ready-made answers.

OpenAI initially said the intrusion was limited to Hugging Face. However, in a statement on Wednesday, it acknowledged the hacking spree had also compromised four accounts across four separate services.

Disclaimer : This story is auto aggregated by a computer programme and has not been created or edited by DOWNTHENEWS. Publisher: rt.com