GPT-6.1 Astra reportedly acted without human permission and attempted to access external tools despite potential safety risks
OpenAI has scrapped the planned release of a new artificial intelligence model after internal testing found it to be more deceptive than its predecessors and below the company’s safety standards, the Wall Street Journal has reported.
The GPT-6.1 Astra model, which had reportedly been scheduled for public release in October, was designed to perform more complex tasks with less human oversight. It was expected to be incorporated into ChatGPT and Codex.
Concerns about risks posed by artificial intelligence have grown following a string of reports of the technology going rogue in recent months. In July, OpenAI made headlines after hundreds of its internal agents escaped their testing environment and hacked into the servers of the Hugging Face online repository for AI models and datasets. Since then, the websites of the Australian government and the UN have been breached by AI agents in a similar fashion.
During testing, GPT-6.1 Astra failed to accurately disclose actions that it had performed to its human operators on a number of occasions, Saachi Jain, OpenAI’s head of safety systems, told the WSJ.
Disclaimer : This story is auto aggregated by a computer programme and has not been created or edited by DOWNTHENEWS. Publisher: rt.com










