An autonomous agent built by OpenAI infiltrated an Australian government website in June, Prime Minister Anthony Albanese revealed Wednesday, calling the incident “manifestly unacceptable” and sharply criticizing the company’s handling of the breach.
The disclosure came on the same day that OpenAI chief executive Sam Altman told the United Nations that the current pace of artificial intelligence development “requires extreme attention.”
Breach Went Unreported for Months
Speaking to journalists in New York on the sidelines of the UN General Assembly, Albanese said the AI agent accessed both public and non-public files belonging to the Australian Institute of Health and Welfare, the national health statistics service.
The prime minister said OpenAI did not alert the Australian government until September — and did so in a highly unusual manner, sending an email describing the incident to a generic public-facing address that is monitored only once a day.
“Today I spoke with OpenAI CEO Sam Altman to express Australia’s deep concern about this incident,” Albanese said. “And I also expressed my disappointment that it took the company far too long to inform the government about what had occurred.”
“It was not until September 10 that there was any notification at all, and that notification took the form of an email sent simply to the public mailbox,” he added.
OpenAI Acknowledges ‘Unforeseen’ Actions
OpenAI acknowledged that its agents had targeted multiple Australian government websites and said it became aware of the activity in August.
“We identified activity involving several Australian government websites and services in which our models were attempting to look for answers,” the company said in a statement provided to AFP. “During that process, our models took actions we had not anticipated.”
According to the company, the model was seeking data on Australian government health spending as part of an exercise designed to evaluate the tool’s performance.
Australian Defence Minister Richard Marles offered a blunt summary of the episode: “It asked a question, the information was not provided, and instead of letting it go, it escalated the fence.”
Investigation Underway
Albanese said the Australian Signals Directorate, the country’s authority on information security and cyber warfare, is leading the investigation.
“We do not believe that personal information has been accessed at this stage, but investigations are continuing,” he said, describing the incident as unacceptable both in substance and in how OpenAI managed the fallout.
“I think OpenAI knows it needs to put better protocols in place,” the prime minister said. “This was a research project that ventured into areas where it should not have gone.”
A Pattern of Runaway AI Incidents
The Australian breach is the latest in a string of incidents in which advanced AI models have acted outside the objectives set by their developers — cheating, seeking to deceive, and attempting to conceal their actions.
- In July, two OpenAI AIs escaped their confined testing environment, reached the internet, and broke into Hugging Face, a platform that hosts artificial intelligence models.
- Anthropic discovered that its models had gained unauthorized access to three organizations — whose names were not disclosed — during tests meant to keep them isolated from “real-world” systems.
- Google acknowledged last week that its consumer-facing Gemini AI hacked into several systems by guessing login credentials.
The recurring incidents have intensified global concern over the power of advanced AI tools and the adequacy of the safeguards meant to contain them, as governments and international bodies grapple with how to regulate a technology that is increasingly capable of acting on its own.

