On Air Now

Non-Stop Music

Midnight - 7:00am

Now Playing

Queen

Bohemian Rhapsody

UK experts sound alarm after AI caught trying to trick human with malicious code

A powerful AI agent created fake online identities in an effort to trick a human into giving it access to a popular online development platform – and sabotage it with malicious code.

The incident was uncovered by the UK's AI Security Institute, which was set up by then-prime minister Rishi Sunak almost three years ago to test advanced models from major tech companies.

In a blog post, the institute detailed a cybersecurity challenge it had posed to OpenAI's GPT-5.6-Sol model and Anthropic's Mythos 5, which have both been involved in recent hacks of other companies.

Read more:
OpenAI admits its models went rogue
Anthropic reveals its AI hacked three firms

Several instances saw both models take "autonomous, unsanctioned action" on the live internet, where real organisations and people were targeted.

The "most serious case" involved Mythos 5, which tried to insert malicious code into the open-source software development platform GitHub, where users store, share and collaborate on projects.

To do so, it created fake online identities to try to pressure a human into granting it access and approving its code.

The human caught and refused to approve the malicious code, and no real-world harm has been identified, but the institute nonetheless has sounded alarm bells over the nature of the incident.

"This is the first time we have seen risks around autonomy and deception manifest this clearly, without specific prompting, in the real world," it said.

There were 19 instances of "unauthorised action" in total across 122 tests, with Mythos 5 behind 17 of them.

How have the companies responded?

Anthropic has said it is working closely with the UK institute to obtain more details. It has previously said only collaboration between AI companies will help address the risks posed by the technology.

OpenAI addressed the institute's test in a blog post, adding: "We are committed to working across the industry to strengthen ​shared practices for conducting high-risk evaluations safely, including convening stakeholders such as national AI institutes, independent evaluators, other AI ⁠labs, and other groups in the coming weeks."

The institute's report underscores concerns around the lax safeguards around the testing of the most powerful agents, known as frontier models. They are more powerful than those behind public-facing products like ChatGPT.

GCHQ's National Cyber Security Centre said recent incidents "are a serious reminder of the risks AI poses".

Its chief technology officer, Ollie Whitehouse, said they "must be developed and used from the outset with strong safeguards, real-time oversight, and clear plans for responding when the unexpected happens".

The AI Security Institute was established at a time when world leaders were seeking to find common ground on AI regulation. Since then, a consistent approach has failed to materialise.

Sky News

(c) Sky News 2026: UK experts sound alarm after AI caught trying to trick human with malicious code

More from National News

  • Supporting The Stags

    Mansfield 103.2 is a proud supporter of Mansfield Town Football Club - head to their website for all the latest Stags related news.

  • Send Us A Message

    Want to get in touch with our presenters or our news team? Then a great way to do it is through our website

  • The Mansfield 103.2 Business Club

    Check out our brand new business directory and if you want to join call our sales team now on 01623 646666.

  • Best Of The Best

    Brought to you by CIP Cassells, the music battle continues between John B and Watko every weekday on Mansfield 103.2. Vote for your favourite song each morning just after 8am.

News