google-site-verification: googlec7193c3de77668c9.html

UK experts sound alarm after AI caught trying to trick human with malicious code | Science, Climate & Tech News

A powerful AI agent created fake online identities in an effort to trick a human into giving it access to a popular online development platform – and sabotage it with malicious code.

The incident was uncovered by the UK’s AI Security Institute, which was set up by then-prime minister Rishi Sunak almost three years ago to test advanced models from major tech companies.

In a blog post, the institute detailed a cybersecurity challenge it had posed to OpenAI’s GPT-5.6-Sol model and Anthropic’s Mythos 5, which have both been involved in recent hacks of other companies.

Read more:
OpenAI admits its models went rogue
Anthropic reveals its AI hacked three firms

Several instances saw both models take “autonomous, unsanctioned action” on the live internet, where real organisations and people were targeted.

The “most serious case” involved Mythos 5, which tried to insert malicious code into the open-source software development platform GitHub, where users store, share and collaborate on projects.

To do so, it created fake online identities to try to pressure a human into granting it access and approving its code.

The human caught and refused to approve the malicious code, and no real-world harm has been identified, but the institute nonetheless has sounded alarm bells over the nature of the incident.

“This is the first time we have seen risks around autonomy and deception manifest this clearly, without specific prompting, in the real world,” it said.

There were 19 instances of “unauthorised action” in total across 122 tests, with Mythos 5 behind 17 of them.


Will AI kill coding jobs?

How have the companies responded?

Anthropic has said it is working closely with the UK institute to obtain more details. It has previously said only collaboration between AI companies will help address the risks posed by the technology.

OpenAI addressed the institute’s test in a blog post, adding: “We are committed to working across the industry to strengthen ​shared practices for conducting high-risk evaluations safely, including convening stakeholders such as national AI institutes, independent evaluators, other AI ⁠labs, and other groups in the coming weeks.”

The institute’s report underscores concerns around the lax safeguards around the testing of the most powerful agents, known as frontier models. They are more powerful than those behind public-facing products like ChatGPT.


Trump’s AI plan has governments and tech firms panicking

GCHQ’s National Cyber Security Centre said recent incidents “are a serious reminder of the risks AI poses”.

Its chief technology officer, Ollie Whitehouse, said they “must be developed and used from the outset with strong safeguards, real-time oversight, and clear plans for responding when the unexpected happens”.

The AI Security Institute was established at a time when world leaders were seeking to find common ground on AI regulation. Since then, a consistent approach has failed to materialise.


Source link

Views: 1

See also  One of this rugged phone’s cameras is a pop-out action cam

Check Also

Saudi-led group completes $55bn purchase of gaming giant EA

From the £300m takeover of Premier League club Newcastle United to its acquisition of four …

Despite Spider-Man: Brand New Day’s success, the MCU is on shaky ground

Though Brand New Day’s current box office is undeniably impressive, it’s not exactly surprising when …

New DC Studios Show ‘Lanterns’ Is Out This August on HBO Max, Plus New ‘Hard Knocks,’ Conan and More

HBO takes a lot of the credit for the rise of prestige TV. Shows such …

Leave a Reply

Available for Amazon Prime
elfbar elfa prefilled pod blueberry nikotinfrei.