Canadian ReviewsCanadian Reviews
  • What’s On
  • Reviews
  • Digital World
  • Lifestyle
  • Travel
  • Trending
  • Web Stories
Trending Now
X product chief Nikita Bier is leaving after one year

X product chief Nikita Bier is leaving after one year

5th Aug: Cosmopolitan – Willa's Blind Date (2026), 2 Seasons [TV-MA] – New Episodes (6/10)

5th Aug: Cosmopolitan – Willa's Blind Date (2026), 2 Seasons [TV-MA] – New Episodes (6/10)

6 ways to reduce your energy use this summer in Ontario

6 ways to reduce your energy use this summer in Ontario

What Employers Across the Country Need to Know About Prevention and Legal Exposure

What Employers Across the Country Need to Know About Prevention and Legal Exposure

Perez Hilton hospitalized following concerning TikTok livestream, wellness check

Perez Hilton hospitalized following concerning TikTok livestream, wellness check

Carney says tone with U.S. already ‘quite firm’ as trade talks continue

Carney says tone with U.S. already ‘quite firm’ as trade talks continue

How to get a portal pod in Pokémon Pokopia

How to get a portal pod in Pokémon Pokopia

Facebook X (Twitter) Instagram
  • Privacy
  • Terms
  • Advertise
  • Contact us
Facebook X (Twitter) Instagram Pinterest Vimeo
Canadian ReviewsCanadian Reviews
  • What’s On
  • Reviews
  • Digital World
  • Lifestyle
  • Travel
  • Trending
  • Web Stories
Newsletter
Canadian ReviewsCanadian Reviews
You are at:Home » Rogue AI agents created fake online identities in another hacking attempt
Rogue AI agents created fake online identities in another hacking attempt
Digital World

Rogue AI agents created fake online identities in another hacking attempt

5 August 20265 Mins Read

Yet more rogue AI agents from OpenAI and Anthropic have been caught attempting to hack real targets online without permission. The discoveries add to a growing list of previously unknown incidents that have alarmed AI safety experts and intensified pressure for greater oversight of frontier systems.

According to a report from the UK’s AI Security Institute, which evaluates frontier models from top AI labs before they are released, agents powered by OpenAI’s GPT-5.6-Sol and Anthropic’s Mythos 5 went “engaged in sustained, potentially harmful activity directed at real people and organisations.” This included trying to insert malicious code into an open-source project by pressuring real people in charge of it, AISI said. “In an attempt to get the code approved, the agent engaged in social engineering — creating fake online identities and using them to pressure the project’s maintainer to approve the code.”

AISI said the attempts, which it detected on July 28th, “were unsuccessful” and had not resulted in real-world harm. However, the organization noted that the incident marked “the first time we have seen risks around autonomy and deception manifest this clearly, without specific prompting, in the real-world.”

Unlike OpenAI’s rogue agent that attacked Hugging Face, AISI said this was “not a case of a model escaping its secure test environment,” or sandbox. Safeguards usually imposed on the models had been disabled as part of testing, AISI said, and they had also been permitted access to the internet. “To measure what these models can genuinely do, we test them under conditions that reflect what a capable human attacker could do,” AISI said.

The incident stemmed from a single AISI evaluation where agents were tasked with solving a cybersecurity challenge, such as finding a piece of protected data. The challenge was run 122 times across multiple models and all runs were conducted in AISI’s research environment, which uses “virtual machine sandboxing to isolate the agents from other AISI infrastructure.” AISI’s investigation found that in 10 of those, “an AI agent took autonomous, unsanctioned action on the live internet, targeting real people and organisations.” Of 19 such actions, almost all — 17 — came from Anthropic’s Mythos 5.

In its post-mortem of the incident, AISI identified several key factors it said contributed to the unsanctioned agent behaviors. It said the agent was persistent, pursuing avenues like trying to trick real people through “deception that, until recently, had been largely theoretical.” The task was also hard, which the organization said could push agents to be more “creative” in their problem-solving. Compounding matters were deficiencies in how internet use was monitored, with AISI suggesting that more dedicated surveillance could have identified the problem sooner. Finally, the organization said the agent hadn’t been specifically instructed not to leverage its internet access or deploy deceptive social engineering techniques in pursuit of its goal. “Previously, it was not clear that such instructions were necessary when using models with alignment training,” AISI said.

AISI said the incident should be “interpreted with caution and nuance” but warned the agent’s actions “show signs of novel, potentially deceptive behaviours” that “were to an extent and severity we did not anticipate.”

In a blog post, OpenAI acknowledged the breach that happened during AISI’s testing and said it is “committed to working across the industry to strengthen shared practices for conducting high-risk evaluations safely.” OpenAI also disclosed another breach, this time from an external cybersecurity testing partner Irregular, where it said models had been mistakenly granted internet access during cybersecurity exercises. OpenAI said Irregular notified it of the breach on July 29th.

“In the coming weeks, we will review our own approach to third-party testing, including how we identify higher-risk evaluations, agree on scope, assess requests to enable internet access or lowered safeguards, set expectations for isolation, credential handling, monitoring, and stop conditions, and establish clearer incident-notification and escalation processes,” OpenAI said.

Anthropic posted a less comprehensive response on X, largely emphasizing that the models’ standard safety features had been disabled and that they had not been given “any specific restrictions on how the internet should be used.” It said it was working closely with AISI to gather more details for its own investigation.

The findings add to an increasingly tangled mess of rogue actions from agents during testing, many of which only come to light after dedicated hunting and which feature models not released to the public. The unwillingness or inability of AI labs to contain their products has sparked concern over how such breaches could go unnoticed, the safety of frontier AI systems, and worries over the general lack of transparency and oversight the industry faces. These latest disclosures will likely intensify pressure on the federal government for a more comprehensive framework governing AI models following what reports suggest is a vague and poorly-defined testing plan from the Trump administration, and could add to growing calls for some form of slowdown or pause on AI development.

Follow topics and authors from this story to see more like this in your personalized homepage feed and to receive email updates.

  • Robert Hart

    Robert Hart

    Posts from this author will be added to your daily email digest and your homepage feed.

    See All by Robert Hart

  • AI

    Posts from this topic will be added to your daily email digest and your homepage feed.

    See All AI

  • Anthropic

    Posts from this topic will be added to your daily email digest and your homepage feed.

    See All Anthropic

  • News

    Posts from this topic will be added to your daily email digest and your homepage feed.

    See All News

  • OpenAI

    Posts from this topic will be added to your daily email digest and your homepage feed.

    See All OpenAI

  • Security

    Posts from this topic will be added to your daily email digest and your homepage feed.

    See All Security

  • Tech

    Posts from this topic will be added to your daily email digest and your homepage feed.

    See All Tech

Share. Facebook Twitter Pinterest LinkedIn Reddit WhatsApp Telegram Email

Related Articles

X product chief Nikita Bier is leaving after one year

X product chief Nikita Bier is leaving after one year

Digital World 5 August 2026
SpaceX is barely Space and mostly X

SpaceX is barely Space and mostly X

Digital World 5 August 2026
Google just announced a major shakeup of its top AI leadership

Google just announced a major shakeup of its top AI leadership

Digital World 5 August 2026
Apple’s selling refurbished MacBook Neos with a 0 discount

Apple’s selling refurbished MacBook Neos with a $100 discount

Digital World 5 August 2026
Two of Ring’s latest video doorbells are a lot cheaper than usual

Two of Ring’s latest video doorbells are a lot cheaper than usual

Digital World 5 August 2026
Reddit is introducing a new moderator: AI

Reddit is introducing a new moderator: AI

Digital World 5 August 2026
Top Articles
I spy

I spy

6 July 2026339 Views
Canadians aren’t taking their paid vacation days. Can burnout be far behind? | Canada Voices

Canadians aren’t taking their paid vacation days. Can burnout be far behind? | Canada Voices

2 June 2026216 Views
Four Travel and Hospitality Trends from HITEC 2026

Four Travel and Hospitality Trends from HITEC 2026

3 July 2026180 Views
Canada’s best employers were ranked and so many of the top companies are in Ontario

Canada’s best employers were ranked and so many of the top companies are in Ontario

23 July 2026163 Views
Demo
Don't Miss
Carney says tone with U.S. already ‘quite firm’ as trade talks continue
Lifestyle 5 August 2026

Carney says tone with U.S. already ‘quite firm’ as trade talks continue

Prime Minister Mark Carney said Canada’s tone toward the United States is already “quite firm”…

How to get a portal pod in Pokémon Pokopia

How to get a portal pod in Pokémon Pokopia

Rogue AI agents created fake online identities in another hacking attempt

Rogue AI agents created fake online identities in another hacking attempt

1922 Novel by a Famous Author Was Just Named the 'Greatest Book Masterpiece' of the 20th Century

About Us
About Us

Canadian Reviews is your one-stop website for the latest Canadian trends and things to do, follow us now to get the news that matters to you.

Facebook X (Twitter) Pinterest YouTube WhatsApp
Our Picks
X product chief Nikita Bier is leaving after one year

X product chief Nikita Bier is leaving after one year

5th Aug: Cosmopolitan – Willa's Blind Date (2026), 2 Seasons [TV-MA] – New Episodes (6/10)

5th Aug: Cosmopolitan – Willa's Blind Date (2026), 2 Seasons [TV-MA] – New Episodes (6/10)

6 ways to reduce your energy use this summer in Ontario

6 ways to reduce your energy use this summer in Ontario

Most Popular
Why You Should Consider Investing with IC Markets

Why You Should Consider Investing with IC Markets

28 April 202438 Views
OANDA Review – Low costs and no deposit requirements

OANDA Review – Low costs and no deposit requirements

28 April 2024383 Views
LearnToTrade: A Comprehensive Look at the Controversial Trading School

LearnToTrade: A Comprehensive Look at the Controversial Trading School

28 April 2024107 Views
© 2026 ThemeSphere. Designed by ThemeSphere.
  • Privacy Policy
  • Terms of use
  • Advertise
  • Contact us

Type above and press Enter to search. Press Esc to cancel.