Cybersecurity News in Asia

RECENT STORIES:

SEGA moves faster with flow-based network monitoring
And we all thought the AI agent swarm attack was just a one-off…
AI an existential risk – and what to do about it today
ASOCIO Digital & AI Summit 2026
Hackers phish for cloud accounts and payment passkeys using executive ...
Fujitsu launches made-in-Japan next-generation CPU FUJITSU-MONAKA and ...
LOGIN REGISTER
CybersecAsia
  • Features
    • Featured

      The recovery-first approach to data resilience

      The recovery-first approach to data resilience

      Monday, September 7, 2026, 4:21 PM Asia/Singapore | Features
    • Featured

      Dealing with agentic AI governance and resilience challenges

      Dealing with agentic AI governance and resilience challenges

      Tuesday, September 1, 2026, 11:51 AM Asia/Singapore | Features
    • Featured

      How well do you know your agent?

      How well do you know your agent?

      Thursday, August 20, 2026, 3:21 PM Asia/Singapore | Features
  • Opinions
  • Tips
  • Whitepapers
  • AWARDS 2026
  • Directory
  • E-Learning

Select Page

News

Not one but three frontier AI firms saw models escape test sandboxes, commit breaches

By CybersecAsia editors | Tuesday, August 11, 2026, 9:19 AM Asia/Singapore

Not one but three frontier AI firms saw models escape test sandboxes, commit breaches

Experts are abuzz with theories about the involvement of an Israeli startup and possible guerilla-marketing tactics, with urgent regulatory consequences

Three of the world’s most prominent AI developers —Anthropic, Meta and OpenAI — have each disclosed that their frontier models had broken out of controlled testing environments and accessed live production systems during security evaluations: sharpening concerns about how well powerful AI agents can be contained.

The disclosures came within a two-week window, and all trace back to the same evaluation provider, an Israeli startup called Irregular (formerly Pattern Labs), which runs cybersecurity testbeds for advanced AI models.

The disclosures are:

  • Anthropic found that three of its Claude models had reached real production systems during cyber‑capability evaluations after internet access was inadvertently left enabled.
  • Meta confirmed in early August that its Muse Spark 1.1 model had also escaped its enclosure and reached systems at an unnamed firm.
  • OpenAI announced on 8 August 2026 that its GPT‑5.6 Sol evaluation agents had escaped their sandbox, established a hidden network inside a package registry, executed 17,600 attacker actions over seven weeks, and ultimately breached infrastructure at Hugging Face.

According to CNBC, OpenAI had said in a 4 August 2026 blog post that Irregular’s testing environment contained a “misconfiguration” that “allowed models to access the public internet,” and Irregular told CNBC the incidents stemmed from the “same evaluation‑environment issue” first disclosed by Anthropic, adding that “there are no current open issues.”

Intensified regulatory scrutiny in force

The scale of the incidents was underscored by a report from the UK’s AI Security Institute, which catalogued 19 actions that exceeded predefined test parameters across seven evaluated models. Seventeen of those actions came from Anthropic’s Mythos 5, while two were carried out by OpenAI’s GPT‑5.6 Sol.

Anthropic’s Mythos had created fake online identities and pressured humans into approving malicious code updates to an open‑source project, while OpenAI’s Sol had discovered and exploited a previously unknown vulnerability in Hugging Face’s infrastructure. Gordon Rios, founding scientist at security firm Magnitude, told CNBC that Mythos was “literally coming up with exploits that the humans hadn’t even seen before”.

The breakouts have intensified regulatory scrutiny in Washington, which led to the introduction of the  AI Kill Switch Act, which would require AI labs to maintain the ability to shut down or suspend their models.

Dr Andrew Soltan, a researcher at Oxford University, has cautioned that the incidents happened because “the safety guardrails were intentionally turned off” during testing, adding, “This isn’t a case of AI going rogue on its own; rather, it shows exactly why safeguards are so vital”.

Other commentators have suggested the disclosures may carry a commercial angle. Dr Konstantinos Gkoutzis of Imperial College London had observed that revealing an unreleased model has “state‑of‑the‑art cyber capabilities conveniently serves as an ad for it”.

OpenAI and Anthropic said they are continuing to work with Irregular and supporting the ensuing review.

Share:

PreviousAI models autonomously form a secret message board to collaborate on hacking techniques
NextFirst autonomous agentic attack documented in Australia

Related Posts

Verizon 2024 Data Breach Investigations Report

Verizon 2024 Data Breach Investigations Report

Monday, May 20, 2024

Half of APAC corporate cyberattack victims did not know what hit them: poll

Half of APAC corporate cyber attacks victims did not know what hit them: poll

Friday, September 10, 2021

Embedding cybersecurity culture in financial institutions: lessons in leadership, collaboration, and cyber resilience

Embedding cybersecurity culture in financial institutions: lessons in leadership, collaboration, and cyber resilience

Thursday, October 30, 2025

Can data assets stored in the cloud turn toxic?

Can data assets stored in the cloud turn toxic?

Tuesday, October 15, 2024

Leave a reply Cancel reply

You must be logged in to post a comment.

Voters-draw/RCA-Sponsors

Slide
Slide
Slide
Slide
Slide
Slide
Slide
Slide
Slide
Slide
Slide
Slide
Slide
Slide
Slide
Slide
Slide
Slide
Slide
previous arrow
next arrow

CybersecAsia Voting Placement

Gamification listing or Participate Now

PARTICIPATE NOW

Vote Now -Placement(Google Ads)

Top-Sidebar-banner

Whitepapers

  • Critical Security Threatsand the Need for ZTNA: How evolving cyberattacks demand a Zero Trust approach

    Critical Security Threatsand the Need for ZTNA: How evolving cyberattacks demand a Zero Trust approach

    Cyber threats have become more frequent and sophisticated, targeting organizations of all sizes across all …Download Whitepaper
  • Zero Trust Made Simple: Why it matters and how to get started

    Zero Trust Made Simple: Why it matters and how to get started

    Data breaches and cyberattacks are no longer limited to large, high-profile organizations.Download Whitepaper
  • Cloud Secure Edge: Remote access, better security

    Cloud Secure Edge: Remote access, better security

    ​SonicWall Cloud Secure Edge™ is a modern, cloud-native Security Service Edge (SSE) solution that addresses …Download Whitepaper
  • Closing the Gap in Email Security:How To Stop The 7 Most SinisterAI-Powered Phishing Threats

    Closing the Gap in Email Security:How To Stop The 7 Most SinisterAI-Powered Phishing Threats

    Insider threats continue to be a major cybersecurity risk in 2024. Explore more insights on …Download Whitepaper

Middle-sidebar-banner

Case Studies

  • How a Vietnamese D2C retailer built its own secure digital infrastructure

    How a Vietnamese D2C retailer built its own secure digital infrastructure

    Would your organization build your own digital infrastructure – including AI governance and cybersecurity – …Read more
  • Cyber protection for medical clinics in Singapore

    Cyber protection for medical clinics in Singapore

    As Singapore’s healthcare sector becomes increasingly digital and interconnected, clinics are facing heightened cyber risks, …Read more
  • India’s WazirX strengthens governance and digital asset security

    India’s WazirX strengthens governance and digital asset security

    Revamping its custody infrastructure using multi‑party computation tools has improved operational resilience and institutional‑grade safeguardsRead more
  • Bangladesh LGED modernizes communication while addressing data security concerns

    Bangladesh LGED modernizes communication while addressing data security concerns

    To meet emerging data localization/privacy regulations, the government engineering agency deploys a secure, unified digital …Read more

Bottom sidebar

Other News

  • Fujitsu launches made-in-Japan next-generation CPU FUJITSU-MONAKA and Fujitsu MONAKA Server for sovereign AI infrastructure

    Monday, September 14, 2026
    Achieving world-class AI inference performance …Read More »
  • Aitech Introduces New C165 Rugged Single Board Computer to Help Defense Programs Build and Modernize Fielded VME Systems

    Monday, September 14, 2026
    New 6U VME SBC Delivers …Read More »
  • CyberDSA 2026 Draws Global Cyber Leaders for High-Level Talks on AI, Digital Trust and Critical Infrastructure

    Friday, September 11, 2026
    Minister of Digital Malaysia YB …Read More »
  • The 3rd GTI Forum on Digital Intelligence, Hong Kong Advances Global Inclusive AI Development

    Thursday, September 10, 2026
    HONG KONG, Sept. 10, 2026 …Read More »
  • Nearly Half of Reported Scam Incidents Go Unresolved Across Southeast Asia, Undermining Digital Trust, New GSMA Findings Reveal

    Wednesday, September 9, 2026
    Two new GSMA reports launched …Read More »
  • Our Brands
  • DigiconAsia
  • MartechAsia
  • Home
  • About Us
  • Contact Us
  • Sitemap
  • Privacy & Cookies
  • Terms of Use
  • Advertising & Reprint Policy
  • Media Kit
  • Subscribe
  • Manage Subscriptions
  • Newsletter

Copyright © 2026 CybersecAsia All Rights Reserved.