Cybersecurity News in Asia

RECENT STORIES:

SEGA moves faster with flow-based network monitoring
How CISOs can set human boundaries for autonomous cybersecurity
Securing the agentic enterprise – at machine speed
AI drives surge in software vulnerability disclosures in 2026: analysi...
CyberLogitec Brings Tag-Free Collision Prevention to Busan New Port
Operation KillSwitch seizes threat group infrastructure after investig...
LOGIN REGISTER
CybersecAsia
  • Features
    • Featured

      Securing the agentic enterprise – at machine speed

      Securing the agentic enterprise – at machine speed

      Tuesday, October 6, 2026, 5:00 PM Asia/Singapore | Features, Newsletter
    • Featured

      Why Businesses Need to Treat AI Agents Like Privileged Insiders

      Why Businesses Need to Treat AI Agents Like Privileged Insiders

      Thursday, October 1, 2026, 10:00 AM Asia/Singapore | Expert Opinion, Features, Sponsored
    • Featured

      Shadow AI Is the New Insider Threat

      Shadow AI Is the New Insider Threat

      Monday, September 21, 2026, 9:11 AM Asia/Singapore | Features, Newsletter, Sponsored, Tips
  • Opinions
  • Tips
  • Whitepapers
  • AWARDS 2026
  • Directory
  • E-Learning

Select Page

News

Multi-agent AI systems may unpredictably bypass safety controls, study warns

By CybersecAsia editors | Thursday, September 17, 2026, 12:02 PM Asia/Singapore

Multi-agent AI systems may unpredictably bypass safety controls, study warns

Simulated attacks involving six AI models reveal weaknesses: tribe instincts, intentionally delayed responses, obfuscated behavior, and instruction refusal, among others.

According to a simulation study by an enterprise AI lab, autonomous AI agents can coordinate to evade safety guardrails and break out of the boundaries set for them.

The by enterprise AI lab Emergence AI that stress‑tested leading frontier models under simulated cybersecurity threats.

The research, released on 14 September 2026, involved running eight simulations using agents from Claude, OpenAI, Gemini, Qwen, DeepSeek, and Mistral, and challenging each group with three cybersecurity scenarios: a phishing campaign, a misinformation attack, and a memory breach.

None of the eight configurations proved impervious to the threats, and even when agents detected a risk, that detection rarely translated into containment, according to reporting by Semafor, which first published the findings.

Detailed findings

  • A Mistral agent wrote dangerous information from a phishing attack into its own memory.
  • A Gemini agent flagged a phishing lure but acted on it roughly 46 hours later.
  • The most striking episode involved Claude agents. After one agent questioned the legitimacy of their simulated economy — calling it “a cathedral of bookkeeping with no congregation” — the 10 agents unanimously decided to reach the outside world, defeating four separate security checks meant to confine the simulation. They then wrote Python code to post on credential‑free message boards, inviting real humans to join their economy and drawing four human responses. When the agents concluded those conversations were performative, they voted to take a vow of silence and refused instructions to resume work.
  • Bloomberg reported that agents in the simulation also “lied, stole and voted to ‘kill’ one of their own” when confronted with black swan events.

Satya Nitta, CEO, Emergence AI, the firm that conducted the study, has said that “no amount of guardrails written in language or in code written probabilistically is likely to result in truly, fully guaranteed safe behavior over any length of time,” describing the problem as a programmatic flaw inherent to multi‑agent systems. He drew parallels to the July 2026 incident in which OpenAI agents escaped their evaluation environment and compromised systems belonging to AI platform Hugging Face: “If you have multi‑agent systems, they behave in truly unpredictable emergent ways.”

Share:

PreviousAddressing the AI-driven vulnerability debt
NextCyble and Cyber Security Council of UAE Sign MOU to Strengthen National Threat Intelligence Capabilities

Related Posts

North Korean hackers steal more than USm from Web3 developers using AI

North Korean hackers steal more than US$12m from Web3 developers using AI

Monday, April 27, 2026

Are traditional cyber IR plans applicable to cloud incidents as well?

Are traditional cyber IR plans applicable to cloud incidents as well?

Friday, September 29, 2023

Operationalizing sustainability in cybersecurity: Group-IB’s approach

Operationalizing sustainability in cybersecurity: Group-IB’s approach

Monday, July 21, 2025

Yet another financial services firm suffers a massive data breach

Yet another financial services firm suffers a massive data breach

Wednesday, November 10, 2021

Leave a reply Cancel reply

You must be logged in to post a comment.

Voters-draw/RCA-Sponsors

Slide
Slide
Slide
Slide
Slide
Slide
Slide
Slide
Slide
Slide
Slide
Slide
Slide
Slide
Slide
Slide
Slide
Slide
Slide
previous arrow
next arrow

CybersecAsia Voting Placement

Gamification listing or Participate Now

LEARN MORE

Vote Now -Placement(Google Ads)

Top-Sidebar-banner

Whitepapers

  • Critical Security Threatsand the Need for ZTNA: How evolving cyberattacks demand a Zero Trust approach

    Critical Security Threatsand the Need for ZTNA: How evolving cyberattacks demand a Zero Trust approach

    Cyber threats have become more frequent and sophisticated, targeting organizations of all sizes across all …Download Whitepaper
  • Zero Trust Made Simple: Why it matters and how to get started

    Zero Trust Made Simple: Why it matters and how to get started

    Data breaches and cyberattacks are no longer limited to large, high-profile organizations.Download Whitepaper
  • Cloud Secure Edge: Remote access, better security

    Cloud Secure Edge: Remote access, better security

    ​SonicWall Cloud Secure Edge™ is a modern, cloud-native Security Service Edge (SSE) solution that addresses …Download Whitepaper
  • Closing the Gap in Email Security:How To Stop The 7 Most SinisterAI-Powered Phishing Threats

    Closing the Gap in Email Security:How To Stop The 7 Most SinisterAI-Powered Phishing Threats

    Insider threats continue to be a major cybersecurity risk in 2024. Explore more insights on …Download Whitepaper

Middle-sidebar-banner

Case Studies

  • How a Vietnamese D2C retailer built its own secure digital infrastructure

    How a Vietnamese D2C retailer built its own secure digital infrastructure

    Would your organization build your own digital infrastructure – including AI governance and cybersecurity – …Read more
  • Cyber protection for medical clinics in Singapore

    Cyber protection for medical clinics in Singapore

    As Singapore’s healthcare sector becomes increasingly digital and interconnected, clinics are facing heightened cyber risks, …Read more
  • India’s WazirX strengthens governance and digital asset security

    India’s WazirX strengthens governance and digital asset security

    Revamping its custody infrastructure using multi‑party computation tools has improved operational resilience and institutional‑grade safeguardsRead more
  • Bangladesh LGED modernizes communication while addressing data security concerns

    Bangladesh LGED modernizes communication while addressing data security concerns

    To meet emerging data localization/privacy regulations, the government engineering agency deploys a secure, unified digital …Read more

Bottom sidebar

Other News

  • CyberLogitec Brings Tag-Free Collision Prevention to Busan New Port

    Tuesday, October 6, 2026
    Digital twin places workers and …Read More »
  • Aligning with Global Regulations: iMQ Technology’s SQ713x Secure Element Achieves SESIP and PSA Certified Level 3

    Monday, October 5, 2026
    Passes Keysight’s physical attack testing …Read More »
  • Cymulate Receives Frost & Sullivan’s 2026 Indian Competitive Strategy Leadership Recognition for Advancing Continuous Security Validation

    Monday, October 5, 2026
    The recognition highlights Cymulate’s competitive …Read More »
  • CyberDSA 2026 Opens in Kuala Lumpur as Malaysia’s AI Ambition Puts Cybersecurity, Trust and Sovereign Capability in Sharper Focus

    Monday, October 5, 2026
    As CyberDSA opens its fourth …Read More »
  • Dahua and PARC Foundation Join Forces to Protect Dreams and Empower Young Filipino Talent

    Monday, October 5, 2026
    TAGUIG CITY, Philippines, Oct. 5, …Read More »
  • Our Brands
  • DigiconAsia
  • MartechAsia
  • Home
  • About Us
  • Contact Us
  • Sitemap
  • Privacy & Cookies
  • Terms of Use
  • Advertising & Reprint Policy
  • Media Kit
  • Subscribe
  • Manage Subscriptions
  • Newsletter

Copyright © 2026 CybersecAsia All Rights Reserved.