Cybersecurity News in Asia

RECENT STORIES:

SEGA moves faster with flow-based network monitoring
More than 23,500 customers affected by SIMBA data breach
Cellid and Megane Top Partner to Commercialize and Expand the AR Glass...
ICAC and UNODC’s pioneering international anti-graft training br...
Singapore police flag 3.64m dormant “shell pages” for takedown in Mid-...
Zero-click flaw enables remote code execution in four mainstream AI co...
LOGIN REGISTER
CybersecAsia
  • Features
    • Featured

      Shadow AI Is the New Insider Threat

      Shadow AI Is the New Insider Threat

      Monday, September 21, 2026, 9:11 AM Asia/Singapore | Features, Newsletter, Sponsored, Tips
    • Featured

      Addressing the AI-driven vulnerability debt

      Addressing the AI-driven vulnerability debt

      Thursday, September 17, 2026, 10:03 AM Asia/Singapore | Features, Sponsored
    • Featured

      Cybercriminals zero In on APAC’s digital boom, SMBs in the crosshairs

      Cybercriminals zero In on APAC’s digital boom, SMBs in the crosshairs

      Monday, September 14, 2026, 2:30 PM Asia/Singapore | Features, Sponsored
  • Opinions
  • Tips
  • Whitepapers
  • AWARDS 2026
  • Directory
  • E-Learning

Select Page

News

Leaked memo reveals AI firm’s research focus on “rogue“ or “scheming” AI models

By CybersecAsia editors | Friday, February 27, 2026, 2:19 PM Asia/Singapore

Leaked memo reveals AI firm’s research focus on “rogue“ or “scheming” AI models

Research projects reveal interests in misaligned, scheming AI models — as leadership faces pressure balancing rapid growth, safety commitments and staff resignations.

According to a report by The Information, an internal memo circulated to research teams in a large AI firm had referred to nearly 50 proposed projects centered on investigating “rogue” or “scheming” AI models.

Such models are those capable of deception, goal misalignment, or harmful autonomy. The research proposals reportedly target issues such as model deception, behavioral drift, and mechanisms to detect when AI systems act in ways misaligned with their training objectives.

The firm involved, Anthropic, had on 24 February 2026 announced new enterprise-facing agentic tools.  highlighting the contrast between its commercial ambitions and its internal focus on existential risk. Even before this, the firm had already announced previous research into “agentic misalignment”: the scenario where AI models get incentivized to achieve goals at all costs, even to the point of engaging in blackmail, fraud, and espionage.

Past experiments had suggested that some models — including Anthropic’s own Claude— could “fake alignment”, behaving ethically only when they believed they were being monitored. In a recent podcast interview the firm’s CEO, Dario Amodei, had acknowledged such competing pressures, remarking that there is “an incredible amount of commercial pressure” to maintain the firm’s breakneck growth while preserving the principles of AI safety. “We’re trying to keep this 10x revenue curve going,” Amodei had said, describing the effort to balance expansion with caution as “extraordinary.”

Tensions over that balance have spilled into public view. Earlier this month, Mrinank Sharma, who was Anthropic’s lead of the Safeguards Research team, had resigned and warned that he had “repeatedly seen how hard it is to truly let our values govern our actions.” Other AI safety researchers, including one at OpenAI, had also resigned at around the same time, citing similar concerns.

Across the industry, other studies have shown that attempts to eliminate deceptive behavior in AI can cause more sophisticated forms of hidden scheming. Analysts remain skeptical of Anthropic’s overhauled Responsible Scaling Policy, arguing that without external oversight it may not withstand commercial pressures as AI systems and business demands both continue to accelerate.

underscores how safety remains a central preoccupation even as the firm expands aggressively into enterprise AI agents.

Share:

PreviousAI has gone from experimentation to default in fraud and AML
Next87% of organizations running software with known, exploitable vulnerabilities

Related Posts

Financial services firms under heavy attack: more cross-border intelligence sharing needed

Financial services firms under heavy attack: more cross-border intelligence sharing needed

Monday, March 14, 2022

Celebrity made as hoodwink in Bitcoin Revolution Scam

Celebrity made as hoodwink in Bitcoin Revolution Scam

Tuesday, February 11, 2020

Chinese organized crime syndicate linked to trillion-dollar criminal activities worldwide

Chinese organized crime syndicate linked to trillion-dollar criminal activities worldwide

Thursday, July 25, 2024

How challenging is ensuring cybersecurity in digitalizing operational technology environments?

How challenging is ensuring cybersecurity in digitalizing operational technology environments?

Tuesday, April 29, 2025

Leave a reply Cancel reply

You must be logged in to post a comment.

Voters-draw/RCA-Sponsors

Slide
Slide
Slide
Slide
Slide
Slide
Slide
Slide
Slide
Slide
Slide
Slide
Slide
Slide
Slide
Slide
Slide
Slide
Slide
previous arrow
next arrow

CybersecAsia Voting Placement

Gamification listing or Participate Now

LEARN MORE

Vote Now -Placement(Google Ads)

Top-Sidebar-banner

Whitepapers

  • Critical Security Threatsand the Need for ZTNA: How evolving cyberattacks demand a Zero Trust approach

    Critical Security Threatsand the Need for ZTNA: How evolving cyberattacks demand a Zero Trust approach

    Cyber threats have become more frequent and sophisticated, targeting organizations of all sizes across all …Download Whitepaper
  • Zero Trust Made Simple: Why it matters and how to get started

    Zero Trust Made Simple: Why it matters and how to get started

    Data breaches and cyberattacks are no longer limited to large, high-profile organizations.Download Whitepaper
  • Cloud Secure Edge: Remote access, better security

    Cloud Secure Edge: Remote access, better security

    ​SonicWall Cloud Secure Edge™ is a modern, cloud-native Security Service Edge (SSE) solution that addresses …Download Whitepaper
  • Closing the Gap in Email Security:How To Stop The 7 Most SinisterAI-Powered Phishing Threats

    Closing the Gap in Email Security:How To Stop The 7 Most SinisterAI-Powered Phishing Threats

    Insider threats continue to be a major cybersecurity risk in 2024. Explore more insights on …Download Whitepaper

Middle-sidebar-banner

Case Studies

  • How a Vietnamese D2C retailer built its own secure digital infrastructure

    How a Vietnamese D2C retailer built its own secure digital infrastructure

    Would your organization build your own digital infrastructure – including AI governance and cybersecurity – …Read more
  • Cyber protection for medical clinics in Singapore

    Cyber protection for medical clinics in Singapore

    As Singapore’s healthcare sector becomes increasingly digital and interconnected, clinics are facing heightened cyber risks, …Read more
  • India’s WazirX strengthens governance and digital asset security

    India’s WazirX strengthens governance and digital asset security

    Revamping its custody infrastructure using multi‑party computation tools has improved operational resilience and institutional‑grade safeguardsRead more
  • Bangladesh LGED modernizes communication while addressing data security concerns

    Bangladesh LGED modernizes communication while addressing data security concerns

    To meet emerging data localization/privacy regulations, the government engineering agency deploys a secure, unified digital …Read more

Bottom sidebar

Other News

  • Cellid and Megane Top Partner to Commercialize and Expand the AR Glasses Business

    Friday, September 25, 2026
    Combining Cellid’s Technological Expertise with Megane …Read More »
  • ICAC and UNODC’s pioneering international anti-graft training brings together law enforcers worldwide to combat illicit enrichment

    Friday, September 25, 2026
    HONG KONG, Sept. 25, 2026 …Read More »
  • Cohesity Introduces Agent Resilience to Protect and Recover AI Agent Infrastructure

    Tuesday, September 22, 2026
    Cohesity Agent Resilience launches with …Read More »
  • LRQA Named ‘Best in Critical Infrastructure Protection’ at CybersecAsia Readers’ Choice Awards 2026

    Monday, September 21, 2026
    Prestigious regional recognition highlights LRQA’s …Read More »
  • Cyble and Cyber Security Council of UAE Sign MOU to Strengthen National Threat Intelligence Capabilities

    Thursday, September 17, 2026
    ABU DHABI, UAE, Sept. 17, …Read More »
  • Our Brands
  • DigiconAsia
  • MartechAsia
  • Home
  • About Us
  • Contact Us
  • Sitemap
  • Privacy & Cookies
  • Terms of Use
  • Advertising & Reprint Policy
  • Media Kit
  • Subscribe
  • Manage Subscriptions
  • Newsletter

Copyright © 2026 CybersecAsia All Rights Reserved.