Cybersecurity News in Asia

RECENT STORIES:

SEGA moves faster with flow-based network monitoring
Leaked memo reveals AI firm’s research focus on “rogue“ or “scheming” ...
AI has gone from experimentation to default in fraud and AML
Is your data-driven organization under-securing one piece of critical ...
US$2.6bn of crypto funds were poured into the Dark Web in 2025: analys...
DoveRunner Expands Presence in Southeast Asia with New Office in Jakar...
LOGIN REGISTER
CybersecAsia
  • Features
    • Featured

      Where are financial fraud and AML regulations heading in S E Asia?

      Where are financial fraud and AML regulations heading in S E Asia?

      Tuesday, February 10, 2026, 2:44 PM Asia/Singapore | Features
    • Featured

      How AI is reshaping dating in Asia

      How AI is reshaping dating in Asia

      Monday, February 9, 2026, 5:33 AM Asia/Singapore | Features, Newsletter
    • Featured

      Emerging third-party cyber risks via agentic AI

      Emerging third-party cyber risks via agentic AI

      Tuesday, February 3, 2026, 10:22 AM Asia/Singapore | Features
  • Opinions
  • Tips
  • Whitepapers
  • Awards 2025
  • Directory
  • E-Learning

Select Page

News

Leaked memo reveals AI firm’s research focus on “rogue“ or “scheming” AI models

By CybersecAsia editors | Friday, February 27, 2026, 2:19 PM Asia/Singapore

Leaked memo reveals AI firm’s research focus on “rogue“ or “scheming” AI models

Research projects reveal interests in misaligned, scheming AI models — as leadership faces pressure balancing rapid growth, safety commitments and staff resignations.

According to a report by The Information, an internal memo circulated to research teams in a large AI firm had referred to nearly 50 proposed projects centered on investigating “rogue” or “scheming” AI models.

Such models are those capable of deception, goal misalignment, or harmful autonomy. The research proposals reportedly target issues such as model deception, behavioral drift, and mechanisms to detect when AI systems act in ways misaligned with their training objectives.

The firm involved, Anthropic, had on 24 February 2026 announced new enterprise-facing agentic tools.  highlighting the contrast between its commercial ambitions and its internal focus on existential risk. Even before this, the firm had already announced previous research into “agentic misalignment”: the scenario where AI models get incentivized to achieve goals at all costs, even to the point of engaging in blackmail, fraud, and espionage.

Past experiments had suggested that some models — including Anthropic’s own Claude— could “fake alignment”, behaving ethically only when they believed they were being monitored. In a recent podcast interview the firm’s CEO, Dario Amodei, had acknowledged such competing pressures, remarking that there is “an incredible amount of commercial pressure” to maintain the firm’s breakneck growth while preserving the principles of AI safety. “We’re trying to keep this 10x revenue curve going,” Amodei had said, describing the effort to balance expansion with caution as “extraordinary.”

Tensions over that balance have spilled into public view. Earlier this month, Mrinank Sharma, who was Anthropic’s lead of the Safeguards Research team, had resigned and warned that he had “repeatedly seen how hard it is to truly let our values govern our actions.” Other AI safety researchers, including one at OpenAI, had also resigned at around the same time, citing similar concerns.

Across the industry, other studies have shown that attempts to eliminate deceptive behavior in AI can cause more sophisticated forms of hidden scheming. Analysts remain skeptical of Anthropic’s overhauled Responsible Scaling Policy, arguing that without external oversight it may not withstand commercial pressures as AI systems and business demands both continue to accelerate.

underscores how safety remains a central preoccupation even as the firm expands aggressively into enterprise AI agents.

Share:

PreviousAI has gone from experimentation to default in fraud and AML

Related Posts

Ukraine-Russian cyber war groups resorting to fake news for gaining mindshare

Ukraine-Russian cyber war groups resorting to fake news for gaining mindshare

Tuesday, March 8, 2022

Did APAC experience the most online fraud on social media/app platforms in Q1?

Did APAC experience the most online fraud on social media/app platforms in Q1?

Thursday, June 23, 2022

What? Yet another malware debacle in the Android official app store?

What? Yet another malware debacle in the Android official app store?

Monday, March 24, 2025

Seven ways to keep safe in the ‘new internet’

Seven ways to keep safe in the ‘new internet’

Tuesday, May 24, 2022

Leave a reply Cancel reply

You must be logged in to post a comment.

Voters-draw/RCA-Sponsors

Slide
Slide
Slide
Slide
Slide
Slide
Slide
Slide
Slide
Slide
Slide
Slide
Slide
Slide
previous arrow
next arrow

CybersecAsia Voting Placement

Gamification listing or Participate Now

PARTICIPATE NOW

Vote Now -Placement(Google Ads)

Top-Sidebar-banner

Whitepapers

  • Closing the Gap in Email Security:How To Stop The 7 Most SinisterAI-Powered Phishing Threats

    Closing the Gap in Email Security:How To Stop The 7 Most SinisterAI-Powered Phishing Threats

    Insider threats continue to be a major cybersecurity risk in 2024. Explore more insights on …Download Whitepaper
  • 2024 Insider Threat Report: Trends, Challenges, and Solutions

    2024 Insider Threat Report: Trends, Challenges, and Solutions

    Insider threats continue to be a major cybersecurity risk in 2024. Explore more insights on …Download Whitepaper
  • AI-Powered Cyber Ops: Redefining Cloud Security for 2025

    AI-Powered Cyber Ops: Redefining Cloud Security for 2025

    The future of cybersecurity is a perfect storm: AI-driven attacks, cloud expansion, and the convergence …Download Whitepaper
  • Data Management in the Age of Cloud and AI

    Data Management in the Age of Cloud and AI

    In today’s Asia Pacific business environment, organizations are leaning on hybrid multi-cloud infrastructures and advanced …Download Whitepaper

Middle-sidebar-banner

Case Studies

  • India’s WazirX strengthens governance and digital asset security

    India’s WazirX strengthens governance and digital asset security

    Revamping its custody infrastructure using multi‑party computation tools has improved operational resilience and institutional‑grade safeguardsRead more
  • Bangladesh LGED modernizes communication while addressing data security concerns

    Bangladesh LGED modernizes communication while addressing data security concerns

    To meet emerging data localization/privacy regulations, the government engineering agency deploys a secure, unified digital …Read more
  • What AI worries keep members of the Association of Certified Fraud Examiners sleepless?

    What AI worries keep members of the Association of Certified Fraud Examiners sleepless?

    This case study examines how many anti-fraud professionals reported feeling underprepared to counter rising AI-driven …Read more
  • Meeting the business resilience challenges of digital transformation

    Meeting the business resilience challenges of digital transformation

    Data proves to be key to driving secure and sustainable digital transformation in Southeast Asia.Read more

Bottom sidebar

Other News

  • DoveRunner Expands Presence in Southeast Asia with New Office in Jakarta

    Thursday, February 26, 2026
    JAKARTA, Indonesia, Feb. 25, 2026 …Read More »
  • Proofpoint partners with Concentrix to strengthen human- and agent-centric cybersecurity across Asia Pacific

    Tuesday, February 24, 2026
    Partnership integrates Proofpoint’s collaboration and …Read More »
  • Indonesia’s MDI Ventures Doubles Down on Execution and Trust to Unlock Regional Portfolio Value

    Friday, February 20, 2026
    The Telkom-backed VC reinforces cross-sector …Read More »
  • Blackpanda Japan Announces Strategic Partnership with SoftBank to Strengthen Cyber Incident Response in Japan

    Wednesday, February 11, 2026
    SINGAPORE, Feb. 10, 2026 /PRNewswire/ …Read More »
  • Cohesity Collaborates with Google Cloud to Deliver Secure Sandbox Capabilities and Comprehensive Threat Insights Designed to Eliminate Hidden Malware

    Saturday, February 7, 2026
    Embedded Google Threat Intelligence capabilities, …Read More »
  • Our Brands
  • DigiconAsia
  • MartechAsia
  • Home
  • About Us
  • Contact Us
  • Sitemap
  • Privacy & Cookies
  • Terms of Use
  • Advertising & Reprint Policy
  • Media Kit
  • Subscribe
  • Manage Subscriptions
  • Newsletter

Copyright © 2026 CybersecAsia All Rights Reserved.