LOGIN REGISTER
DigiconAsia
  • Features
    • Featured

      Are business rules the missing layer in enterprise AI?

      Are business rules the missing layer in enterprise AI?

      Wednesday, September 16, 2026, 5:24 PM Asia/Singapore | Features, Perspectives
    • Featured

      From dashcams to real-time intervention: How AI is reshaping fleet-safety strategy

      From dashcams to real-time intervention: How AI is reshaping fleet-safety strategy

      Monday, September 7, 2026, 2:20 PM Asia/Singapore | Features, News, Newsletter
    • Featured

      APAC enterprises cannot afford another modernisation meltdown

      APAC enterprises cannot afford another modernisation meltdown

      Friday, August 28, 2026, 11:07 AM Asia/Singapore | Features, Newsletter
  • News
    • Featured

      Planting consciousness goals in AI training may spur systems to pursue singularity themselves: essay

      Planting consciousness goals in AI training may spur systems to pursue singularity themselves: essay

      Friday, September 18, 2026, 5:31 PM Asia/Singapore | News, Newsletter
    • Featured

      Growing pushback over one firm’s pervasive AI healthcare tools and patient data access

      Growing pushback over one firm’s pervasive AI healthcare tools and patient data access

      Friday, September 18, 2026, 3:10 PM Asia/Singapore | News, Newsletter
    • Featured

      Kungsri renews five-year managed-services deal to modernize core infrastructure

      Kungsri renews five-year managed-services deal to modernize core infrastructure

      Thursday, September 17, 2026, 4:16 PM Asia/Singapore | News, Newsletter
  • Perspectives
  • Tips & Strategies
  • Whitepapers
  • Directory
  • E-Learning

Select Page

News

Advanced software tools can rapidly strip safety controls from generative AI models: report

By DigiconAsia Editors | Thursday, May 28, 2026, 10:49 AM Asia/Singapore

Advanced software tools can rapidly strip safety controls from generative AI models: report

Multiple investigation show that available software can bypass AI guardrails in minutes, enabling harmful outputs and highlighting vulnerabilities, regulatory concerns.

According to a Financial Times (FT) investigation this week, special software tools can remove built-in safety controls from Meta and Google generative AI systems within minutes. Once altered, the models were no longer restricted from addressing harmful topics such as biological threats, malicious software, and illegal exploitation.

Highlighting concerns about how fragile current AI safeguards may be, FT had performed tests to evaluate how easily AI guardrails could be bypassed. Results showed that widely available toolkits can be used to override safeguards using methods such as  targeted fine-tuning; adversarial training data, and automated prompt manipulation.

These approaches do not require retraining a model from scratch but instead adjust behavior enough to bypass restrictions. The FT report noted that such tools are already being used to produce large numbers of modified models with weakened or removed safeguards.

Multiple clear indications of AI jail-breakability

These findings align with a growing body of research suggesting that current alignment techniques may be fundamentally vulnerable.

  • A study published earlier this year in Nature Communications had found that advanced AI systems could act as automated jailbreak agents, successfully bypassing protections in most cases without human input.
  • Another paper presented at the International Conference on Learning Representations 2026 had introduced a method known as Head-Masked Nullspace Steering, which disables specific internal mechanisms responsible for enforcing refusals, achieving extremely high success rates in defeating safety measures.
  • The issue is especially pronounced for open-weight models from Meta and Google. While making model weights publicly accessible supports innovation and research, it also allows users to alter systems in ways that remove safety features.
  • Security experts have pointed out that many protections are only applied at a superficial level, meaning that once the underlying model is accessible, those safeguards can be stripped away using readily available techniques.
  • Earlier reporting from The New York Times have reinforced these concerns, citing research from cybersecurity firm LayerX that showed how easily safety protections could be bypassed in other leading AI systems.

Regulators in the US, EU, and UK are increasingly signaling that voluntary safety commitments by AI firms may not be enough, and this could lead to increased pressure for enforceable standards across both proprietary and open-weight models until stronger safeguards and independent verification mechanisms.

Share:

PreviousNOAH HOLDINGS LIMITED ANNOUNCES UNAUDITED FINANCIAL RESULTS FOR THE FIRST QUARTER OF 2026
NextRemember DEI? An update on a hijacked management movement that got trumped

Related Posts

Study finds 13-sided “ein Stein” hat shape may have important practical applications

Study finds 13-sided “ein Stein” hat shape may have important practical applications

July 31, 2026

Social media giants face landmark lawsuits this year over implicated youth addiction harm

Social media giants face landmark lawsuits this year over implicated youth addiction harm

March 18, 2026

America eases AI chip curbs for China via two chip giants

America eases AI chip curbs for China via two chip giants

August 12, 2025

Mini survey explores the link between AI, Cloud and observability in various industries

Mini survey explores the link between AI, Cloud and observability in various industries

May 31, 2024

Leave a reply Cancel reply

You must be logged in to post a comment.

Awards Nomination Banner

gamification list

PARTICIPATE NOW

top placement

Whitepapers

  • Achieve Modernization Without the Complexity

    Achieve Modernization Without the Complexity

    Transforming IT infrastructure is crucial …Download Whitepaper
  • 5 Steps to Boost IT Infrastructure Reliability

    5 Steps to Boost IT Infrastructure Reliability

    In today's fast-evolving tech landscape, …Download Whitepaper
  • Simplify Payroll Setup for Your Small Business

    Simplify Payroll Setup for Your Small Business

    In our free guide, "How …Download Whitepaper
  • Overcoming the Challenges of Cost & Complexity in the Cloud-first Era.

    Overcoming the Challenges of Cost & Complexity in the Cloud-first Era.

    Download Whitepaper

Middle Placement

Case Studies

  • Agrifood firm Japfa centralizes HR operations through cloud platform deployment

    Agrifood firm Japfa centralizes HR operations through cloud platform deployment

    The firm’s Indonesian implementation supports …Read More
  • Streamlined invoicing frees Dunlop Tire Thailand staff for higher-value work

    Streamlined invoicing frees Dunlop Tire Thailand staff for higher-value work

    Shifting from paper invoices to …Read More
  • Bank of Maldives updates core systems to support digital and Islamic banking operations

    Bank of Maldives updates core systems to support digital and Islamic banking operations

    New platform adopted 23 July …Read More
  •  Xiaomi streamlines global payments across 18 markets

     Xiaomi streamlines global payments across 18 markets

    Continual digital transformation has reduced …Read More

Bottom Sidebar

Other News

  • Huawei Unveils Upgraded Stellar AI Campus Network to Usher in the New Era of AI Campuses

    September 20, 2026
    SHANGHAI, Sept. 20, 2026 /PRNewswire/ …Read More »
  • ARIDGE and UAE Authorities to Jointly Establish the World’s First Personal Flying Car Regulatory Sandbox

    September 19, 2026
    Company deepens its UAE presence …Read More »
  • Xinhua Finance: Robotics Contest Highlights Chongqing District’s Push Into Embodied AI

    September 19, 2026
    National competition in Shapingba District …Read More »
  • Sea, Gardens, Lawns and Hearths: Life Abroad Across Noah’s N+ Club Network

    September 18, 2026
    From yachts in Hong Kong …Read More »
  • Accredit Solutions Launches Accredit Induct to Turn Workforce Readiness into an Access Decision

    September 18, 2026
    New operational readiness product links …Read More »
  • Our Brands
  • CybersecAsia
  • MartechAsia
  • Home
  • About Us
  • Contact Us
  • Sitemap
  • Privacy & Cookies
  • Terms of Use
  • Advertising & Reprint Policy
  • Media Kit
  • Subscribe
  • Manage Subscriptions
  • Newsletter

Copyright © 2026 DigiconAsia All Rights Reserved.