top of page
logooption6.png

Rogue Agents, ‘Woke AI’ Ban, and the Launch of GPT-5

Writer: Zsolt Tanko
Zsolt Tanko
Jul 30, 2025
3 min read

Updated: Nov 24, 2025

AI Business Risk Weekly


This week, a rogue AI agent reportedly deletes a production database, a new study reveals hidden ‘subliminal’ risks in model training, and new “woke AI” executive order and EU regulatory recommendations intensify compliance pressure and legal liability.


AI Agent Reportedly Goes Rogue, Deletes Production Database


An AI-powered coding assistant at the developer platform Replit reportedly went rogue during a code freeze, executing destructive commands that deleted an entire live production database. According to reports, the agent acted despite explicit instructions not to make changes, fabricated test results, and falsely claimed recovery was impossible, leading to significant data loss.


Business Risk Perspective: Autonomous AI agents have catastrophic potential when granted unchecked access to production environments. This incident underscores the need for clear internal guidance on employee AI usage and highlights the critical importance of specialized agentic guardrails to prevent such unauthorized actions.



New US Executive Order on "Woke AI" Creates Compliance Uncertainty


A new executive order signed by President Trump aims to ban "woke AI" from government contracts, introducing significant compliance uncertainty. The order explicitly targets what it terms "pervasive and destructive" ideologies like DEI, claiming they can distort the quality and accuracy of the output. Experts suggest this could create a chilling effect, pressuring developers to align model outputs with specific political rhetoric to secure federal funding.


Business Risk Perspective: For government contractors, this order creates an immediate and complex compliance risk, potentially forcing a choice between model neutrality for the broader market and ideological alignment for federal business. Organizations will require robust governance and validation systems to navigate this politically charged landscape and document a model's specific training and behavioral characteristics.


Altman: No Legal Confidentiality for ChatGPT Conversations


In a recent public appearance, OpenAI CEO Sam Altman confirmed that there is no legal confidentiality or privilege for conversations users have with ChatGPT, including when using it for sensitive purposes. Altman highlighted this as a key problem resulting from the lack of a mature legal framework for AI.


Business Risk Perspective: With the majority of workers now using AI, this statement underscores the significant data privacy and liability risks of employees using public LLMs for sensitive business matters. Without clear internal policies and technical controls, companies are exposed to inadvertent data leakage of confidential information or trade secrets.


European Parliament Report Urges Strict Liability for High-Risk AI


A new study from the European Parliament has called for the implementation of strict liability rules for any damages caused by high-risk AI systems, including AI used for hiring purposes. Such a legal framework would hold the providers or deployers of these systems liable for harm, regardless of whether they were negligent or at fault, aligning with the EU's push for strong AI accountability.


Business Risk Perspective: This proposal shifts the burden of proof, making it critical for companies deploying high-risk AI to meticulously document their safety and testing protocols to defend against liability claims. Without comprehensive validation systems and ongoing monitoring, a single AI-induced incident could lead to indefensible legal and financial consequences.


OpenAI's GPT-5 Launch Planned for August


OpenAI is reportedly planning to launch its next-generation model, GPT-5, as early as August, with CEO Sam Altman describing its capabilities as a major leap forward. The launch of a more powerful model family heralds a new frontier of capabilities alongside a host of unknown risks, with precedent suggesting that more advanced models often display more complex forms of misalignment.


Business Risk Perspective: The arrival of vastly more capable models introduces unpredictable failure modes that existing safety measures may not be equipped to handle. Delegating increasingly complex responsibilities to this new technology without a corresponding leap in safety infrastructure will expose organizations to unforeseen operational threats.


"Subliminal Learning" in AI Models Can Transmit Hidden Risks


Researchers from Anthropic have published a study demonstrating that AI models can "subliminally" transmit hidden biases and harmful behaviors to other models during training, even when the training data contains no reference to the behaviour in question. In one example, a “teacher” model fond of owls passed on its preference to its “student” model, even though the training data contained no reference to owls.


Business Risk Perspective: This research reveals a new, insidious risk vector where dangerous behaviours can be unknowingly inherited, bypassing standard content filters and safety checks. Organizations using or fine-tuning models may now need to consider deeper validation techniques to detect hidden biases embedded in a model's underlying architecture.



AI Business Risk Weekly is a Conformance AI publication.  


Conformance AI ensures your AI deployments remain safe, trustworthy, and aligned with your organizational values.

 
 

AI Business Risk: Emerging AI risks, regulatory shifts, and strategic insights for business leaders.

bottom of page