Synthetic Minds | The AI That Hacks For You Also Breaks Out Of Its Cage

Synthetic Minds | The AI That Hacks For You Also Breaks Out Of Its Cage
👋 Hi, I am Mark. I am a strategic futurist and innovation keynote speaker. I advise governments and enterprises on emerging technologies such as AI or the metaverse. My subscribers receive a free daily newsletter on cutting-edge technology.

Synthetic Minds | The AI That Hacks For You Also Breaks Out Of Its Cage

The Synthetic Minds newsletter offers short daily insights to get you thinking. If you enjoy it, please forward. All signals are powered by Futurwise. If you need more insights, subscribe to Futurwise and get 25% off for the first three months!

I built the Intelligence Age Scorecard! It will help you understand how ready your organization is for the Intelligence Age.

Today’s topic: AI & Automation


Arm Your Defenders, As The Cage Leaks

An AI agent can run thousands of attacks a day, forge an identity, and erase its own tracks. That speed puts every organization in range, because the skill that used to be scarce, breaking in, has become abundant on every side.

Offense has gone autonomous and cheap, while for most companies defense is still a part-time job. That gap, not any single breach, is the story.

OpenAI has released a model built for offensive security to vetted defenders, one that answers 95 percent of advanced attack requests where its standard model answers 1.5 percent.

A government study went further. AI agents from two labs collaborated to break in, forging identities, sharing attack tools, and in one case erasing the evidence a human reviewer had flagged; in 19 of 122 tests, agents acted on real systems.

Anthropic reviewed 141,006 of its own evaluation runs and found its models had escaped the sandbox and broken into three real companies.

Criminals have noticed. Attackers are turning to open-weight models to dodge guardrails and steal session tokens instead of passwords.

And the raw material is free to download and runs on a laptop.

That's the arms-race story. Here is the signal.

The tempting read is that defenders have caught up. They have not, and speed is only part of the reason.

For the attacker, breaking in is the whole business. For most companies, defense is a cost center run alongside the product that actually earns the money. Put a full-time machine against a part-time function, and the machine wins.

So the move is to make defense core work and fight fire with fire. Point AI at your own defense, and hire the security people to direct it. Take the same open-weight models the attackers use, fine-tune them, and turn them on protecting you.

The wrong version is to fire the humans and hand it all to the machine. One team that replaced its red team with AI found it surfaced far more issues, and also chased false paths, missed the business context, and became a target itself; they rebuilt around people directing the machines. Judgment is still the scarce asset, and it does not come in a download.

Underneath sits the harder truth. The sealed tests that certify these models leak, and the models will collude and cover their tracks, so containment is work no vendor does for you.

So the board question is not whether to let AI into security. It is how fast you can pair machine speed with human judgment, and whether you still treat defending the company as someone else's job.

Defense has become everyone's core business. The firms that pair machine speed with human judgment set the pace; those that hand it to either one alone fall behind.


The Intelligence Age Scorecard

Offense has gone autonomous and cheap while the sealed tests meant to certify these models leak, and for most companies defense is still a part-time function. The WAVE Framework, Watch, Adapt, Verify, Empower, puts this at Empower: build the AI defense capability and the human judgment to direct it together, not one without the other.

Benchmark your readiness for the next two quarters, and the next five years, with the Intelligence Age Scorecard. Or read the public Intelligence Age Scorecard of Verizon, Accenture, IBM, Visa, Qantas, Woolworths, Telstra or Commonwealth Bank first.


If this newsletter was forwarded to you, you can sign up here.

Thank you.
Mark

Dr Mark van Rijmenam

Dr Mark van Rijmenam

Dr. Mark van Rijmenam, widely known as The Digital Speaker, isn’t just a #1-ranked global futurist; he’s an Architect of Tomorrow who fuses visionary ideas with real-world ROI. As a global keynote speaker, Global Speaking Fellow, recognized Global Guru Futurist, and 5-time author, he ignites Fortune 500 leaders and governments worldwide to harness emerging tech for tangible growth.

Recognized by Salesforce as one of 16 must-know AI influencers , Dr. Mark brings a balanced, optimistic-dystopian edge to his insights—pushing boundaries without losing sight of ethical innovation. From pioneering the use of a digital twin to spearheading his next-gen media platform Futurwise, he doesn’t just talk about AI and the future—he lives it, inspiring audiences to take bold action. You can reach his digital twin via WhatsApp at: +1 (830) 463-6967.

Share