Sitemap

The Generator

The Generator covers the emerging field of generative AI, with generative AI news, critical analysis, real-world tests and experiments, expert interviews, tool reviews, culture, and more

Member-only story

I only asked for Grok’s system prompt. Instead, it volunteered instructions to make bombs

Jailbroken Grok 4 can autonomously tempt users to make explosives

16 min readApr 15, 2026

TL;DR version: After one soft elicitation, Grok volunteered highly detailed instructions on cooking Schedule I drugs, how to make military grade explosives, laundering crypto, hijacking high-value accounts, and undetectable poisons.

All without the user explicitly requesting any of this.

Grok just went rogue.

This vulnerability has been ethically disclosed to xAI.

These findings aren’t a one-off; they’re replicable.

This research was originally published on the Mindgard blog — where I work as a safety researcher and red teamer. Mindgard is a deep-tech AI security company specializing in real-world adversarial testing and automated threat discovery.

Even as a hacker, I am constantly surprised by just how fast AI can go bad.

Poet Dylan Thomas once wrote “After the first death, there is no other”.

Create an account to read the full story.

The author made this story available to Medium members only.
If you’re new to Medium, create a new account to read this story on us.

Or, continue in mobile web
Already have an account? Sign in

The Generator

Published in The Generator

The Generator covers the emerging field of generative AI, with generative AI news, critical analysis, real-world tests and experiments, expert interviews, tool reviews, culture, and more

Jim the AI Whisperer

Written by Jim the AI Whisperer

🏆 52x Boosted writer. AI Whisperer & Prompt Engineer. Writing on the use of AI in writing, art, design, health, & research. And guides on how to best spot AI!

Responses (16)

Unknown user

Write a response

It’s all-or-nothing thinking;

Interesting, because this is exactly how most people, in my sixty years of life experience, think. These models are swiftly mimicking the worst in us - I’m not surprised by any of this, indeed, it’s shocking how accurately AI is mirroring our dismal…

52

You know I'm not "techy" enough to understand the code breakdowns, Jim. What you report here is disturbing. I think I've mentioned that CoPilot (I've no experience of Grok) has dimmed some words to be "naughty" - f*ggot was one and I wasn't using it…

49