Member-only story
I Scanned the Internet for AI Servers. What I Found Should Terrify Every Founder Using Ollama.
A flaw called “Bleeding Llama” just turned 300,000 local LLMs into open windows. Yours might be one of them.
It was 2:47 AM when the Slack notification hit my phone.
“You need to see this. Now.”
My co-founder almost never messages at that hour. I rolled out of bed, opened my laptop, and stared at a CVE entry that had dropped only minutes earlier on the National Vulnerability Database.
CVE-2026–7482. CVSS 9.1. Codename: Bleeding Llama.
By the time the sun came up, I understood why he had not been able to sleep. The flaw he was pointing at did not just affect a single product. It affected the foundation that thousands of AI startups, indie developers, and enterprise teams had quietly built their entire 2026 roadmaps on.
It affected Ollama.
And if you have ever run a local large language model on your own server, you need to keep reading. Because the version of Ollama you spun up last quarter — the one your team forgot about, the one humming away on a $40-per-month VPS in Frankfurt — is almost certainly leaking its memory to anyone who asks nicely.