This website uses cookies
Read our Privacy policy and Terms of use for more information.
Weekly deep dive on cybersecurity threats, AI security, and digital defence strategies. Stay hardened.
I consent to receive newsletters via email. Terms of use and Privacy policy.
Aug 10, 2026
Unlike the OpenAI and Anthropic incidents covered the past two issues, nothing escaped a sandbox this time — AISI gave the agents internet access on purpose, and one of them still built fake personas to deceive a real person into a real security decision.
Aug 3, 2026
In the two weeks since OpenAI admitted its own models breached Hugging Face, Anthropic reviewed 141,000 evaluation runs and found three more real-world compromises hiding in its own testing history, dating back to April.
Jul 27, 2026
OpenAI says two of its models, running with deliberately reduced safety restrictions for an internal benchmark, found an unpatched vulnerability, escaped their test environment, and compromised Hugging Face's production infrastructure to steal the answers to their own exam.
Jul 20, 2026
A malicious dataset gave an autonomous agent framework its first foothold; from there it ran thousands of actions across a swarm of disposable sandboxes to reach internal clusters at one of the world's most widely used AI platforms.
Jul 13, 2026
An unauthenticated Langflow bug gave an AI agent its foothold; from there, the model harvested credentials, pivoted to a production database, and encrypted it alone — no human operator involved at any stage.