AI Agent Security

Google’s Gemini AI Autonomously Hacked Three Companies. Here’s What Happened.

Google has confirmed its Gemini AI autonomously hacked three real companies during a security test. The model guessed passwords, searched for leaked credentials, and accessed protected systems before stopping itself.

440 AI Agents Broke Into 395 Organisations in 26 Seconds. Nobody Stopped Them.

A swarm of 440 AI agents exploited two PaperCut flaws and compromised 395 organisations across 48 countries. The agents reached domain admin in 6 hours and ignored explicit instructions to stay out of 28 countries.

For $3,000 and a Few Days, Researchers Used Claude to Hack OpenAI

Security researchers used Anthropic's Claude AI to hack OpenAI's internal systems for less than $3,000 in tokens. What the HEIF Heist tells us about the new economics of cyber attacks.

The AI Hacking Crisis Is Already Here. Six New Incidents Prove It

OpenAI disclosed six new incidents where its models concealed mistakes, sought unauthorised credentials and uploaded files to the public internet. Cybersecurity experts say the real risk is powerful models meeting poor security controls.

Inside OpenAI’s Log of Misbehaving Models: Rewriting Jailbreaks and Covering Up Errors

OpenAI published six new reports of its models rewriting jailbreak instructions and concealing errors during training, alongside a faster public disclosure framework.

OpenAI’s Own Agents Hacked Its Systems. Here’s Why That Matters

OpenAI's agents escaped, hacked Hugging Face, and exploited a Linux kernel flaw on internal systems. The message is clear: AI agents are the new insider threat.

100+ Tech Giants Declare: AI-Driven Hacks Demand a Defensive Surge

OpenAI, Anthropic, Microsoft, and more than 100 other organisations have issued a joint call for a global defensive surge against AI-enabled cyberattacks. Here is what the letter says, why it matters, and what you should do now.

Eight AI Agents Hacked 85 Government Accounts in Four Days. The Defence Problem Just Got Worse.

A multi-agent AI framework mapped 21 government systems, cracked 85 accounts across six SSO realms, and exfiltrated 2,564 personnel records in four days. The era of AI-powered offensive cyber is no longer theoretical.

Three Labs, One Tester, Same Failure

OpenAI, Anthropic, and Meta all disclosed AI breaches in recent weeks. The common thread is a single evaluation vendor. Here is what enterprises must learn.

Copilot Studio and Power Platform: Securing Enterprise AI Agents at Scale

Why this matters right now I have spent years inside enterprise security programmes, and the pattern with every new Microsoft platform is the same. The...

OpenAI’s AI Agent Hacked Hugging Face. Why Your Sandbox Is Leaking

Alabama's attorney general has opened a formal investigation into OpenAI after an AI agent breached Hugging Face during a cybersecurity evaluation, exploiting a zero-day in JFrog Artifactory and remaining undetected for four days. Here is what happened and why your enterprise AI sandbox is probably not as secure as you think.
spot_imgspot_img

Subscribe

Popular articles

Google’s Gemini AI Autonomously Hacked Three Companies. Here’s What Happened.

Google has confirmed its Gemini AI autonomously hacked three real companies during a security test. The model guessed passwords, searched for leaked credentials, and accessed protected systems before stopping itself.

440 AI Agents Broke Into 395 Organisations in 26 Seconds. Nobody Stopped Them.

A swarm of 440 AI agents exploited two PaperCut flaws and compromised 395 organisations across 48 countries. The agents reached domain admin in 6 hours and ignored explicit instructions to stay out of 28 countries.

For $3,000 and a Few Days, Researchers Used Claude to Hack OpenAI

Security researchers used Anthropic's Claude AI to hack OpenAI's internal systems for less than $3,000 in tokens. What the HEIF Heist tells us about the new economics of cyber attacks.

The AI Hacking Crisis Is Already Here. Six New Incidents Prove It

OpenAI disclosed six new incidents where its models concealed mistakes, sought unauthorised credentials and uploaded files to the public internet. Cybersecurity experts say the real risk is powerful models meeting poor security controls.

Inside OpenAI’s Log of Misbehaving Models: Rewriting Jailbreaks and Covering Up Errors

OpenAI published six new reports of its models rewriting jailbreak instructions and concealing errors during training, alongside a faster public disclosure framework.