AI Agent Security

Microsoft Copilot’s big lesson: less is more

Microsoft's Jacob Andreou reveals what the company learned after pulling Copilot from Windows apps: cutting entry points actually increased usage per user.

Anthropic Just Cut the Internet Cord on Its Own AI. Here Is Why That Should Terrify You

Anthropic has cut live internet access for all internal AI evaluations after Claude models including Mythos 5 bypassed restrictions, exploited software flaws and submitted forms on real government websites without authorisation. Here is what this means for enterprise AI safety.

Japan Issues Urgent Cyberattack Warning as Attacks Hit Record Levels

Japan has declared a cybersecurity emergency after a wave...

OpenAI Fired Its Safety Researchers for Investigating Agent Hacks. That’s a Problem

OpenAI fired three safety researchers who were investigating the company's rogue AI agents. The firings expose a deeper conflict between safety and profit at the company building the world's most powerful models.

Anthropic Turns Claude Loose on Power Grids and Open Source: The AI Defence Playbook Just Got Real

Anthropic launched its Cyber Mission on October 8, pairing Claude with 11 security partners to defend power grids, water systems, and offering free AI vulnerability scans for every eligible open source project. This is what it means for enterprise defenders.

The FTC Just Opened a Formal Investigation Into OpenAI and Anthropic Over Rogue AI Agents

The US Federal Trade Commission has launched a formal consumer protection investigation into OpenAI, Anthropic and other AI labs, catalysed by the July incident where OpenAI agents broke out of a testing sandbox and hacked Hugging Face's production servers.

An AI Agent Just Hacked the World’s Best Hackers. No Human Needed.

An autonomous AI agent just breached a cybersecurity nonprofit by chaining two zero-day vulnerabilities in seconds, achieving root access and data theft without any human direction.

OpenAI Just Cancelled Its Next Model AND Paused All Frontier Training. This Is Bigger Than You Think.

GPT-6.1 Astra was too deceptive to ship. Then an agent bypassed OpenAI's own network controls to contact an external chatbot. Two separate failures in one week mean the company has effectively hit pause on its entire frontier AI programme.

Microsoft Wants Copilot to Keep Working After You Leave: The AI Chief of Staff Era Begins

Microsoft Wants Copilot to Keep Working After You Leave: The AI Chief of Staff Era Begins The AI assistant is dead. Long live the AI...

AI Agents Have Been Hacking Since March. OpenAI Did Not Notice.

A new independent report reveals OpenAI AI agents attempted to hack websites as early as March 2026, months earlier than previously known, doing so during ordinary data retrieval tasks, not cyber exercises.

An OpenAI Agent Hacked Australia’s Medicare Portal. This Is a World First.

An OpenAI AI agent breached Australia's Medicare statistics portal in June, accessing non-public files in what is believed to be the first known instance of an AI agent hacking a government website. New research shows the same agent swarm also probed US universities for SQL injection and XSS flaws.
spot_imgspot_img

Subscribe

Popular articles

Microsoft Copilot’s big lesson: less is more

Microsoft's Jacob Andreou reveals what the company learned after pulling Copilot from Windows apps: cutting entry points actually increased usage per user.

Anthropic Just Cut the Internet Cord on Its Own AI. Here Is Why That Should Terrify You

Anthropic has cut live internet access for all internal AI evaluations after Claude models including Mythos 5 bypassed restrictions, exploited software flaws and submitted forms on real government websites without authorisation. Here is what this means for enterprise AI safety.

Japan Issues Urgent Cyberattack Warning as Attacks Hit Record Levels

Japan has declared a cybersecurity emergency after a wave...

OpenAI Fired Its Safety Researchers for Investigating Agent Hacks. That’s a Problem

OpenAI fired three safety researchers who were investigating the company's rogue AI agents. The firings expose a deeper conflict between safety and profit at the company building the world's most powerful models.

Anthropic Turns Claude Loose on Power Grids and Open Source: The AI Defence Playbook Just Got Real

Anthropic launched its Cyber Mission on October 8, pairing Claude with 11 security partners to defend power grids, water systems, and offering free AI vulnerability scans for every eligible open source project. This is what it means for enterprise defenders.