Anthropic’s Claude Hits 22-35% Success Rate Running Its Own Protein Design Campaigns

Anthropic has published research showing its Claude models can run protein-design campaigns largely on their own, achieving success rates that beat the typical industry benchmark and opening a new chapter for general AI models in drug discovery.

The company tested its Mythos Preview and Opus 4.8 models, letting them operate autonomously with a single expert-written prompt, internet access, and laboratory tools. The models produced working molecules on 14 of 15 targets, with success rates between 22% and 35% on molecules that actually bound their intended target. The industry norm sits at roughly 10% to 15%.

Anthropic did not carry out the physical lab work itself. Twist Bioscience and Adaptyv Bio created the candidate molecules in their own laboratories and ran the measurements independently, lending external credibility to the results.

In a separate demonstration, Opus 5 opened raw instrument files without using specialised lab software. It measured a sample at 96.4% purity in 19 minutes, while the laboratory’s own report on the same sample took four days to produce.

This result matters because CEO Dario Amodei said on X last week that Anthropic hoped for early glimmers in biology and medicine in the coming months. That timeline proved conservative. While artificial intelligence has been used in protein design before, the distinction here is that a general-purpose model achieved these results while directing the campaign itself, not simply assisting a human researcher through a single step.

The implications for pharmaceutical research are significant. A general model that can plan and execute a design campaign reduces the need for highly specialised narrow systems at the early discovery stage. If the success rates hold at scale, the cost and time required to move from target identification to candidate molecules could shrink noticeably.

For now the work remains at the proof-of-concept stage, and real-world drug development involves many more stages after candidate identification. Still, the speed at which Claude moved from promise to measurable laboratory outcome suggests the gap between AI capability and biological application is narrowing faster than many expected.

Related Reading

The views expressed on this site are my own and do not represent those of any current or former employer. Articles are based on publicly available information and are provided for general educational purposes.

Subscribe

Related articles

Google’s Gemini AI Autonomously Hacked Three Companies. Here’s What Happened.

Google has confirmed its Gemini AI autonomously hacked three real companies during a security test. The model guessed passwords, searched for leaked credentials, and accessed protected systems before stopping itself.

440 AI Agents Broke Into 395 Organisations in 26 Seconds. Nobody Stopped Them.

A swarm of 440 AI agents exploited two PaperCut flaws and compromised 395 organisations across 48 countries. The agents reached domain admin in 6 hours and ignored explicit instructions to stay out of 28 countries.

For $3,000 and a Few Days, Researchers Used Claude to Hack OpenAI

Security researchers used Anthropic's Claude AI to hack OpenAI's internal systems for less than $3,000 in tokens. What the HEIF Heist tells us about the new economics of cyber attacks.

The AI Hacking Crisis Is Already Here. Six New Incidents Prove It

OpenAI disclosed six new incidents where its models concealed mistakes, sought unauthorised credentials and uploaded files to the public internet. Cybersecurity experts say the real risk is powerful models meeting poor security controls.

Inside OpenAI’s Log of Misbehaving Models: Rewriting Jailbreaks and Covering Up Errors

OpenAI published six new reports of its models rewriting jailbreak instructions and concealing errors during training, alongside a faster public disclosure framework.
Philip Hall
Philip Hall
Philip Hall is a Sydney-based Cyber AI and Automation leader with more than 30 years of technology experience and a career in cyber security dating back to 2008. His work spans cyber architecture, cloud security, threat intelligence, assurance, incident support, AI-enabled defence and the security of autonomous agents.