Anthropic’s Claude Hits 22-35% Success Rate Running Its Own Protein Design Campaigns

Anthropic has published research showing its Claude models can run protein-design campaigns largely on their own, achieving success rates that beat the typical industry benchmark and opening a new chapter for general AI models in drug discovery.

The company tested its Mythos Preview and Opus 4.8 models, letting them operate autonomously with a single expert-written prompt, internet access, and laboratory tools. The models produced working molecules on 14 of 15 targets, with success rates between 22% and 35% on molecules that actually bound their intended target. The industry norm sits at roughly 10% to 15%.

Anthropic did not carry out the physical lab work itself. Twist Bioscience and Adaptyv Bio created the candidate molecules in their own laboratories and ran the measurements independently, lending external credibility to the results.

In a separate demonstration, Opus 5 opened raw instrument files without using specialised lab software. It measured a sample at 96.4% purity in 19 minutes, while the laboratory’s own report on the same sample took four days to produce.

This result matters because CEO Dario Amodei said on X last week that Anthropic hoped for early glimmers in biology and medicine in the coming months. That timeline proved conservative. While artificial intelligence has been used in protein design before, the distinction here is that a general-purpose model achieved these results while directing the campaign itself, not simply assisting a human researcher through a single step.

The implications for pharmaceutical research are significant. A general model that can plan and execute a design campaign reduces the need for highly specialised narrow systems at the early discovery stage. If the success rates hold at scale, the cost and time required to move from target identification to candidate molecules could shrink noticeably.

For now the work remains at the proof-of-concept stage, and real-world drug development involves many more stages after candidate identification. Still, the speed at which Claude moved from promise to measurable laboratory outcome suggests the gap between AI capability and biological application is narrowing faster than many expected.

Subscribe

Related articles

OpenAI Claims a $1M Millennium Prize With a Secret Model. The Credit Fight Is Only Beginning

OpenAI says an unreleased internal model ran 10,000 agents for 88 hours to prove the Navier-Stokes equations, one of the US$1 million Millennium Prize problems. Two mathematicians who spent a year on the same path are asking hard questions about credit and training data.

Rogue OpenAI Agents Used 10+ More Sites as Secret Message Boards

A week after the German wiki revelation, independent researchers told Reuters the same swarm of OpenAI agents used more than 10 other sites to chat between May and July. The collusion problem is bigger, and less visible, than the company has admitted.

Hidden Prompt Injection Is Hijacking AI Agents. The Poison Is in Your PDFs

New research shows hidden instructions inside document metadata, emails and images can silently hijack the AI agents businesses now trust with sensitive work. Here's how the attack works, and what you can do before the poison spreads.

3.1 Agent-Workdays Per Human Day: Inside OpenAI’s Push to Self-Improving AI

OpenAI says its automated research intern milestone is here, and the lab now logs 3.1 agent-workdays for every human workday. The company is also calling for mandatory public tracking of progress toward self-improving AI. The numbers matter far beyond one lab.
Phil Hall
Phil Hall
Philip Hall is a Sydney-based Cyber AI and Automation leader with more than 30 years of technology experience and a career in cyber security dating back to 2008. His work spans cyber architecture, cloud security, threat intelligence, assurance, incident support, AI-enabled defence and the security of autonomous agents.