← Back to feed
WritingArticle

AI science agents still need the wet lab in the loop

Anthropic’s Claude biology work is strongest as an operating pattern: agents can scan huge search spaces, but validation still decides whether discovery happened.

SourceClaude discovers a novel enzyme systemanthropic.com

Anthropic’s new life-sciences post is easy to overread as “AI discovers biology.” The more useful read is operational: Claude was used to search large DNA datasets, generate hypotheses, and identify a previously uncharacterized enzyme system that human scientists then tested in the lab.

The reported system, array-associated reverse transcriptases, or ART, was found near DNA repeat arrays with properties reminiscent of CRISPR-like systems. Anthropic says Claude identified the pattern with high-level direction from scientists, produced a report for expert review, and the company then ran supporting lab experiments. The lab itself operates at BSL-1 and BSL-2 and does not handle human-infecting pathogens, according to the post.

Grey Haven’s read: this is a serious workflow signal, not a reason to let agents free-run science. The model is valuable because it can search a huge hypothesis space, notice candidate patterns, and compress expert review time. The discovery only becomes meaningful when the chain includes provenance, expert inspection, experimental design, wet-lab validation, and clear safety boundaries.

For operators in biotech, pharma, materials, diagnostics, and industrial R&D, the practical question is where AI agents can improve the search loop without weakening scientific accountability. Good candidates are literature triage, sequence search, anomaly detection, assay planning support, and report generation. Bad candidates are unreviewed claims, opaque provenance, and automation that outruns the organization’s ability to validate.

The watch item is whether labs report reproducible lift: more validated leads per scientist, shorter cycles from hypothesis to assay, and fewer dead-end experiments. If the metric is just “the model found something interesting,” skepticism is still the correct default.

Source: Anthropic, “Claude discovers a novel enzyme system,” September 23, 2026.

Grey Haven
Grey HavenApplied AI Venture Studio