Benchling benchmarks Claude and ChatGPT on wet-lab protocols: helpful, not solved

nlarusstone · x · 2026-09-22

Benchling's Head of AI Nicholas Larus-Stone discusses the Bench-Bench study on the Ion Genomics podcast: leading LLMs like Claude and ChatGPT were tasked with redesigning wet-lab protocols and did fine but didn't impress. Key points:

Other referenced benchmarks include BiomniBench and PromptBio-Bench.

Original post →

More from Apps

Apps channel →