GamowLabs to launch RareBench challenge: can you tell synthetic genomes from real ones?

danielmckinn0n · x · 2026-09-03

GamowLabs explains why improving clinical genomics agents is hard: real cases are either trivial for frontier models or unsolved and require physician verification, leaving synthetic data as the only systematic path — yet LLMs now spot artifacts left by existing workflows. In one test, a blinded evaluator agent (Codex with GPT 5.6 Sol) compared Reseq2-generated sequencing reads from a synthetic child VCF against real 1000 Genomes data and identified the synthetic sample at 96% self-reported confidence. The team has invested in generating synthetic genomes indistinguishable from real ones and will launch a RareBench challenge asking the community to tell real from synthetic samples.

Original post →

More from Research

Research channel →