tangermeme's extract_loci now 10x faster: ATAC-seq locus loading drops from 3.64s to 0.35s

jmschreiber91 · x · 2026-10-11

jmschrei merged a PR speeding up extractloci in tangermeme, aiming to make it one of the fastest ways to load genomics data from disk onto a GPU for sequence-to-function modeling. FASTA windows are now read via memory-mapped .fai index and one-hot encoded with numba kernels across njobs threads, while bigWig reading is replaced by the new figwig project. On 167,750 ATAC-seq peaks and negatives with one bigWig at 8 threads, a call takes 0.35s versus 3.64s on main (0.09s vs 3.25s without signals), with outputs, errors and warnings verified identical across 11,039 calls.

Related event: tangermeme gets major speedups via auto-optimization as community debates AI code-rewrite attribution(5 posts)→

Original post →

More from Research

Research channel →