CLSS: Mapping Protein Sequence and Structure into a Shared AI Space

bravo_abad · x · 2026-08-07

Guy Yanai and coauthors introduce CLSS, a contrastive protein language model that learns a shared representation of protein sequence, structure, and even subsequences.

Architecture & Training

Results & Significance

CLSS goes beyond multimodal alignment by learning a unified "protein space." In this space, sequence and structure yield nearly the same global organization, successfully recovering expert-curated evolutionary hierarchies that the model never saw during training.

Original post →

More from Research

Research channel →