Meta's hybrid GPU-CPU retrieval serves personalization over a billion docs plus 20x CPU inventory

_reachsumit · x · 2026-09-21

Meta details a production-deployed hybrid GPU-CPU co-serving system for ultra-large-scale personalized search. A deep GPU pathway fuses retrieval and interaction pre-ranking over a billion-document pool, while a breadth-first CPU pathway searches an inventory 20x larger with lightweight personalized scoring. A full-system A/B test against the legacy CPU-only setup improved both model-scored relevance and substantive engagement.

Original post →

More from Infra

Infra channel →