Apple's CoGR: Co-Evolving Generative Retriever with Reinforcement Learning

_reachsumit · x · 2026-09-02

Apple researchers propose CoGR, a framework training LLMs to directly construct retrieval representations (keywords) for both query and item sides, rather than just augmenting queries. It matches keywords via inverted index for infrastructure compatibility. The training involves SFT and co-evolving RL, alternately optimizing both sides to maximize retrieval F1. Experiments show CoGR outperforms 10 baselines.

Original post →

More from Research

Research channel →