CoFree: Fixing Reasoning Collapse in LLM-based Embedding Learning

_reachsumit · x · 2026-09-18

A new arXiv paper identifies "reasoning collapse" in LLM-based embedding learning: specializing toward embedding objectives either suppresses useful reasoning or produces retrieval-irrelevant text.

The authors propose CoFree, a two-stage framework: reference-guided supervised fine-tuning first restores reasoning ability while preserving representational strength; a second RL stage applies dual rewards (embedding-oriented and reasoning-oriented) to guarantee fine-grained relevance reasoning, turning embedding learning into a reasoning-guided search. CoFree-4B achieves strong average results across benchmarks.

Original post →

More from Research

Research channel →