Attention Relay makes embedding models instruction-aware without training via LLM attention weights

_reachsumit · x · 2026-10-06

An arXiv paper proposes Attention Relay, a training-free method that passes an instruction-tuned LLM's attention weights into an embedder's attention pooling, transferring instruction-following ability to text embeddings.

Original post →

More from Models

Models channel →