Never run EmbeddingGemma 2 in FP16: use BF16 or FP32 to avoid NaNs

tomaarsen · x · 2026-10-07

tomaarsen warns that EmbeddingGemma 2 must run in BF16 or FP32, not FP16: the model's activation range exceeds FP16's dynamic range, and the model card warns of NaNs or silently degraded embeddings. BF16 is recommended where natively supported; FP32 elsewhere, including most CPUs. Sentence Transformers users: pip install -U sentence-transformers transformers.

Original post →

More from Models

Models channel →