SentenceTransformers V6 Unifies Text, Vision, and Audio in Single API
ManuelFaysse · x · 2026-08-21
SentenceTransformers V6 is released with a unified API that supports ColPali, ColQwen, and other visual retrieval models out of the box. This update enables training and inference for text, vision, and audio within a single framework. The standalone ColPali repository will be deprecated in favor of the main library, ensuring long-term support from the Hugging Face team.
More from Models
- Ox Alpha Test: Slow and Frequent Stops — kevinkern · 2026-08-21
- Anonymous "Ox Alpha" model researches, builds and QA's its own site in one prompt — Acceptable-Object390 · 2026-08-21
- Ox Alpha Coming Soon, Excitement Builds — zephyr_z9 · 2026-08-21
- Ox-Alpha's take on researcher Janus: real contributions, polarizing figure — jd_pressman · 2026-08-21
- mLateOn matches 8B models on Japanese JMTEB, beats Japan-specific models — IgorCarron · 2026-08-21
- User Report: Opus 5 Regresses, Context Compaction Losing Track — springrod · 2026-08-21