Liquid AI details d1-omni-600M: 600M params for text+image or text+audio

JosephJacks_ · x · 2026-10-08

Liquid AI's experimental d1-omni-600M combines LFM2.5-Encoder-350M with vision and audio encoders, accepting text+image or text+audio. It leads their text benchmarks in toxicity detection and paraphrase identification, targeting voice-command routing, on-device moderation, and intent classification.

Original post →

More from Models

Models channel →