Liquid AI releases d1-omni-600M, a 600M multimodal model for on-device tasks.
Original titled1-omni-600M is an experimental 600M-parameter model for text + image or text + audio.
AISummary
Liquid AI has released d1-omni-600M, an experimental 600M-parameter model that handles text plus image or audio input. It combines LFM2.5-Encoder-350M with vision and audio encoders and leads the company's text benchmark comparison on toxicity detection and paraphrase identification. The post suggests uses such as voice-command routing, on-device moderation, and intent classification.
Source: Liquid AI · x.comPublished · added here