Liquid AI releases LFM2.5-VL-3B, a 3B vision-language model for edge devices
LFM2.5-VL-3B: A Better and Faster Vision-Language Model for the Edge
AISummary
Liquid AI released LFM2.5-VL-3B, an open-weight 3B vision-language model that it says rivals models twice its size while running faster on CPU and GPU. Benchmarks show large gains over LFM2-VL-3B, including ScreenSpot-v2 averaging 80.7, RefCOCO precision@1 rising from 57.1 to 87.9, and ToolSandbox rising from 26.4 to 59.5. The model is available on Hugging Face and decodes 228 tokens/s on an Apple M5 Max.
AIWhy it matters
The post pairs benchmark gains with on-device and GPU throughput figures, showing how a 3B vision model trades size against speed and accuracy.
Source: Liquid AI Blog · liquid.ai