Liquid AI releases LFM2.5-VL-3B-DSpark drafter for faster vision-language decoding
Original titleLiquidAI/LFM2.5-VL-3B-DSpark
AISummary
Liquid AI released LFM2.5-VL-3B-DSpark, a speculative-decoding draft model for its LFM2.5-VL-3B vision-language model. The source reports decoding up to 2.66× faster on a single H100 with SGLang, up to 3.13× on Apple M5 Max with MLX-VLM, and up to 2.14× on Apple M3 Ultra with llama.cpp, with output unchanged under greedy decoding.
Source: Liquid AI · new models on Hugging Face · huggingface.coPublished · added here