Liquid AI releases LFM2.5-8B-A1B-DSpark draft model for faster LFM2.5 decoding
Original titleLiquidAI/LFM2.5-8B-A1B-DSpark
AISummary
Liquid AI released LFM2.5-8B-A1B-DSpark, a 327.7M-parameter speculative-decoding draft model for its LFM2.5-8B-A1B target. In SGLang on one H100 with batch size 1, mean accepted tokens per step reached 7.21 across five benchmarks, and decoding ran about 2.6× faster. The model also runs on Apple silicon through the Metal backend, with a 1.18× mean speedup on an M4 Max.
Source: Liquid AI · new models on Hugging Face · huggingface.coPublished · added here