Liquid AI released an experimental speculative-decoding draft model (LFM2.5-VL-DSpark) for its LFM2.5-VL-3B vision-language model
Model release Provisional 84% confidence first seen
Liquid AI released an experimental DSpark draft model, LFM2.5-VL-DSpark, designed to accelerate speculative decoding for its LFM2.5-VL-3B vision-language model while aiming to keep output quality unchanged. The coverage reports that the draft model improves decoding speed (up to 3.13x on Apple silicon) and is available via Hugging Face weights with support in tools such as SGLang, MLX-VLM, and llama.cpp.