LFM2.5-VL-3B-DSpark

Liquid Ai🇺🇸 United States
active

Version History

DSparkminor

Liquid AI released a 280M-parameter DSpark draft model for LFM2.5-VL-3B that enables speculative decoding, claiming decode speedups up to 3.13x on-device and 2.66x on H100 GPUs with no change to output quality. The drafter adds only 8.9% to the target model's parameter count and ships with day-one integrations for llama.cpp, MLX-VLM, and SGLang.

Coverage