LFM2.5-VL-DSpark Announced with Speculative Decoding for Vision-Language Models
The company has released an experimental DSpark draft model for their vision-language model (VLM) LFM2.5-VL-3B, featuring speculative decoding to enhance inference speedup on both CPU and GPU without affecting output quality.