LLMs1 min read
LFM2.5-VL-DSpark speeds up vision-language models with minimal memory increase
A new draft model for vision-language tasks offers faster decoding on devices and GPUs, with small memory costs and day-one support for popular frameworks.
From Hugging Face blog




