# Liquid AI releases LFM2.5-VL-DSpark for faster vision-language model inference

Digest AI · Generative AI & Models · published 2026-09-24T14:08:57Z

Canonical: https://digestai.news/story/liquid-ai-releases-lfm2-5-vl-dspark-for-faster-vision-language-model-i

## Summary

Liquid AI announced **LFM2.5-VL-DSpark**, a draft model for its **LFM2.5-VL-3B** vision-language model. The model uses speculative decoding to speed up inference by up to **3.13x on-device** and **2.66x on H100 GPUs**, with end-to-end gains of **2.62x and 2.27x**, respectively. It adds **280M parameters** (8.9% increase) but maintains output quality, according to the company’s benchmarks across tasks like VQA, captioning, and multi-turn conversation.

The model supports **day-one integration** with **llama.cpp, MLX-VLM, and SGLang**, targeting edge and GPU deployments. Liquid AI emphasizes its open-weight approach, allowing unrestricted fine-tuning and deployment. The release aligns with the lab’s goal of AI running across devices, from base models to specialized variants like audio and vision.

## Key points

- LFM2.5-VL-DSpark speeds up vision-language model inference by up to 3.13x on-device and 2.66x on H100 GPUs
- Adds 280M parameters (8.9% increase) to LFM2.5-VL-3B with no output quality trade-off, per Liquid AI
- Supports day-one integration with llama.cpp, MLX-VLM, and SGLang for edge and GPU deployment

## Why it matters

Faster vision-language model inference reduces latency for applications like real-time captioning and multi-modal reasoning, critical for edge devices and latency-sensitive workflows.

## Sources

1. [Accelerating vision-language models with LFM2.5-VL-DSpark](https://huggingface.co/blog/LiquidAI/lfm2-5-vl-dspark) (Hugging Face, 2026-09-24, primary source)

Part of the developing story: [Rise of Local Open-Source AI Tools](https://digestai.news/thread/kdnuggets-lists-seven-open-source-chatgpt-alternatives-that-run-locally) (3 stories)

## Cite

Digest AI, "Liquid AI releases LFM2.5-VL-DSpark for faster vision-language model inference", 24 September 2026, https://digestai.news/story/liquid-ai-releases-lfm2-5-vl-dspark-for-faster-vision-language-model-i

---

Written by Digest AI's editorial model from the linked sources; the sources are the record. Headlines, digests and key points are written by Digest AI and may be quoted with a link to the story page. Linked articles belong to their publishers. Terms: https://digestai.news/terms#reuse
JSON: https://digestai.news/story/liquid-ai-releases-lfm2-5-vl-dspark-for-faster-vision-language-model-i.json
