# Liquid AI releases DSpark draft models for faster LFM2.5 inference

_Thursday, August 20, 2026 at 12:52 PM EDT · AI · Latest · Tier 2 — Notable_

Liquid AI released DSpark draft models for its LFM2.5 family and said they deliver up to 3.18-times GPU throughput improvement and up to 2.87-times on-device improvement. The company said the models have day-one integrations for llama.cpp and SGLang. Its measurements used a single H100 80 GB GPU for SGLang and an M4 Max MacBook Pro for llama.cpp, with batch size one and temperature zero. Liquid AI says speculative decoding preserves greedy output because the target model verifies proposed tokens.

## Sources

- [Hugging Face - Blog](https://huggingface.co/blog/LiquidAI/lfm25-dspark)

---
Canonical: https://techandbusiness.org/newswire/SdiYvzIS6Ve0UH3__2i2WW
Retrieved: 2026-08-21T00:08:32.325Z
Publisher: Tech & Business (techandbusiness.org)
