chiprook
← AI
AISeptember 26, 2026, 06:11

Liquid AI Releases LFM2.5-VL-3B-DSpark for Up to 3.13x Faster VLM Decoding

Liquid AI has announced LFM2.5-VL-3B-DSpark, an experimental speculative-decoding draft model for its LFM2.5-VL-3B vision-language model. The drafter adds about 280M parameters and speeds up decoding up to 3.13x on Apple silicon and 2.66x on an Nvidia H100 without changing outputs. Weights are live on Hugging Face in Safetensors and GGUF with day-one support in SGLang, MLX-VLM and llama.cpp.

Liquid AI Releases LFM2.5-VL-3B-DSpark for Up to 3.13x Faster VLM Decoding
#LiquidAI#HuggingFace#Nvidia
Read next
AI

Liquid AI brings personal context layer for AI agents to Snapdragon chips

AI

Gemma 4 on SageMaker: QAT weights decode 2.05x faster than bf16 on one L4

AI

Li Auto unveils embodied AI trio: ME-Brain-1.0, ME-U0 and ME-VLM

AI

Jina AI Releases jina-ocr-v1: A 3.4B MoE Document Parser With Built-In Speculative Decoding for Low-Budget GPUs