LFM2.5-DSpark: Up to 3.2x Faster Inference from H100 to MacB
Liquid AI released DSpark draft model checkpoints for three LFM2.5 models, enabling speculative decoding that speeds up inference by up to 3.18x on GPUs and 2.87x on-device without changing output qua…