cd/entity/DSpark· home› entities› DSpark
grep -l @dspark /news/*.json | wc -l → 25

DSpark

mentions 25 type Organization page 1/2 feed RSS

// recent coverage 25 mentions

14:14
2026-09-24
runtimewire.com
artificial-intelligence

Liquid AI adds a 280M draft model to speed up its 3B vision model

Liquid AI released an experimental 279.5-million-parameter draft model, DSpark, for its LFM2.5-VL-3B vision-language model on September 24th, which the company says cuts decoding time by up to 3.13x o…

00:00
2026-09-11
mindstudio.ai
large-language-models

MiniCPM5-2B: A 2B Open Model That Beats 4B Rivals

OpenBMB released MiniCPM5-2B, a 2.52 billion parameter open-weight language model that posts an average benchmark score of 53.9 across nine capability axes, ahead of Qwen3.5-4B (51.1), granite-4.2-3B …

14:00
2026-08-31
kdnuggets.com
large-language-models

Speed Up LLM Inference with DSpark Speculative Decoding

DeepSeek's DSpark speculative decoding technique, which combines parallel drafting with a lightweight sequential component, can improve local LLM generation speed on the same GPU, with DeepSeek report…

16:52
2026-08-20
huggingface.co
artificial-intelligence

Up to 3.2x Faster Inference with LFM2.5-DSpark

Liquid AI released DSpark draft model checkpoints for its LFM2.5 family, claiming up to 3.18x throughput improvement on GPU and up to 2.87x on-device, with day-one support for llama.cpp and SGLang. Th…

00:00
2026-08-14
mindstudio.ai
large-language-models

How to Run DeepSeek V4 Pro Locally with vLLM or SGLang

DeepSeek released DeepSeek-V4-Pro-0813, the production successor to its V4 Pro preview, under an MIT license with open weights, adding a DSpark speculative decoding module and stronger agentic benchma…

02:22
2026-08-10
github.com
artificial-intelligence

DeepSeekV4SSD: DeepSeek-V4-Flash-0731 on an M-series Mac

DeepSeekV4SSD, an experimental app from developer yanun0323, streams routed experts from SSD to run all 284 billion parameters of DeepSeek-V4-Flash-0731 on an M-series Mac with about 30 GB of memory, …

15:44
2026-07-27
vllm.ai
artificial-intelligence

Kimi K3 on vLLM: Up to 370 Tokens/sec

VLLM announces efficient day-0 support for Moonshot AI's Kimi K3, a 2.8-trillion-parameter Mixture-of-Experts model, achieving up to 370 tokens per second with speculative decoding on 16 NVIDIA GB300 …

07:46
2026-07-25
promptcube3.com
large-language-models

DSpark: Solving LLM Inference Bottlenecks

DSpark introduces a sharding strategy that reduces GPU communication overhead to solve the KV cache memory bloat in long-context LLM inference, targeting latency spikes in distributed setups. The appr…

18:14
2026-06-30
cryptobriefing.com
artificial-intelligence

DeepSeek’s DSpark complicates Nvidia’s latest hardware deals

DeepSeek launched DSpark, an open-source speculative decoding module that boosts AI inference speed by up to 400% on existing chips, potentially reducing demand for Nvidia's high-end accelerators. The…

06:29
2026-06-30
venturebeat.com
large-language-models

DeepSeek Open Sources DSpark

Chinese AI firm DeepSeek open-sourced DSpark, a speculative decoding system that accelerates large language model inference by up to 85% without altering output quality, releasing it under the MIT lic…

page 1 / 2 next →
// co-occurs with top 8 entities
// topics top 6 topics