cd/entity/DSpark· home entities DSpark
grep -l @dspark /news/*.json | wc -l → 17

DSpark

mentions 17 type Organization feed RSS

// recent coverage 17 mentions

02:22
2026-08-10
github.com
artificial-intelligence

DeepSeekV4SSD: DeepSeek-V4-Flash-0731 on an M-series Mac

DeepSeekV4SSD, an experimental app from developer yanun0323, streams routed experts from SSD to run all 284 billion parameters of DeepSeek-V4-Flash-0731 on an M-series Mac with about 30 GB of memory, …

15:44
2026-07-27
vllm.ai
artificial-intelligence

Kimi K3 on vLLM: Up to 370 Tokens/sec

VLLM announces efficient day-0 support for Moonshot AI's Kimi K3, a 2.8-trillion-parameter Mixture-of-Experts model, achieving up to 370 tokens per second with speculative decoding on 16 NVIDIA GB300 …

07:46
2026-07-25
promptcube3.com
large-language-models

DSpark: Solving LLM Inference Bottlenecks

DSpark introduces a sharding strategy that reduces GPU communication overhead to solve the KV cache memory bloat in long-context LLM inference, targeting latency spikes in distributed setups. The appr…

18:14
2026-06-30
cryptobriefing.com
artificial-intelligence

DeepSeek’s DSpark complicates Nvidia’s latest hardware deals

DeepSeek launched DSpark, an open-source speculative decoding module that boosts AI inference speed by up to 400% on existing chips, potentially reducing demand for Nvidia's high-end accelerators. The…

06:29
2026-06-30
venturebeat.com
large-language-models

DeepSeek Open Sources DSpark

Chinese AI firm DeepSeek open-sourced DSpark, a speculative decoding system that accelerates large language model inference by up to 85% without altering output quality, releasing it under the MIT lic…

20:16
2026-06-27
github.com
machine-learning

GitHub DeepSeek-AI/DeepSpec

DeepSeek-AI released DeepSpec, an open-source codebase for training and evaluating draft models for speculative decoding, supporting three draft model algorithms (DSpark, DFlash, Eagle3) and requiring…

13:27
2026-06-27
byteiota.com
large-language-models

DeepSeek DSpark Goes Live with 80% Inference Speed Gains

DeepSeek released DSpark, a speculative decoding framework now live in its DeepSeek-V4 Flash and Pro production API, delivering 51 to 400 percent throughput gains and up to 80 percent latency reductio…

// co-occurs with top 8 entities
// topics top 6 topics