cd/entity/Splitwise· home entities Splitwise
grep -l @splitwise /news/*.json | wc -l → 4

Splitwise

mentions 4 type Organization feed RSS

// recent coverage 4 mentions

09:00
2026-08-11
blog.doubleword.ai
large-language-models

The case for disaggregated LLM serving

Disaggregated LLM serving, which runs prefill and decode on separate GPU pools and transfers KV caches over the network, should always be used in practice under sufficient load, according to a technic…

04:00
2026-08-03
arxiv.org
artificial-intelligence

Topology-Aware Data Movement for Disaggregated GPU Inference

A new arXiv paper (arXiv:2607.28633v1) proposes a topology-aware transfer orchestrator for disaggregated GPU inference, claiming existing systems like DistServe, Splitwise, and Mooncake ignore that ba…

// co-occurs with top 8 entities
// topics top 6 topics