cd/entity/BurstGPT· home entities BurstGPT
grep -l @burstgpt /news/*.json | wc -l → 4

BurstGPT

mentions 4 type Organization feed RSS

// recent coverage 4 mentions

18:35
2026-07-31
systems.seas.harvard.edu
large-language-models

Bursty arrivals speed up LLM inference

A benchmark study by an independent researcher found that burstier request arrivals speed up LLM inference, contradicting standard intuition. The analysis of vLLM serving shows that higher burstiness …

15:30
2026-07-27
systems.seas.harvard.edu
large-language-models

MorphServe: Making Model Precision Elastic for Bursty LLM Serving

MorphServe, a system developed by researchers and detailed in a paper on arXiv, enables elastic model precision for bursty LLM serving by temporarily swapping low-impact FP16 layers to INT4 quantizati…

// co-occurs with top 8 entities
// topics top 4 topics