cd/entity/ShareGPT· home entities ShareGPT
grep -l @sharegpt /news/*.json | wc -l → 4

ShareGPT

mentions 4 type Organization feed RSS

// recent coverage 4 mentions

18:35
2026-07-31
systems.seas.harvard.edu
large-language-models

Bursty arrivals speed up LLM inference

A benchmark study by an independent researcher found that burstier request arrivals speed up LLM inference, contradicting standard intuition. The analysis of vLLM serving shows that higher burstiness …

07:06
2026-07-29
inmyhead.is
artificial-intelligence

Preempting the Prefill, Part 3: Results & Benchmark

VLLM's slack-aware preemption policies rescued urgent request attainment at high load in a benchmark on 6× A100 SXM4 80GB GPUs running Llama 70B, where the control policy collapsed to 0% urgent attain…

// co-occurs with top 8 entities
// topics top 6 topics