cd/entity/EngineCoreยท homeโ€บ entitiesโ€บ EngineCore
grep -l @enginecore /news/*.json | wc -l โ†’ 2

EngineCore

mentions 2 type Organization feed RSS

// recent coverage 2 mentions

07:06
2026-07-29
inmyhead.is
large-language-models

A Scheduler as a Lens into LLM Inference

A developer built a small scheduler in Go to understand vLLM's scheduler for LLM inference, tracing each design decision back to its vLLM equivalent. The scheduler operates in a tick loop with three pโ€ฆ

05:46
2026-07-29
inmyhead.is
artificial-intelligence

Preempting the Prefill

A new paper, FlowPrefill by Hsieh et al., proposes preempting long LLM inference prefills mid-forward-pass to rescue urgent requests that would otherwise miss their time-to-first-token (TTFT) service-โ€ฆ

// co-occurs with top 6 entities
// topics top 4 topics