Reliable confidence estimation is increasingly central to the trustworthy deployment of language models: a calibrated estimate of the probability that an output is correct decides what to ship, what to escalate, and what to retry. Existing confidence estimators, however, share one design premise: th
LLMjacking: When Stolen AI Tokens Turn Your Account Into Someone Else’s Compute