Show HN: Determinstic LLM inference for lowest price Gemma 4, with Windows XP TokenDelivery.ai launched a preview of deterministic LLM inference for the open-weight Gemma 4 model, offering byte-for-byte identical answers on every call and free access with no credit card required during the preview. The service is OpenAI-compatible, streams answers, accepts images, exposes every sampling parameter, and supports streaming, tools, structured outputs, logprobs, seeds, and a token-ids endpoint at base URL https://api.tokendelivery.ai/v1, with list prices per million tokens to apply once billing starts. The company said the desktop interface requires landscape orientation. Open-weight models, fully deterministic. The same answer every time, byte for byte. Free while we're in preview. Sign up in a dialog, make a key in API Keys. It's free while we're in preview: no card. OpenAI-compatible: point any client at the base URL with your key. The playground streams answers, takes pictures, and exposes every sampling parameter. Questions: the Contact window. Free during the preview. List prices per million tokens for when billing starts. Streaming, tools, structured outputs, images, logprobs, seeds. A token-ids endpoint for callers that tokenize themselves. Base URL: https://api.tokendelivery.ai/v1 This desktop is a landscape thing, like it's 2001. Rotate to landscape to use TokenDelivery.ai.