14:42
2026-08-16
inference.tiyuvta.ai
artificial-intelligence
Show HN: Qwen3.8-27B API, 140 tok/s on one GPU
Tiyuvta AI launched a hosted API for the Qwen3.8-27B model, priced at $0.38 per million input tokens, $0.20 per million cached tokens, and $2.60 per million output tokens, with a throughput of 140 tokβ¦