Pulled every model × provider listing from the HF router (/v1/models) — 14 providers, 107 models, 295 listings.
- openai/gpt-oss-120b: 10 providers, 4.41x price spread (same weights, same outputs)
- Worst spread: Qwen3-235B-A22B 4.66x
- deepinfra cheapest for most models; latency leaders vary
Full matrix with cheapest/fastest provider per model: hardik90/hf-inference-pricing-matrix (hardik90/hf-inference-pricing-matrix · Datasets at Hugging Face)
If your inference bill matters, check your provider pin before your next month’s spend. Refreshed monthly.