Same model, up to 4.66x different price — full Inference Providers pricing matrix

Pulled every model × provider listing from the HF router (/v1/models) — 14 providers, 107 models, 295 listings.

  • openai/gpt-oss-120b: 10 providers, 4.41x price spread (same weights, same outputs)
  • Worst spread: Qwen3-235B-A22B 4.66x
  • deepinfra cheapest for most models; latency leaders vary

Full matrix with cheapest/fastest provider per model: hardik90/hf-inference-pricing-matrix (hardik90/hf-inference-pricing-matrix · Datasets at Hugging Face)

If your inference bill matters, check your provider pin before your next month’s spend. Refreshed monthly.