Models served via Hugging Face Inference

5 models from 4 makers that you can call through Hugging Face Inference — each one tested by us with fixed prompts, outputs kept, costs measured.

Makers: black-forest-labs · qwen · stabilityai · tongyi-mai

5 models 🧪 all independently tested

Qwen Image

N/A

SD XL

N/A

FLUX.1-dev

N/A

stable-diffusion-3.5-medium

N/A

Z-Image-Turbo

N/A

About Hugging Face Inference

aggregator

Hugging Face is the largest open-weight model hub — over 100,000 models, most of which can be called through its Inference API without standing up your own infrastructure.

A free tier exists for light usage; dedicated Inference Endpoints are the path to production-grade serving. For a model with no other hosted API, Hugging Face is often the only way to call it without self-hosting.

  • The largest open-weight model hub — most of its 100,000+ models are callable through one Inference API
  • Free tier available; dedicated Inference Endpoints for production workloads
  • Best fit when a model has no other hosted API — this is often the only way to call it

Affiliate program: None found.

Hugging Face Inference pricing, by model

This is what Hugging Face Inference itself charges for each model — not the cheapest price across every provider (that's what the cards above show).

ModelMakerPrice via Hugging Face InferenceContext
black-forest-labs/FLUX.1-devblack-forest-labs
Qwen Imageqwen
stabilityai/stable-diffusion-3.5-mediumstabilityai
SD XLstabilityai
Tongyi-MAI/Z-Image-Turbotongyi-mai

Latency: Per-provider latency isn't independently measured yet — the probe rig currently tests one route per model, not one per provider (see PROBE_RIG.md ledger #10).