How many audit probes it takes to catch a substituted model, as a function of how subtle the swap is. A 1/Δ² curve over real quantization accuracy gaps: a model swap is caught in a few hundred probes; FP8 needs ~150k and falls off the cliff into the economically-invisible zone.
You pay per token for a named model; the provider picks the precision. FP8 quantization costs 0.6 MMLU points and is near-invisible to output auditing — so inference markets bond and attest instead of detect.