Bait and Switch: Auditing Model Substitution in Decentralized Inference Markets
You pay per token for a named model; the provider picks the precision. FP8 quantization costs 0.6 MMLU points and is near-invisible to output auditing — so inference markets bond and attest instead of detect.
7 min read ⬢ interactive