asked 9 models behind my own gateway to name a random number 1-100, forty times each. no logprobs, no weights, one token per query. then reconstructed the whole routing table from the answer histograms. models are reliably bad at random, and the badness is a fingerprint.