"""Measure real MTP acceptance for a model by parsing the server's MTP[ lines. Controlled A/B: identical prompts, identical sampler, one model then the other. Prompts are deliberately NOT drawn from the training corpus -- the head was fitted to that distribution, so measuring on it would flatter the trained head. These are fresh held-out prompts in the model's normal serving register. Usage: python3 bench_accept.py