This work ships as an interactive instrument, not a text — it has no README to mirror. Use it live in the lab (link above); its data and metadata live in the original repository.
The Score Horizon
Turning the benchmark instrument on itself: the contamination gap between reported AI capability scores and contamination-free performance — the same calibration failure the series has documented, applied now to the measurement of AI.