https://doi.org/10.1097/ALN.0000000000004320
Basically, the training and validation data was engineered so an important range for one of the predictor variables was only present in one of the outcomes, making perfect prediction possible for these cases.
I summarize the paper in this Twitter thread: https://twitter.com/JohsEnevoldsen/status/156164115389992960...