Given that these are machine learning algorithms their performance will very much depend on the training dataset. So it is probably not (just) that “technology has moved on a lot”, but that the engineers working on it curated new training sets. It is not entirely unreasonable to think that they too read the paper you are talking about and made measures in an attempt to correct for the effect.