That doesn't affect Jabbles's point in any way.
Run enough tests, and you will get statistically significant but bogus results.
Run enough tests, and you will get statistically significant but bogus results.
Many times, when we declare a winner in test and if you keep running it, you may see that eventually it does not perform as good.
But I agree my comment doesn't negate Jabbles' point.