Of course, that's the whole point of declaring statistical significance. With 95%, you will have false positives. With 99%, you will have false positives. There are no guarantees.
Many times, when we declare a winner in test and if you keep running it, you may see that eventually it does not perform as good.
But I agree my comment doesn't negate Jabbles' point.