If you use one of these tools it's completely safe to run tests until the test says win.
There are still caveats to account for the real world - e.g., only stop the test after an integer number of weeks - but statistically this is a solved problem.
[1] The blog post under discussion was written before Optimizely's StatsEngine and VWO's SmartStats.
Standard statistical tests used in a/b testing are based on one check. If someone is checking repeatedly on a test until they get a 'significant' result, your chance of getting a getting a false positive is many X the stated significance.
Best practice - set a pre-defined end, and one or two defined early check-in points where only make an early call if result is overwhelmingly significant or if the business has fallen off a cliff.
We try to run as many experiments as our traffic can handle, but we always estimate the sample size upfront when planning the test.