Because what appears to be missing is any sort of feedback mechanism by which to improve the accuracy of it's predictions on the same scale as which it is making its predictions.
In other words, if the prediction was rain in 8 minutes for 15 minutes, how can the program determine if it was accurate? (And suggests the question, is rain in 4 minutes for 18 minutes an acceptable level of inaccuracy?). Where does the data that it was 18 minutes of rain come from?