Although their status page was completely green, for the last two weeks we've been hitting weird errors and disconnections when accessing their London cluster.
Since this Monday, the error rate has grown greatly and we found that, not only we had problem sometimes when writing to the object storage ( something we could try to mitigate retrying the upload ) but also some files that where confirmed as uploaded by the API, weren't later available to download and gave 404s.
After a lengthy call with their support this morning they've retroactively updated their status page, but the only response I have to the issue is: "Engineers have increased the storage capacity in the region; however, it may take a few days for mitigation to take effect." A few days!
The status update is hilarious:
Further additional storage capacity will continue to be added in a controlled manner balancing object replication and deletion process to further restore the environment's health. Engineers continue to recommend that customers retry upload and delete actions as needed as this is the only workaround currently available.
Did they forget that their storage was filling up?I know that not even Rackspace knows what the hell that company is about, but letting their storage fill up is beyond ridiculous.
The impact in our business has been, lets say, noticeable. And of course, we're migrating away as fast as we can transfer everything we have there.