Errors for the Twilio Rest API impacting multiple services
status.twilio.com
status.twilio.com
I say that jokingly, but wow does it suck. I've been on both sides. I think it's actually worse to be not laid off and left picking up the pieces for incompetent management. One time I had to get admin rights and log into a "shadow IT" developers machine (who had been laid off) to get his Eclipse workspace directory and get the code and decipher what the state of his 5 apps were. Good times.
Deserved to be fired for that one but man did it suck to deal with.
I worked at two orgs, and did worse or better.
One had a crappy backup policy, I replicated the prod database to my workstation and took snapshots twice a day. That saved us TWICE!
The other time and I believe I still have an account running in their systems, they wanted to force devs to use Python 2.4 because that is what was in the base system included. I built my own Python plus deps and rsynced it to all the prod machines.
When I left and they disabled my account they enabled it a couple hours later. It was reportedly still running years later.
Sometimes getting laid off from a company that’s floundering (with personal finances in order enough to be able to handle it) can be a blessing.
Towards the end you start to wonder who the lucky ones really are. Not kidding.
Eng 2: Bob left in December. Who replaced Bob?
Eng 1: I don't know, his boss' email is coming back as invalid in outlook.
Eng 2: Fk it I'll just do it myself.
1 hour later
Eng 2: So backups stopped when Bob left in December. He was doing them manually due to an error in the automation. I'll try to restore from December.
5 hours later
Eng 2: backups are corrupt and won't work. Looks like we've never restore from backup to production
--------
Nation State Actor: WTB: Twilio access Recently Laid off Employee: I like money.
Do you often restore backups to production for testing purposes?
“According to Twilio’s latest earnings release, the company had 8,992 employees as of September 30, 2022 and expected to lay off 816 employees for the 2022 round of layoffs. Based on these figures, around 1,400 people will be impacted by this year’s layoffs.”
So that’s 2216 total out of 8992, 24.6%
https://techcrunch.com/2023/02/13/twilio-cuts-17-of-its-work...
The first layoff left you with 89%. Of that 89%, you lost another 83%. 0.89 * 0.83 = 0.7387.
1 - 0.7387 is how much you lost = 0.2613 = 26.13%.
It's probably clearer to reason about if you have two consecutive 50% layoffs. Clearly you didn't have a 100% layoff. Your workforce got cut in half twice, leaving you with 25% of what you had originally (or a 75% layoff).
Why else would BMW spend time and effort adding yellow lights to each corner of the vehicle?
He got laid off two weeks before we were set to do a huge annual update. I stormed into the interim CTO's office (the CTO got laid off too) and explained that laying this guy off was the dumbest idea ever and that it had put the entire project in jeopardy and no one had the knowledge or access to systems that this guy had, and that laying him off was going to cost the company millions of dollars because the release was going to fail.
I don't know what wound up happening at that release. Two weeks later I was working elsewhere; a few years later that company completely evaporated.
There are some unsupported methods a Google search away. I have had luck with this one in the past but haven't used in over a year, so can't vouch for it still working.
https://gist.github.com/gboudreau/94bb0c11a6209c82418d01a59d...
I backup my TOTP seeds in KeepassXC. I also have an offline backup of my Bitwarden vault that is in a separate KeePassXC vault. I agree that I don't like the idea of one vault holding both bits of info.
Exporting the configuration was a bit tricky, but I found a guide on GitHub: https://gist.github.com/gboudreau/94bb0c11a6209c82418d01a59d...
I use 1Password - its UI leaves some things to be desired, and it's not cheap, but it has zero incentive to cancel the account of any paying customer!
They did fix it with another update, but that was a seriously un-fun few days. Luckily I was just logged in to my AWS account so I could disable the 2FA.
Edit/Update: the title of the submission got changed.
This post is nothing of the sort, seems unlikely to provoke thoughtful discussion.
Edit: title has been updated since I wrote this comment.
Here’s a related HN post describing a hiring effort during the layoffs https://news.ycombinator.com/item?id=34804077
“For example, Twilio 2 days ago announced they were laying off 1000+ people, and here they are opening a staff software engineering role a day later.”
That’s just more evidence of significant incompetence at the leadership level that needs to be weighed when considering a service offering.
It seems we are at the stage of layoffs were we can argue about how you measure the percentage sacked..
That's 26.1%
Anecdata, but in several years of being a Twilio customer, I have never seen such a long outage.
EDIT: fixed % of %.
Unless the 11% and 17% are both talking about the original number of employees, the percentages don't simply add up like that. If it was 11% of the original number of staff then 17% of the leftovers, it'd be 26%:
.11+.17*.89=0.26
Still, that's not a huge difference so I don't know why people are getting up in arms over it.
Often times you have 2-3 ppl out of 6-8 that carry the cart while 3-4 are juniors or diversity hires.
Firing those 3 will have devastating consequences, firing the other 4 wont make a dent.
Firing based on excel is the most stupid thing you can do as a manager.
Vs placing an adamant bet on paying SF rent, SF salaries and urging others to do the same? https://www.sfgate.com/local/article/Twilio-CEO-wants-tech-c...
iirc the context of the article you're referring to is more of Tesla, Oracle moving out of California
(I was thinking the same though.)
Regardless, though, it seems to me a bit disingenuous to suggest that an outage now is correlated with a layoff from 5 months ago, let alone causal. Even with the more-recent layoff, Twilio is still well above its 2020 employee count, which should be more than sufficient to keep things running (and then some).
(Disclosure: former long-tenured Twilio employee; resigned a year ago.)
That being said, a 5 month delay is not at all surprising in the time between critical people were laid off and major issues arose.
Most systems are designed to run on their own (developers like to be able to go on holiday). The real problems arise when those systems are changed, and/or upgraded.
Regardless, I'm glad the title has been changed, as it was just unnecessary editorializing.
Any engineer at the company can add a small extra feature here and there to a system where the expert is gone.
But the issue is when that expert was maintaining a vision over their part of the platform, and keeping all of the smaller changes in-line with proper architecture for that system (knowing why we don’t do x, etc). The expert is able to push back against features that may cause problems, or suggest better ways to solve the goal. Laying off that expert will mean the things that they were protecting in the past are no longer protected. So the newbies comes in and add some change that creates an n+1 query, which might not be a big deal. But then they change some other functionality later that makes it n^2+1. And by not understanding the system, their changes compound over time to bring major issues to the system.
But small changes are fine. That’s why it takes so long for quality to suffer when you remove the experts.
And you don’t always know who the experts are either, especially as company leadership. It’s not always the person with a big title. Managers may know better, depending on how technical they are.
Harder when so many are let go though.
And then the entire tech world promptly ignored that lesson the very next year.
Building a product that falls over within hours of reducing the head count doesnt suggest they did a very good job, but they were keeping a fragile system up all this time ...
It’s not really worth speculating about fired employees performance, but there’s probably internal chaos after yet another layoff round.
That's an interesting metric worth considering.
If you're solely judging that based off complete downtime, I guess, but I think it's misleading to claim there's been no impact from those layoffs that affected service availability.
Robust? No. I mean, if the only thing you look at is your feed, maybe? But that's not Twitter.