Uber's Employee Ratings Put Women at a Disadvantage, Suit Says
bloomberg.com
bloomberg.com
While pretty much all employment in the US is "at will" employment, "good faith" is assumed in the system....
Stack Ranking by default is a "bad faith" system, where the employer, will fire people, even though the employees were good performers (if you are in a team of all good performers, it is common place that stack ranking punishes good performers/employees as well).
It is very similar to the "decimation" system that the Roman Army practiced, when a particular unit or legion performed bad or cowardly in battle, one in ten was chosen (via short sticks), and culled to death by their companions. Very brutal, and can't see how it improves morale even if the unit itself is actually performing well. (which stack ranking is, constant decimation). This causes rampant politicking and favoritism to the point that systematic discrimination arises from it.
Decimation was a system for BUILDING cohesion. You don't flee battle because you might be decimated. You don't let your comrades flee battle because you might be decimated. It's for when you want to severely punish a collective group of people but can't afford to hurt them all.
Stack ranking is the opposite. You are actively competing against your colleagues. You are incentivized for your colleagues to fail.
Both have the "Culture of Fear" as a central point, and yes, stack ranking has its roots on the Roman Decimation system.
http://www.hrmagazine.co.uk/article-details/performance-impr...
But five minutes of thought will tell you that people will hate working on that environment and will quit or just not apply to work there.
It is perhaps a natural fear that a Stack Ranking system would lead to a "Rank and Yank", but the mere ordering of employees during a performance management process doesn't give me any particular concern, nor do I consider it a bad faith system.
In fact, it forces people to make a sometimes hard choice, which I believe has more benefits than drawbacks to an organization, else you run the risk of having a massive "undifferentiated middle" and no clear development plans because you have no feedback from rich discussion about why Jane is better than Joe; instead Jane and Joe are simply declared to be equal without any discussion about what Joe could do to improve or what Jane should continue doing, etc.
But I think it's fair to assume that in most minds stack ranking includes a fire quota, and that's what most commenters mean here.
1: https://www.vanityfair.com/news/business/2012/08/microsoft-l... 2: https://www.glassdoor.com/Reviews/Microsoft-stack-ranking-Re... 3: http://www.nytimes.com/2015/08/16/technology/inside-amazon-w...
I thought it was effective in forcing deeper conversations as part of the comparison process, but became unwieldy and unworkable when we passed about 50 engineers which was a long time ago, so I'm not sad to see it go.
With Romans, specifically, if cowardice was punished, there would be noone left. They were disciplined and had great quality equipment, but their morale was not the highest. For instance, before they met the Germans on the field, having heard of the military prowerss of their enemy, the Romans would spend their time crying and making their will, certain that they'd die in battle - Cesar has written about this.
The specific problem with stack ranking is the one-on-one ranking method. I'm absolutely convinced that I myself would succumb to all sorts of biases when asked to rate people like that: I'd tend to think the silent, introvert guy who stays late and has no fashion sense is smarter than the outgoing jock leaving at 5pm sharp to be with his family. I'd somehow convince myself that the guy I enjoy eating lunch with is more important than the Belgian guy who doesn't speak english well enough to socialise... The article makes a similar point: “My observation was whoever was the best orator got their people ranked higher.”
On aggregate, it's quite obvious that these biases could amount to gender-specific discrimination, even if not a single evaluator had gender on their minds.
[0] For some reasonable value of "never"
[1] https://hbr.org/2017/10/a-study-used-sensors-to-show-that-me...
When managers are faced with stack ranking employees, they will often decide based on factors not purely related to the productivity or value of the employee.
----
That, and as an employee in a stack-ranked organization, you know you need to be competitive. Not in terms of productivity but in terms of "in the eyes of your manager"-- Perception is more important than performance.
Studies have shown that men compete more aggressively than women, on average. So it's certainly plausible that men enjoy a gender-driven advantage that isn't purely due to gender-discrimination. Rather, it is the system that is biased.
This reminds me of that "racist algorithm" debacle.
[0] https://hbr.org/2017/10/a-study-used-sensors-to-show-that-me...
I can appreciate that. It's almost certainly true. Most people are biased.
The point of my comment wasn't argue differently, or to defend Uber or their managers, but to show how choosing a sexist system is the same as being sexist, in effect.
People have good reason to be doubtful of social sciences, and the replication crisis has only gotten worse now affecting biological science.
https://www.nature.com/news/over-half-of-psychology-studies-...
Studies have also shown that when women exhibit identical "aggressive" behavior as men, they are labeled more negatively than their male counterparts.
Most people who use ridesharing see it only as a cheaper, easier option, which it is.
Most of the people I know who use or drive either because they seem them as commodities and don't really care. They just pick the better priced or closest one.
It is a real trend, but doesn't represent 'most' users.
In fact... I have only been to one place where I could not Lyft, and it turned out I also couldn't Uber (very rural).
Problems like this are endemic to zero-sum schemes like stack ranking.
Edit: The parent post was edited.
Are other employee evaluation schemes free of appearance of sexism (or actual sexism), or is it harder to see?
What I would like to know is, suppose you need to make a system in a large organization in an attempt to lay off your 10% worst employees. It's pretty reasonable to expect sexist rankings, even made by non-sexists, and to also expect beliefs by individuals, wrongly held, that their personal low rank was the result of sexism. How do you determine whether an organization's ranking system is sexist, in which direction, and by how much?
[] https://blog.impraise.com/360-feedback/how-performance-revie...
Edit: By non-sexists, I mean, like, non-garbage people.
On a serious note, tech companies, large and small, have varying amounts of lawsuits and news coverage. I agree that stack ranking and other rating methods make illegal discrimination more likely. I do not agree with your assertion that unreported discrimination based off sex doesn’t occur.
No you can't, unless you have a method for dividing a project up into blocks of exactly equal difficulty. The person who takes more time to do genuinely difficult things is probably closing fewer JIRAs than the one quickly cranking through easy ones, which would score more highly on that metric, but which is the more valuable to your organisation?
Why would you ever want to evaluate employees by any metric not directly tied to their contribution to company success?
Usually from the point of view of the problematic person, everyone else is the asshole. At some point the evaluation has to boil down to a judgement call, rather than an objective measurement. This process is always fraught, for instance when the bosses are the assholes themselves or support them (ahem, Travis), but I don't love the idea of just not considering interpersonal factors when evaluating performance. Bad behavior shouldn't have to rise to the level of an HR complaints before consequences are faced, and good behavior should be rewarded.
What metric would you recommend?
This isn't, by the way, an original concept; the dangers of the inappropriate use of micro-level metrics that aren't tied to business outcomes in decision-making (and the individual employee level is about as micro as you can get) was one of the key problems W. Edwards Deming identified.
Stack ranking is nothing more than a more aggressive form of regular hiring promotion and advancement.
Moreover, it's just as likely that female managers prefer females.
There's nothing inherently sexist about 'stack ranking' - rather, possible sexism is derived during performance reviews - which are commonplace.
One where no one ever will complain? No, but so what?
But absolutely eliminating sexism isn't the bar most people are held to. Most people are held to the bar of taking reasonable steps to mitigate the issue and improve culture.
In the vast majority of cases, you can iterate your way to a fair system with a clear rubric that controls for unconscious biases and establishes a substantial trail of evidence to support each judgement of performance. Actually doing that, especially when giving criticism, tends to be a matter of managerial skill.
I'm not sure this is true. A system with a clear rubric that will control for unconscious bias will be accused of reducing people to pure numbers and not consider factors that cannot be quantified. It's also entirely possible that the rubric may control for unconscious bias on the part of the manager, but introduce a whole new set of biases against one group or another.
This isn't to say we should make the effort. I actually feel that your point about managerial skill is perhaps the most important. Middle managers, the ones to whom ranking and performance evaluation fall, are often promoted from lower positions and do not have the background or training to be good managers. I'm not sure how to solve this problem, but in most organizations I've worked at this only way up is to management. Good workers turn into bad managers.
In turn, giving up on building a less sexist system because people will always complain is like giving up on building a more secure infrastructure because people will always complain.
Unless you are extraordinarily lucky at both the humans building what you're building and the inputs to your system (constraints about language and tech stack and problem domain on the security side; constraints about hiring sources and existing culture/public reputation and leadership on the employment side), you're always going to have security flaws in your infrastructure, and you're always going to have sexism in your employment processes. And even if you are this lucky, you're not going to eliminate people poking at your system and suggesting places where it seems to fall short. The important question is what you do about it, how severe the problems are now, and how the severity changes over time.
Just like there are plenty of ways to write software that take it as a given that implementors will screw up and attackers will be clever, there are also plenty of ways to design a performance review system that takes it as a given that some fraction of humans in your organization will have biases (conscious and unconscious) and will treat their coworkers according to their biases.
Honestly, some of the mitigations are the same: the first thing I'd reach for is defense-in-depth and auditing. If your system is consistently giving different results (whether it's hiring, or retention, or promotions, or salary) to people in different demographic groups in a way that seems to disprove the null hypothesis that the demographic is irrelevant to job performance, look very hard at it, add additional reviews to make sure the results are consistent even when different people are involved, and keep measuring. There might be good reasons that the disparity is legitimate (just like there might be good reasons that one of your servers is regularly opening outbound TCP connections to random machines in China), just make sure you are actually noticing that something unexpected is going on and try to find an explanation for it.
Uber seems to be full of bad ideas, so maybe it's a good fit.
[1]: https://wiki.lspace.org/mediawiki/Gaspode
>is only still alive because the various diseases are too busy fighting each other to kill him
> Microsoft Corp. and Goldman Sachs Group Inc. have faced similar legal challenges; both firms and, more recently, Uber, have abandoned the practice.
I don't know. They've basically changed the world, made a 60 billion dollar company quickly, and become a brand synonymous with a service.
That's very special.
Just as I think they probably have used some slightly underhanded things on their way there, I significantly doubt anything they're doing is considerably more sketchy than other companies.
As the article states: 1/3 of F100 companies use stack ranking - so how can we possibly call them out for it. Moreover, that someone 'says they are sexist' does not make them sexist.