Someone wrote and tested this algorithm, and either:
a) didn't test it on pictures of women, or,
b) didn't notice that it cropped breasts rather than faces, or,
c) didn't think that was a problem.
If they had noticed and cared, this wouldn't be the approach in use.
Maybe they weren't thinking about boobies a whole lot, tried like 5 random test images and shipped it?
Presumably there are a ton of failure modes for that algorithm, why get so moralistic and high-horse about just one?
Not doing so; just observing the facts.
If a proposed 'unsupervised' algorithm of this simplicity highlighted women's faces perfectly, but zoomed in on men's receding hairlines, it wouldn't have made it past the drawing board. Indeed, it's reasonable to believe that nobody would have noticed that it consistently worked for women. We certainly wouldn't know that this algorithm existed or be talking about it here.
We observe a bias in what is considered important to check before shipping.
The simplest case is just to pick the most central square, and then you could probably improve that by picking a standard square according the rule of thirds. Those are the naive algorithms - this choice of alternative algorithm is deliberate and isn't as naive or simplistic as you're claiming.
The algorithm is only considered useful because it appears to do better than that, on whatever examples that the developers tried (i.e. there was a business case for using it), and against other possible code.
Including, likely, pictures of their own selves. That's certainly what I would test it on, until it vaguely worked.
What 'awful lot of assumptions' do you think I am making? I don't imagine we are in disagreement about this.
Years ago reddit was, IIRC, not very staffed at all compared to their traffic. It's a pretty privileged take to say they should have done expensive QA entirely around your particular things that you care about.
If you don’t have the capacity to use new technologies without increasing harm, maybe you don’t have the capacity to use them.
And it was a naive image cropping algorithm, years ago, and not making use of any sophisticated 'new technologies'. The beauty of the algorithm is that it was a simple function that could have been written in 1975 and required no training, deep learning or any of that. If you want to talk self-driving cars, you've got a much more relevant measure of harm and I'm right there with you.
As it is, I'd say there's a disconnect between where you and those years-ago shoestring developers stand on Maslow's hierarchy of needs. They were being scrappy with limited resources, and you're mad at them for not having an amount of QA that would have seemed unbelievable to them under their resource constraints.
Right. I’m saying that not caring is the problem. We mostly agree on the facts. I’m objecting to not caring as an acceptable position.
We’re not talking about shoestring developers. We’re talking about platforms that millions of people including heads of state use to affect billions of people’s lives.
Jedberg was >5 years ago, maybe closer to 10 for this algorithm. They were relatively shoestring compared to now, and especially compared to current-day FAANG.
I’m not objecting to an algorithm that exposes me to more cleavage. I couldn’t care less. I’m objecting to an algorithm that exposes millions of men to a subtle but routine equation of men with faces, and women with breasts. It’s not a privilege for women to expect not to be casually sexualized because someone thought their algorithm was helpful but didn’t bother to find out it wasn’t.
Likewise the impact of racial bias. I’m not offended by seeing white faces. But it’s not a privileged position for POC to object to being basically erased from images because whoever developed the algorithm didn’t bother acknowledging more than one skin tone exists.
These things have real impact, not just on how it affects the people who are directly underrepresented. It also affects how the people who are overrepresented perceive and ultimately act toward them.
If we are aware of a negative principle that infests and invades every part of life, though we do our utmost to repel it and mitigate it and remediate it, how can we maintain the conviction that it is not the natural order of things, but is a product of solvable human faults?
I'm not denying the importance of anti-racism or of examples of bias as mentioned here; my point is that when faced with an all pervasive adversary, whether a personified devil or something more abstract, human minds tend to find relief in submitting to the seemingly inevitable via some rationalization.
Clearly there is human-derived input in the system (otherwise... What's the point just crop randomly)
https://github.com/reddit-archive/reddit/blob/753b17407e9a9d...
But in short, it's a histogram of the values of the pixels.
Thanks for the insights!