Google is investigating the actions of another top AI ethicist
axios.com
axios.com
If true, that is damning, and would demonstrate once again that being an “ethics researcher” does not mean that you are are any more ethical than the average person. It just means you are more interested in the subject.
As a side note, I wish this field was more interested in meta-ethics than it is in forcing machines to abide by the personal ethics of the humans involved.
"Emancipatory research: Research that exposes underlying ideologies in order to liberate those oppressed by them." (Zina O'Leary textbook, used by many schools, over 2K citations for multiple editions)
So if a corporate behemoth such as Google wants to ethicswash itself by hiring individuals with the above approach, what outcome did it expect?
Let's take Timnit's paper for example. She found out that Google had been using a biased language model in search. The bias was like - 'male' association with 'doctor' as opposed to 'woman' association to 'nurse'. But she didn't show how this is harming anyone in a concrete way. Just theoretical. And then she offered no solution, just blaming the work of others, using her paper as a soap box to raise scandal and make herself holier than thou.
Hiring someone (perhaps with aspirations to buy their loyalty?) entails risks. Instead of pulling funding from the "independent think-tank", google now finds itself in the midst of a potential discrimination/political scandal involving an employee.
She routinely offered suggested methods, experiments, and even new datasets [1] that fixed what she saw as wrong. She does have a fair bit of practical ML experience as shown in the computer vision papers in her publication list.
1. Buolamwini, Joy, and Timnit Gebru. "Gender shades: Intersectional accuracy disparities in commercial gender classification." Conference on fairness, accountability and transparency. 2018.
> We developed the Pilot Parliaments Benchmark (PPB) to achieve better intersectional representation on the basis of genderand skin type. PPB consists of 1270 individuals from three African countries (Rwanda, Senegal, South Africa) and three European countries (Ice-land, Finland, Sweden) selected for gender parity in the national parliaments.
A dataset of 1270 images is hardly a breakthrough, the kind I expect to see in a small university project. But it doesn't lead to better models because it's not nearly large enough to train on. What it can do is rate existing models. Basically - useful to critique, not to improve.
A small nitpick: why just two races in a de-biasing dataset? Where are the Asians?
To roll with the example “man” ~ “doctor”, “woman” ~ “nurse”, the harm is having a giant and widely used search engine reinforce baseless gender biases, ie that there is no underlying reason why women should be nurses and men doctors. What is the harm you may ask? The harm may be subtle, eg being surprised when you find out your next doctor is a woman or your next is a man. It could suppress career choices and aspirations, and it could even be financial, eg reinforcing systemic pay gaps.
You’ve essentially created a fictional data set because it’s biased due to the underlying prejudice (preconceived opinion that is not based on reason or actual experience) that men ought to be nursing more, despite that not being reality.
We’re in a strange situation where we have large concerted efforts by activists to inject fiction in to our facts (whatever the medium) with the aim of distorting perceptions in such a way as to some how correct what they perceive to be injustice in the real world.
Why is the aim of exposing ideology not worthy of being researched in your opinion?
The belief that research has to involve activism has been demonstrated by the individuals under discussion. I cite that this is taught as a valid position in the field.
Political motivation of scientific activity is worthy of research. Both sides involved in the conflict at google may be described as demonstrating political, and apparently opposing, viewpoints.
A large part of them believe in other delusions such as "silence is violence", coming to work on time is "Whiteness", and "objective, rational" thinking is normalized racism.
In their deranged minds, the ends justify their means, so they can be the antithesis of their being because their crusade is holy, and just.
That's the complete opposite of being a researcher.
For example Andrew Ng is considered "white adjacent" because he's the only non-white in an article about the history of AI. Asians are not favored with this group of activists.
It seems that the occurrence of people who bemoan that social scientists are activists, that they are not supposed to actually develop solutions to what they study, has increased in recent years. It's bad logic. A sociologist studying the effects of poverty shouldn't be interested in solving poverty is a mind-boggling idea.
There's something more pressing, now. People can concur here. I've found that this line of thought, that social scientists shouldn't be activists, is an idea that's been drummed about by the so-called 'intellectual dark web'. This rag-time team of pundits propagated a lot of conspiracy theories about the 'cultural marxism', the 'Frankfurt school', etc. It feels like an attempt at policing the content of social research under the cover of conservative/christian propriety.
There is also a reason to be cautious if a sociologist studying the effects of poverty were basing their suggestions on what would fix their own poverty.
As for your second point, that's true, but it's also why research is always open to criticism. Casting social scientists and ethicists in this case as activists feels like a political swipe.
I'd prefer they label themselves as partizan ethicists or activists.
> , that they are not supposed to actually develop solutions to what they study
On the contrary, they should develop solutions, not just scandals. The problem is with activists who just want to criticize without contributing a solution. I suspect they are more interested in making a name for themselves and using ethics as a club.
Pointing out a serious problem is a good thing to do, even if you don't have a solution.
Ethics is an old field, and also one that's been applied for a long time. It changed too. For the better even -- since social Darwinism was seen as ethical to some extent at the start of the 20th century.
[1] https://www.nytimes.com/2016/06/26/opinion/sunday/artificial...
She's lost her job; if she's lucky that's the end of it. She can probably sue to get her job back or a compensation depending on whether a judge or jury rules that leaking the data served a higher purpose, but it's debatable.
The other activist was fired when she deployed political messaging code in production while hiding the whole process from her team and manager.
Do those strike as a people that will honestly present their story and would be good to work with? Ones that happily lie and fudge the truth to drive their agendas?
Because in my experience people who act like this, no matter what skin color they have, are corrosive and abusive to work with.
See US news for example.
Of course, if they weren’t trying to poison the well about an imminent revelation of some greater unethical behavior on their part, they probably wouldn’t have engaged in the obvious unethical behavior. So...
Edit: another comment here also claims that Google’s statement was made because Axios reached out to them regarding this story after Timnit Gebru’s tweet. So there you have it.
Yes, Google’s ethical responsibility in a current employer/employee relationship with Mitchell is different than Gebru’s ethical obligation to her former employer with whom she is already in a public, contentious battle.
And even if the ethical obligations were identical, Gebru’s violation toward Google wouldn’t excuse Google’s toward Mitchell.
What about Mitchell's violation towards Google?
If we accept Google's own claims, they have an automated indication which leads to suspicion of that and on ongoing investigation, not even something where they are prepared to claim an actual violation. i.e., exactly the circumstances where every half competent organization would decline comment (potentially citing “personnel matters” until they'd actually completed an investigation.)
What substantiates this conclusion? They can have an investigation ongoing, and share the cause for said investigation. In the statement they explicitly establish that this doesn't imply guilt of the account owner.
Why are you so triggered by this clarification?
My experience over a lifetime as a news consumer of seeing how organizations deal with media inquiries about personnel matters.
> Why are you so triggered by this clarification?
Grow up.
Have a good day, stay safe.
In the narrow ethical world of top down organizations, be they for profit corporations or the Army, it's of course a mortal sin, as is every other attempt to bring accountability to upper levels and redirect the flow of bullshit from top-down to bottom-up. I'll let you guess who decides what "ethics" stands for in this situation.
Doesn't apply in this case. Snowden acted for the country, not for his dear friend who was fired in a scandal. Their supporters on Twitter raised hell and even made a cancel list of twitter users (AI researchers) who dared oppose their comments (Anima A.).
it's not damning and not unethical, it's the definition of whistle-blowing.
Sure, and, if false, its also damning — in both cases, of Google management.
Either:
(1) They have received an early indication which may indicate either unauthorized or legally protected activity, and are publicly naming and blaming and specific employee before completing an investigation, which is merely grossly unprofessional and unethical though probably not actually illegal, or
(2) They are lying and libeling a current employee.
I’d bet the employee leaking documents did, and Google just responded to a request for comment.
It’s more the style of those players.
Why do you think I think that? I never said anything about who reached out to the press.
> I’d bet the employee leaking documents did
I bet they didn’t, because none of the articles include anything that the employee would have given them, only that they could not immediately be reached for comment. If that employee was the first source of information, then there would actually be some information from their perspective.
> and Google just responded to a request for comment.
The normal response (for a variety of good reasons, including legal and ethical ones) from a company on a matter like this when asked by the press when they haven’t completed an internal investigation would be to decline comment on personnel matters. Google’s behavior is grossly unethical here irrespective of who reached out to the media first.
They're not doing this if the press reaches out to them for comment with the specific details first. Given the bad faith actions of Gebru already (and apparently she's at least partially the source here from the other reply to your comment) it makes sense for Google to clarify with the response they gave (and in that response Google also did not mention the name of the employee).
This issue is way too heated for real productive discussion on internet forums. It's obviously tribal, flamewar bait, with a massive undercurrent of partisan politics (and perceived partisan politics on the side of the person you're disagreeing with).
I found what Gebru did earlier to be wrong, my prior is that the behavior here by the employee is also likely to be wrong. I've been generally disappointed in the level of discussion from the political fringes (both 'woke' and pretty much the entire GOP at this point). I find a lot of overlap between this Google AI ethicist community and critical race theory woke politics.
I suspect in the end when all the details are out Google will be in the right.
Without waiting there's no point for all the arguing that's going to go back and forth in these comments.
The press didn’t have the specific details until given them by Google, which is the “naming and blaming”.
The press had a report from another party of an apparently nonfunctional corporate email address. Describing the existence and basis of an internal investigation that had not been completed from that is grossly unprofessional; I’d be mildly surprised if it was done by a small outfit where media inquiries were regularly fielded by someone with no corporate PR, HR, or legal training, advice, or guidance, but for anything at Google’s scale it is unimaginable as anything but a conscious, deliberate breach of norms with the intent of harming the subject employee in public without developing a full picture of the facts, and that’s assuming Google’s statement is completely truthful as far as it goes.
To give a counter-argument to this: If party A makes a claim or accusation, and party B just says "no comment for now", it'll be near-impossible to reduce the (potentially unjustified) fallout of the premature verdict made. Public opinions matter, and stories told one-sided should not be desired by anyone interested in an objective discourse.
- Gebru tweeted the name of the employee [1]
- Axios then reached out to Google, who then made the following statement:
> Our security systems automatically lock an employee’s corporate account when they detect that the account is at risk of compromise due to credential problems or when an automated rule involving the handling of sensitive data has been triggered. In this instance, yesterday our systems detected that an account had exfiltrated thousands of files and shared them with multiple external accounts. We explained this to the employee earlier today.
Context matters.
[1] https://twitter.com/timnitGebru/status/1351698317550432256
No, its not.
> Gebru tweeted the name of the employee
Sure, more to the point Gebru tweeted that Mitchell’s corporate email appeared to be nonfunctional, sure.
> Axios then reached out to Google
That seems likely to be the sequence of events, sure.
Usually and ethically, a company that was in exactly the circumstances Google described would have:
(1) Declined comment, or
(2) Confirmed the email was nonfunctional and declined further comment, or
(3) Explicitly declined comment on personnel matters (especially if the framing of the question from Axios raised the issue of it being a disciplinary action of some kind; raising a personnel issue when it wasn’t part of the framing of the question would itself be somewhat unusual.)
> Context matters.
As an abstract truism, sure; while the narrative you describe is exactly what seems like the most likely scenario to me, I didn’t describe it because that fact was already considered in the description of scenario #1. Google’s behavior is (even assuming that they are being completely honest) grossly unethical in the context described.
If you're accused of a SECOND action taken against an ethicist, and AGAIN it's not because of anything you did that was bad, then yes your hand is actually being forced to say more than "no comment".
It's absolutely incredible how much Google has not spoken out to defend itself against the lies upon lies upon inconsistencies that Timnit has thrown out. And now another case pops up?
How about “the guy burning the person they are currently in a business relationship with without getting all the facts is wrong”. Or “a wrong by party A against party B does not justify a wrong by B against C.”
> It's absolutely incredible how much Google has not spoken out to defend itself against the lies upon lies upon inconsistencies that Timnit has thrown out. And now another case pops up?
Er, nothing in Google’s story, even taken as gospel truth, indicates that either Gebru’s fact claims in this case were lies or that her speculations were unreasonable or inconsitent in her position given the observable facts. So, your characterization seems...misplaced, at best, even if your characterization of her past actions was accurate. (And in the cases where Google has presented contrary stories to Gebru on other points, Google’s own stories have been outright self-contradictory whereas Gebru’s were at least internally consistent, so I can either trust Gebru or neither.)
I don't even know what you think are the actions people have been taking, to make you interpret things in a way to make you say that.
> Er, nothing in Google’s story, even taken as gospel truth, indicates that either Gebru’s fact claims in this case were lies or that her speculations were unreasonable or inconsitent in her position given the observable facts.
It's very damning that she chose to call out Jeff Dean as being the person who fired her.
Her manager was not a man. Her manager's manager (who actually delivered the news) is not a man. The CEO is not white. She chose probably the ONLY person in her entire reporting chain who happens to be a white man (and in engineering circles famous), and she points to him and says "He! He did this!".
And that's just a start.
Face it, you don't even have to read Google's official account, much less believe it, without seeing that her story absolutely does not add up.
Even the headline does not match the content in any article about her.
The researcher who tweeted out the name of an employee who's email has been blocked, and throwing out theories about crackdowns and firings - that's all good.
Explaining the email account has been blocked due to mass-leaking documents - that's beyond excusable?
Sure. When facts contradict your opinion, you shouldn't hate the facts.
That said, if Google is "investigating" her because she was trying to find evidence of discrimination or bad treatment of Dr. Gebru, that seems borderline criminal behavior. (Assuming her attempts at trying to find evidence did not involve exfiltrating secret information).
Hopefully, if it was just normal conversations and after identities had been redacted, their access is restored. Still, analysing this information on their work machine itself and then sending out a summary seems like a better course of action than this.
——
Separately, from the drama I’ve seen around the recent AI ethics stuff, I’d bet money Google is on the right side of this and the activists are not.
Why is your username 'foss user' when you're showing you are for intellectual property rights by calling this person a crook?
This seems like a contradictory criticism to make. The Less Wrong folk, in my perception, mostly dabble in thought experiments (such as Roko’s basilisk) that aren’t terribly relevant to real world AI work. Timnit’s work on identifying racial and gendered inaccuracies in facial recognition [1] seems much more actionable.
[1] https://www.technologyreview.com/2020/06/12/1003482/amazon-s...
https://news.mit.edu/2018/study-finds-gender-skin-type-bias-...
Miri is mostly focused on the AGI control problem and thinks that AGI is closer than others believe. If true, all other problems are pretty irrelevant.
There is real work in ML bias and fairness to be done too, but it seems overrun by toxic personalities and partisan politics.
That’s fair, I haven’t visited LW in a while. I did thoroughly enjoy HPMOR, but thought that most of Eliezer’s other work fell closer to speculation than practice.
As other commenters have mentioned, these AI algorithms are already being implemented in the real world with real consequences. Their associated ethical concerns therefore seem more urgent, and they can be acted upon.
[1] https://www.technologyreview.com/2020/12/04/1013294/google-a...
> “It is past time for researchers to prioritize energy efficiency and cost to reduce negative environmental impact and inequitable access to resources,”
What? She should read previous work on the topic first - quantization, sparsification, distillation, fine-tuning pre-trained models, etc. I couldn't find her actionable ideas because she doesn't have any original ones. It's a hard topic which requires a complete rethinking of both algorithm and hardware. And she doesn't recognize that Google is investing in both to do that - TPUs on the hardware side, and improved algorithms such as the Reformer - a variant of the Transformer, reducing complexity from O(n^2) to O(n).
Then regarding dataset bias - I couldn't find their actionable ideas. I mean, except telling people to be careful about how they select training data, but nothing about how to replace the model they criticize with an unbiased model that works just as well in the general case.
It's easy to point out problems while not providing any solutions and making yourself the critic of those who are doing the hard work. See how she treated Yann LeCun on Twitter - she basically told him to go educate himself on her papers and fuck off because she doesn't have time for debating him, that after denouncing his tweets.
Also, if you're into the LessWrong mindset and are susceptible to think that that which can be destroyed by the truth should be, you might find this tidbit interesting:
> Buried in the recent trillion parameter language model paper is how the dataset to train it was created. Any page that contained one of these words was excluded: https://github.com/LDNOOBW/List-of-Dirty-Naughty-Obscene-and... Two sample banned words: "twink" and "sex" https://twitter.com/willie_agnew/status/1350551463718621184
For example, there's also prevention of mass manipulation (AI learns human psychology, $BAD_ACTOR uses that to manipulates masses) and AI based profile generation of individuals with the goal of creating leverage for blackmail/extortion/whatever. Or, how would one form a controlling body of unethical usage in AI in such areas. There are quite a few aspects never mentioned in these discussions, and a search via google scholar turns up lots of the sex/gender/race papers, but none of the others (at least on the first few pages)-
There's a sense of starting small - redlining is illegal. No company wants to redline - it cuts into their profits and dings their compliance scores. So they're willing to work with us. Once we get that right, then, maybe we can start dealing with the cases where there aren't millions of dollars of funding assisting with making the tech more ethical -- and then, cases where it may actually be working against the money.
They have their proprietary algorithm to calculate a credit score, for example you almost cannot find an apartment without presenting your SCHUFA review and there is now way to know how their score is calculated (it has even been shown in the past that they use incorrect or outdated data for their calculation).
That’s the definition of a black box in my book.
Is there an analysis comparing the automated to manual approaches? I mean, people are biased, too. The models are more systematically biased, but can also be more systematically evaluated and de-biased. Which is better, the old or the imperfect new?
0: https://arxiv.org/abs/1809.07842 2: https://arxiv.org/abs/2004.07213 also has an overview of our current combative systems.
I mean, come on. There’s no way someone in a leadership position would legitimately think that’s okay. I suppose to some degree we should reserve judgment until hearing her side of the story - maybe this was just unambiguous evidence destined for her lawyer and the EEOC - but it’s hard to avoid some pretty negative conclusions about the prevailing culture in this group.
I wonder if this one collected all those internal threads where Timnit was being toxic, that insiders kept describing.
No? I guess this ethicist thought it best to only exfiltrate things matching her narrative?
I have to assume not, otherwise everybody using this feature would trigger the same exfiltration-lockout.
Legitimately curious here. "Download all my mail from the mail server" is something I do multiple times every day. I guess my workflow would be verboten in corporate environments? Is it really such an odd behavior that they have an automated check for this, and it isn't throwing false positives every hour? Or is this one of those things they tell you at the "onboarding" like "hey, you might download all your personal gmail via IMAP regularly, but don't do that with your corporate email account or you'll get a phone call from security".
Maybe I'm just out of touch with how corporate email works. It's been a decade since I was an employee of a publicly traded company.
It's not like the IMAP protocol sends the host MAC address or /etc/machine-id as part of the command stream.
I'm really having a hard time imagining a plausible automated security system that would have led to this outcome.
* Google automatically flags this by the volume of data transferred and alerts human eyes
* Google's alert erroneously fired and they made a press release about it
* Google is lying
* Your premise that it is a common and allowed use case is false
Why stop when your conclusion is absurd. These paths don't appear to take great creativity to identify.
According to a source, Mitchell had been using automated scripts to look through her messages to find examples showing discriminatory treatment of Gebru before her account was locked.
Google's own statement, quoted by Axios, says:
Yesterday our systems detected that an account had exfiltrated thousands of files and shared them with multiple external accounts.
You're taking those as if they say the researcher was just using her mail client the way everybody does, but that's not what they say at all. The first one ("using automated scripts to look through her messages") says nothing about downloading at all; the second, official one, says that "thousands of files" were "shared with multiple external accounts." The allegation is clearly that the researcher wrote scripts to select (an apparently large number of) specific messages and forwarded those messages to people outside the company.
Nobody is saying "OMG SHE USED IMAP!". They're saying "She sent internal corporate emails to people outside the corporation."