How Amazon's Mechanical Turkers Got Squeezed Inside the Machine
spectrum.ieee.org
spectrum.ieee.org
I stopped after I realized that buying a higher tier of tax software just to handle multiple income sources ate at least the first $15 of income (which is a significant time investment in Mechanical Turk).
In other words the income was so low that even paying the taxes on it ruined the whole venture. Even Bing's rewards have a higher rate of return. The $2/hour from this article entirely mirrors my experience (although my average might be a little lower).
Here you would basically be defining how much you'd be willing to pay per work item, regardless of how it's done; you wouldn't get the secret sauce, just the black box output.
It's not really a great idea, but in fairness I don't think it's much worse than what's out there either.
I wonder if there's been research into the effort in time or persons to perpetuate this fraud. Knowing Amazon, I'm certain they've analyzed this internally if not constantly at least a few times over the years. I wonder if there's a public attempt at the same analysis.
So even if you come up with a scheme that rewards WhiteHat Turks, the BlackHats will inevitably crack it unless you insure it's a moving target.
Side note, if anyone has a way to automate the above in a way better than Hunter.io, there's big money in that.
Turkers who work hard to get the reputation to do high paying jobs won't risk losing the reputation by cheating. So if you want good results you need to pay up to get those Turkers on your job.
It's this just schilling or...?
Yes, this is true, but this is true outside MTurk as well. If you rely on the YouGov panel (people complete surveys to get points for t-shirts), or really anything else, there's a strong incentive to cheat.
This is why good survey research is going to involve attempting to bound fraud through a variety of measures: attention checks ("What's your favorite color? Ignore this question and answer yellow."), looking for straight-lining (respondents always picking the leftmost answer), looking for unusual contradictions (i.e. asking the same question in two, opposite ways and looking for people who don't have the expected relation in their responses), looking at the distribution of completion speed and scrutinizing the lower quantiles, attempting to log participants who take the survey multiple times through dummy accounts, etc.
Oddly, my experience was that Turkers were fairly conscientious. I know one theory for this is that many Turkers are in fact Turking on the job, and so their reserve wages are not about what they are being paid for the MTurk hit, they're about what they're being paid to sit in their desk and not work.
I’ve noticed something similar to that in some of the YouGov surveys I’ve answered over the last decade. I wonder what it says about me that I sometimes write, in the errors/comments section at the end of the surveys, that I have seen surprising questions such as “what do you think about Channel 4?” when I have previously answered that I haven’t watched Channel 4 recently.
As you say though, it's the same problem with just about any study of this type whether MTurks, some other online panel, or students earning a little bit of beer money.
Does mTurk have some sort of quality-related metrics respectively higher rewards linked to it?
For example when I submit some work can I specify "I want only at least 4-stars-people (out of 5, where "5" is for people that rarely make mistakes) to work on this and yes I will pay 4$ extra fees for 4-stars and 8$ extra fees for 5-stars rated workers"?
Thx