"Although translations are managed automatically, it is possible for users to suggest alternative translations manually.
However, the BBC understands that this was not how the errors were introduced."you can only say something isn't automatic if someone with authority over the system executed the manual step.
Edit: found related link https://www.reddit.com/r/pics/comments/cygfx/4chan_is_using_... https://i.imgur.com/oCa4d.png
From the ideological side, as long as Google profits from the work of users due to their monopoly, but doesn’t provide it as open dataset under a noncommercial license, I don’t wish to support reCAPTCHA.
I used to type aa before noticing that and this also worked like a charm.
[0] https://www.reddit.com/r/pics/comments/cygfx/4chan_is_using_...
Helping a company that uses such a business model is not in my interest, as it means less volunteer power is available for actually open projects.
If Google wants actually meaningful manual OCR from me, they can have it – under GPL license, of course.
Things like "raining cats and dogs" doesn't probably translate well literally, for instance. It probably failed to categorise these universities as things that should not be translated in this manner.
Danish "... 15 meters ... 100kr" becomes "... 15 feet ... $100" which isn't helpful. (45 feet would not be helpful either.)
As it makes use of parallel corpora, there were examples of documents mentioning (say) a famous person in the United States, but when the document was manually translated into French, the human translator selected a different famous person more well-known in France. It made perfect sense in the context of the document, but it threw a wrench into the automated translation based on those documents.
http://itre.cis.upenn.edu/~myl/languagelog/archives/005485.h...