Firefox is getting language translation
zdnet.com
zdnet.com
In my experience Deepl is consistently, without fail, considerably better than Google Translate. I basically use Google now only for more exotic language pairs, or full-page translations.
I've found that sometimes what Google does for these is translate from Language A > English, and then English > Language B, which leads to bizarre results.
So yes, there's an intermediary language. But it's not English.
https://www.newscientist.com/article/2114748-google-translat...
For example, translating "рубанок" ("plane", in the sense of the carpenter's tool) from Russian to Polish used to produce "samolot" ("airplane") in Google translate up until sometime earlier this year, because in the intermediate representation "plane" was ambiguous just like it is in English. It looks like that particular bit is fixed now, which is at least progress! Maybe they've been adding more non-English text pairs...
And I gave the Hindi input in devanagari (राम आ गया), but still the Nepali translation ended up being the equivalent of 'the sheep came' (भेडा आयो), so somewhere along the line it seemed to be treating the name राम (rām) as equivalent to the English string 'ram' and translating accordingly.
So if the intermediate language isn't English, it certainly has some English-like properties....
Considering Translate is such an important product, I can't fathom why they just don't hire a single linguist (or just anyone who isn't completely clueless, really) per language to register decent translations, or at least import them from a real dictionary...
How many people on Google Translate actually are able to reasonably verify translations? Relatively few - and those qualified few who might poke at Translate out of curiosity are just as likely not to feel inclined to offer free labour to Google.
> Considering Translate is such an important product, I can't fathom why they just don't hire a single linguist (or just anyone who isn't completely clueless, really) per language to register decent translations, or at least import them from a real dictionary...
Perhaps professional translators rather than linguists. I imagine they have some linguists on the project, but they're likely to be more NLP-type linguists.
The difficulty is that they're interested not just in word-level meaning/translation-accuracy, but also phrase- and clause-level accuracy, and those are really large (i.e. theoretically infinite) spaces.
I've heard that the Spanish<->English translations aren't too bad.
One example is technical/service documentation for a heavy machinery company I worked for. The tech writers were based in Germany, but O&M Manuals were required in Japanese for sale in Japan. Those docs were translated using English as a "pivot language". Usually dictated by pricing (fewer German+Japanese translators = much higher cost).
And in any case, for German->Japanese, going through English probably has a lower cost. But for Hindi->Nepali, you'll lose a lot of information as Hindi and Nepali are closely related and similar not only in terms of correspondences between vocabulary items, but also grammatical structures, which is effectively 'thrown away' if there's an English, or close-enough-to-English-to-effectively-be-English intermediate translation language. (Not to mention the inefficiencies of the equivalent of sending a package from Delhi to Kathmandu via London.....)
And recently everything I had to look up while reading Gabriel García Márquez, DeepL didn't know. After enough failures I gave up and returned to Google Translate for the remainder of the book.
Moreover, Google often provides a single target word; whereas DeepL allows you to select from a range of synonyms clicking on a word, and will adjust the sentence accordingly to use the new word. When Google gets the context wrong and provides the wrong meaning for the translation, DeepL's capability to translate with a different meaning is invaluable.
For example it will always translate 'order' in a sentence as 'Befehl' (I order you to fix a steak) instead of 'Bestellung' (I order a steak).
Both are correct, but completely different, and in our context we never mean the former.
The Teacher said that it was amazing and that many native students she had couldn't write that well.
The bigger problem is likely to be lack of training data. Unless they have a pile of cash to pay professional translators to produce a parallel corpus, the alternative is to scrape translations from the internet. Basically, crawl the same site multiple times with different Accept-Language headers and try to align the results. Crucially, this depends on an existing ecosystem of bilingual websites with high-quality human translations.
According to DeepL's website they're a spin-off of Linguee, who provide a search service for exactly that kind of parallel data. So before DeepL starts supporting any given language pair, you should expect it to appear in Linguee first. https://en.wikipedia.org/wiki/Linguee
Edit: It took me a while to figure out how to select a different language on https://linguee.com (Their UI seems broken using mobile Firefox Preview.) Appending /english-chinese and /english-japanese to the URL shows that they already support those two, and the alignment of translations appears reasonable to me. No /english-arabic, though.
Google has clearly made it their mission to solve that problem, and I'd say they've been rather successful.
only has 9 language options(8 if you exclude the dest language).
But once you step away from it, quality goes down. I tried translating random pieces of Russian literature and it makes obvious mistakes. It can't even manage the structure of sentences, never mind word choice.
Translations from English are also bad. For example, it translated "I never felt that the translation is off" as "I never felt that the translation is turned off".
Specifically, the page is marked up with such tags as these:
<link rel="alternate" hreflang="en-GB" href="https://www.mozilla.org/en-GB/" title="English (British)">
I would prefer the browser to first suggest “do you want to switch to the English (British) version of this page?”Sure, I may actually want a machine translation of the text I see on page—the content may be different on the other language’s page; but if the page indicates alternatives are available, I’d prefer to try that first.
<link rel='prev' href='..." title="Previous page (about X)">
<link rel='next' href='...' title='Next page (about Y)' />
<link rel="copyright" href="..." title="Our imprint and contact infrormation">
<link rel="search" type="application/opensearchdescription+xml" href="..." title="Our Website title">
<link rel="alternate" type="application/rss+xml" href="..." title="News from our website" />
<link rel="alternate" href="..." hreflang="de" title="German translation of XYZ">
From these six examples, modern desktop Firefox/chrome only displays adding the search engine (when clicking in the search/omni box). RSS died a long death (the availability was removed in Firefox a few months ago). And all other information was only ever displayed by the good old Opera browser, who had a special symbol bar for relational data and even overloaded the prev/next browser navigation buttons with prev/next header suggestions from the page.As an aside, I don't know who needs to know this, but we are really lucky right now in that we now have tools that enable pretty much anyone to train a language model and translate with it, not just giant or small companies. It's not necessarily going to be any good, but the tools are all right there for us plebes and it's pretty fun.
I pulled together a tutorial just this weekend walking through doing this on Google Colab (https://sevenminuteserver.com/post/2019-10-17-machine-transl...) and EC2 Spot instances (https://sevenminuteserver.com/post/2019-10-17-machine-transl...) if anyone wants to play around with it.
It is better.
Google is now just a sluggish tech bully and only shows world-beating performance through acquisitions (such as, sadly, their acquisition of DeepMind).
The Google of 1999 - that a lot of us thrilled to - is long dead.
It says "[t]ranslate from any language", but doesn't seem to support Chinese or Japanese at all.
To me, translating other Western languages to English is never a problem. I can use any half-decent online translation services and the result is totally understandable.
But in the case of deepl, the result is not only understandable but also often (but not always) nearly perfect.
AFAIK, only the devil can translate from any language (not even Google Translate can translate from any language, in spite of attributed evilness to Google).
According to this article, they had 22 employees at launch and back then it was already lauded as quite good: https://www.gruenderszene.de/allgemein/deepl-maschineller-ue...
According to the Data page [1], it at least makes use of ParaCrawl multilingual corpus [2].
It is amazing that they work so closely together with Firefox, so the project result will be really of a use for not only European citizens but people on all the world.
There will always be a performance gap between multi-gigabyte server ML models and client side ones though - and it's up to the users how they prefer the privacy Vs performance tradeoff.
So maybe link to the actual source instead of the blog spam?
[1] https://web.archive.org/web/20190520103652/https://careers.m...
I really really hope this is a coincidence vs. Mozilla having some sort of tiff with Google and us being in the crossfire.
I don't use chrome that much so maybe it's a configuration thing I never looked into, but the nagging bar asking me if I want to translate every single english page is very annoying. Not a native english speaker but confortable enough to never need the page translated.
I hope when this is included in a stable release, it'll convince some of them to make the shift.
https://addons.mozilla.org/en-US/firefox/addon/translate-who...
Google Meta-translator finds the meaning not the words.
I switched to Firefox a couple of weeks ago and ML translations would be very low in my list of priorities.
It is a good translation tool (and I've used several, Google, Bing) and sure, it's accurate and all but DeepL doesn't compare when it comes to conversational translation or, voice to text etc. If it's simple copy and pasting it does that sufficiently well.
I know that translation became a lot better (using deepl [0] personally, check it out if you haven't heard of it!) in recent years, but on-site-translation always seemed quite goofy to me.
I wonder if this caters more to a broader (elder) audience, making Firefox more attractive to less tech-affine folks? Seems like a critical feature for "your grantparents computer", doesn't it?
I'm sure it's a nice feature to lots of people, but I never want to see it. It's just another annoyance that I don't wish to deal with. Dealing with multiple language is still something browsers and websites struggle with. Yes, I know I have a Danish IP, and my language settings is British English, but just give me the original Swedish version of the content it's FINE.
Much more useful feature, for my use case, would be language detection of input fields, so I don't have to switch between dictionaries to get the right spellchecker.
Given these settings, how would you expect the browser to know you want to see the Swedish content? You have explicitly indicated that you want British English (if available).
I will click No, keep content language and tick don't ask me ever again and be super happy.
If you really want me to like you, put a tickbox in the settings so I can save this as default into my other browsers (that don't use cookies / don't save data)
Thank you !
For instance I never want the Danish version of Wikipedia, but I do want the Danish version of a Danish news site, or the Swedish version of a Swedish website. There's no real way to indicate those preferences to the browser, so it would be best if it never try to deal with language.
I'm in that situation and i pretty much use Chrome exclusively for on-site translation.