It's not nearly as speedy as Google Translate, but I'll take that happily if it means keeping it local.
It's not nearly as speedy as Google Translate, but I'll take that happily if it means keeping it local.
It sounds like after Bergamot funding ended there were some communication issues and the Translate Locally group that was working with the Firefox Translate group stopped working together and now have their own extension, as mentioned in another comment:
https://news.ycombinator.com/item?id=33795141
That one can be installed outside the browser and would likely give better performance, although I haven't tried that yet.
The other interesting open source translation software I've seen is Apertium:
I don't think they have a browser extension unfortunately but it is an entirely rule based translation rather than AI models. I haven't tried this one yet either but hope to soon. I did try their web interface a few times:
Then this came along. All those nits are gone. Personally I find the translations easier to understand than Google's. (When you're overseas easier to understand trumps grammaticality perfect every time.) I'm not a fan of the banner at the top - the could move it to a tool bar icon like Ublock Origin does, but apart from that - it's damned good.
Now we need a replacement for Google Lens. For all it's flaws, Lens seems near magical to me.
I also think that the domain and the type of language used on Wikipedia is pretty consistent which will help a lot with unseen sentences.
By no means are these models bad! It’s just that Wikipedia is a particularly easy test for them.
Both of these methods have a bootstrapping problem, but at this point in the MT for many languages we have enough data to get started. Previous iterations of ParaCrawl used things like document structure and overlap of named entities among sentences to identify matching pairs. But this is much less robust. I don't know how they solve this problem today for low-resource languages.