My IELTS score is 7.5, but my writing band is 6.0.
I write my thoughts and comments in Chinese first and then use AI to translate them. The entire article was also translated from my original Chinese manuscript.
My IELTS score is 7.5, but my writing band is 6.0.
I write my thoughts and comments in Chinese first and then use AI to translate them. The entire article was also translated from my original Chinese manuscript.
Thank you very much for the article, it was super interesting. The mystery in the story draws people in, and people surely won't mind a couple of grammatical mistakes. But you have to watch out: the use of AI makes it easy for people to suspect that the story might've been embellished. For the second part, it might be better to try translating it manually; the same goes for writing replies.
Thank you for sharing your story. It makes the world a better place.
Low-level English is normal and accepted.
On the job, I never speak above the level of a 13-year-old.
AI generated English is hated.
Consider using English, not software translation.
I've been seeing this take on HN a lot recently, but when it comes to translation current AI is far, far superior to what we had previously with Google Translate, etc.
If the substack was written in broken English there's no way it would even be appearing on the front page here, even less so if it was written in Chinese.
Of course, even this can be faked, sadly.
That's why translation is a job in the first place and you don't see publishers running whole books through Google translate. No one, least the authors, would accept that.
Contrast this with the faux polite, irritating tone of the AI, complete with fabrications and phrases the author didn't even intend to write.
Authenticity has value. AI speech is anything but authentic.
The author acknowledged they used AI to translate. Is the translation they decided to publish among the given tools they had available to them not by definition the most authentic and intentional piece that exists?
All of this aside, how do you think tools like Google Translate even work? Language isn't a lookup table with a 1:1 mapping. Even these other translation tools that are being suggested still incorporate AI. Should the author manually look up words in dictionaries and translate word by word, when dictionaries themselves are notoriously politicized and policed, too?
Maybe or, most likely this is the same for writing: there are people that think correct grammar and punctuation and no help on achieving this, means writing.
The core algorithm behind modern generative AI was developed specifically for translation, the task which arguably these chatbots are the most suited! It’s the task that they’re far the best at, both relative to older translation algorithms (which were also AI), and relative to their capabilities other tasks that they’re being put to. These LLMs are “just” text-to-text transformers! That’s where the name comes from!
“Stop using the best electric power tool, please use the outdated steam powered tool.” is what you’re saying right now.
You’re not even asking for something to be “hand crafted”, you’re just being a luddite.
Indeed! And yet, generative AI systems wire it up as a lossy compression / predictive text model, which discreetly confabulates what it doesn't understand. Why not use a transformer-based model architecture actually designed for translation? I'd much rather the model take a best-guess (which might be useful, or might be nonsense, but will at least be conspicuous nonsense) than substitute a different (less-obviously nonsense) meaning entirely.
Bonus: purpose-built translation models are much smaller, can tractably be run on a CPU, and (since they require less data) can be built from corpora whose authors consented to this use. There's no compelling reason to throw an LLM at the problem, introducing multiple ethical issues and generally pissing off your audience, for a worse result.
Because translation requires a thorough understanding of the source material, essentially up to the level of AGI or close to it. Long-range context matters, short-range context matters, idioms, short-hand, speaker identity, etc... all matters.
Current LLMs do great at this, the older translation algorithms based on "mere" deep learning and/or fancy heuristics fail spectacularly in the most trivial scenarios, except when translating between closely related languages, such as most (but not all) European ones. Dutch to English: Great! Chinese to English: Unusable!
I've been testing modern LLMs on various translation tasks, and they're amazing at it.[1] I've never had any issues with hallucinations or whatever. If anything, I've seen LLMs outperform human translators in several common scenarios!
Don't assume humans don't make mistakes, or that "organic mistakes" are somehow superior or preferred.
[1] If you can't read both the source and destination language, you can gain some confidence by doing multiple runs with multiple frontier models and then having them cross-check each other. Similarly, you can round-trip from a language you do understand, or round-trip back to the source language and have an LLM (not necessarily the same one!) do the checking for you.
> This can avoid the taste of AI, but it may be very bad to read, I first used machine translation translation, many parts become very wordy, and at the same time puzzling.
Perfectly clear and comprehensible. It's not fluent English, there are comma splices everywhere, and it translated "machine translation翻译" as "machine translation translation", but I understand it – and I'm confident it's close to what you actually meant to say. I can spot-check with my Chinese-to-English dictionary, and it seems like a slightly-better-than-literal translation. My understanding of your comment:
> This can avoid the smell of AI, but it may be a struggle to read. I initially used a dedicated machine translation system, but many parts became verbose (/ very wordy) and incomprehensible.
Generative models don't solve the 令人费解 problem: they just paper over it. If a machine translation is incomprehensible, that means the model did not understand what you were saying. Generative models are still transformer models: they're not going to magically have greater powers of comprehension than the dedicated translation model does. But they are trained and fine-tuned to pretend that they know what they're talking about. Is it better for information to be conspicuously lost in translation, or silently lost in translation?
Please, be willing to write in your native language, with your own words, and then provide us with either the original text, or a faithful translation of those words. Do you really want future historians to have to figure out which parts of this you wrote yourself, and which parts were invented by the AI model? I suspect that is not the reason you wrote this.