> there is no "si" in the hiragana table, so s_ + (i) = shi. […] this is why it's important that you don't actually "think in" romaji. […] i'm using romaji as a convenient way to refer to phonetics in text. however, your "mental algebra" should match the hiragana table.
If you fix all the errors that are in the article, at best there is an argument buried here that Hepburn romanization should not be used to teach Japanese to English speakers—but I think that point is really my own argument that I’m making with the fragments of the article that make sense.
Romanization can be more consistent with Japanese phonetics or it can be more consistent with English phonetics, and the Hepburn romanization is more consistent with English phonetics, which is why it’s a good choice for English speakers that don’t know Japanese, but a bad choice for English speakers who are trying to learn Japanese.
You may argue with my choice, or maybe you can argue that referring to cells in Hiragana table solely by my chosen romanization is somehow bad, and I should instead be inconsistent and give the same mora two different romanizations within a single article. Is that what you’re suggesting?
Second: am I arguing that the choice of using Hepburn here is somehow bad? Yes, that’s correct. I think Hepburn is a bad choice here. A good choice is Nihon-shiki. JSL romanization is also fine.
I think I agree that Nihon-shiki and explaining it upfront would’ve made the article more elegant. One constraint I wanted to hit is that a person should be able to read this article with zero knowledge of Japanese, and walk away with being able to conjugate almost every verb to every suffix correctly. This is more of a challenge to myself as a writer than any practical need but hope it shed some light on the choices and the framing. I liked Hepburn because it’s closer to how it sounds. You can imagine I’m using IPA instead if you want.
> hanas* + (i)masu = hanasimasu (wrong!)
I cannot wrap my head around how this line in the article could be defensible. Like, if I don’t understand how Japanese is pronounced or written, and I just rely on Hepburn, I guess pasting these fragments of Hepburn together don’t produce the right Hepburn in the end?
YMMV indeed, but I think the lesson here is “this is why you don’t use Hepburn when you’re writing an article about Japanese verb conjugations”.
Hepburn does make sense for somebody with zero knowledge of Japanese but it just gets in the way when you are trying to explain how Japanese works. So lesson zero is “don’t rely on Hepburn” and IMO if you are interested in pronunciation and listening you should be using audio as your primary source.
I sympathise with your point about the benefits of Nihon-shiki romanization here. It might’ve been a better choice for this article.
> I cannot wrap my head around how this line in the article could be defensible
I think the reader would just read the next section where I use your argument to critique my own approach? And then make up their own mind whether it’s defensible to do something in the article, to raise pros/cons for why I did it, and then to keep on with the choice.
I wanted to illustrate this confusing point, and that’s how I chose to illustrate it. I think it’s confusing either way. I trust that a reader who actually wants to learn, and isn’t just being a pedant, would carry away the right set of conclusions, and would understand the isomorphism (again — see EDIT below) after those two sections.
> Like, if I don’t understand how Japanese is pronounced or written, and I just rely on Hepburn, I guess pasting these fragments of Hepburn together don’t produce the right Hepburn in the end?
Yeah. So that’s a learning opportunity that kana row shifting doesn’t quite follow rules you might expect from many other languages. Maybe that’s a clunky way to introduce it. I personally like this framing. As I noted somewhere else, you could imagine that I’ve chosen IPA notation instead.
—
EDIT: Actually wait, Hepburn is not bijective for zu and ji. I haven’t thought about that. It’s not relevant to any of the conjugations so it doesn’t break the article, but that may be a good argument that it’s not worth the effort rescuing Hepburn.
I think that’s a long wait; I don’t want to rely too heavily on analogies but it is like teaching somebody arithmetic roman numerals and then explaining in a parenthetical that there are other ways to do arithmetic (but not naming them). Maybe the reader can make up their own mind—but I don’t think the pros and cons are raised in the article, or if the are raised, I couldn’t find it.
I don’t want to pile on here but it sounds like you are, in this conversation, learning about why the different romanizations exist and what the pros and cons are. Or if you already knew, you are getting what they call an object lesson. (Like you noted—in Hepburn, ji and zu correspond to two different kana each.)
> As I noted somewhere else, you could imagine that I’ve chosen IPA notation instead.
This just resurfaces a similar problem with different symbols—if you put your IPA notation in slashes // you get phonemes, which will get you something mostly equivalent to Kunrei-shiki romanization. If you put your IPA in brackets [] then you get something sort of equivalent to Hepburn (in that it’s designed to show pronunciation). Both choices will on some level obscure a regular pattern that could be revealed with kana or romaji. Orthography is funny like that; in both Japanese and English it can show the origin of words even when the pronunciation changes.
I think the other lesson here is that students will mostly learn morphophonology intuitively by absorbing examples with some light explanations of the rules, and if you overexplain the rules you end up with too much “scaffolding” which gets in the way. Like when people use mnemonics or try to memorize kanji by thinking pictorially.
In general, I find your attitude a bit condescending. This is what I wrote about my choice:
> note i could also have used a different romanization that renders し as "si", つ as "tu", and ち as "ti" for this article. i decided to not because everyone else uses romaji, and once you understand this point once, you shouldn't have a difficulty doing this in your head
My main mistake seems to be meaning “[Hepburn] romaji” by writing “romaji”. I was obviously aware of other systems because that is what the sentence says but I thought it’s acceptable to refer to Hepburn as just “romaji” as a sort of the default one. Maybe that’s wrong.
Other than this terminology nit, I think I’ve made myself quite clear there. I genuinely don’t think it’s a big deal. Maybe I overestimate my readers’ intelligence but I don’t find this difficult to live with at all once you get it.
Roman numerals is a funny parallel but it doesn’t hold very well. The difficulty of using Hepburn is O(1) shortcut: for conjugation, you only have to “remember” three special cases and they’re always applied just-in-time. It’s just substitutions — and are arguably inherent phonetically. Arithmetic with Roman numerals requires many stacked adjustments where you have to match pairs of things. And lack of orders really screws with ability to do multiplication. This just isn’t an intellectually honest comparison.
Re: your last point I actually kind of agree. I’m that annoying student who likes to un-extrapolate backwards from examples to the rules, knowing which gives me a warm fuzzy feeling, after which I can go back to examples. My article is for people like me. Maybe there’s a few more of them.
Yeah—I can understand why I’d come across as condescending. There’s a balance here—I want to be clear when I say that I have problems with the article, but I don’t want to be hurtful and I don’t want to make criticisms that are not supported by the text.
Rather than defend my comments as “correct” let’s say that I failed in my goals of not coming across as condescending. The reason I want to frame it this way is that similarly, I think the article failed in its goals as coming across (to me) as “look at this neat thing about Japanese”.
It is just kind of the nature of written communication that it takes a lot of editing and polish to make it clear, correct, and concise. I had the good fortune to sign up for Japanese 101 when my professor was in the middle of writing a new Japanese textbook—it was pretty exciting, with the changing lesson plans, the flock of master’s students hanging around, revisions and drafts to teaching materials, and those endless hours of classroom observation. The teachers occasionally gave us a “peek behind the curtain” and explained why they chose to teach things a certain way or another. I’ve rarely gotten that kind of explanation in any class that I’ve taken so I thought it was pretty special.
I don’t expect you to put in the textbook-level of polish into your article but there is a kind of verbosity (the article is long, which makes it kind of hard to respond to because there is just so much to sift through), there are some problems with clarity (the issue of romanization and orthography is mixed in with the conjugation, and maybe it would be better to separate those issues) some problems with correctness (various) and some problems with completeness (the patterns omit some conjugations that I think you don’t know, and I don’t think they follow the pattern).
I have certainly put effort into articles that have gotten brutal negative feedback; I think it was right for me to write the article, and then feel like shit from the feedback, and then maybe retract and revise it. If there is one actual error here, a true error, I think the error is fighting out criticism in the HN comments.
So, on verbosity: that’s a stylistic choice. Not for everyone. For romanization: point taken and I agree frontloading it would’ve been more elegant. Though I kind of don’t like that it sounds wrong for an unprepared speaker.
For correctness: please provide specific issues. I’ll try to fix them. This is the part I actually care about.
For completeness: yes, some things I put out of scope break the pattern (or rather extend it — the mechanism of concatenation is the same but it actually may be easier to hard-split it by godan/ichidan). I genuinely think that by the point you learn those, you don’t need the scaffolding anyway, and the model has done its job.
I don’t feel like shit from the feedback. This is not my first rodeo. Where I have correctness issues, I would like them pointed out so I can fix. The handwringing about it being a weird way to teach — not so much. I know it’s weird; I wrote it because that’s what worked for me.
And fighting out the criticism in HN comments is half of the fun, isn’t it? :)
Again, the conceit of the article is you can learn almost the entire conjugation system in a single evening with no prior knowledge of the language. I invite you to step back for a moment, to accept that conceit as valid, and then to judge the article based on that conceit. For a serious learner, think of it as a fever dream that helps the concepts click next time you see them “properly”. For a tinkerer, think of it as a spark that gets you curious about the language.
I assume no prerequisites at first. So my reader has never seen a kana table and doesn’t know which syllables exist.
I choose to teach conjugation first. That’s an unorthodox choice but I like it! That’s what I set out to do. So we get far enough until it breaks down. And it breaks down when a rule (which worked so far) doesn’t help with “s” because saying “si” would sound wrong.
That’s the moment I use to teach kana table and its importance. This “you made a mistake” is a pedagogical vehicle for introducing kana rows. And we go over the exact ones that you’d make a mistake with. So each special case is walked through.
At this point we could discard Hepburn but I choose to keep going because if you know special cases, there’s no issue. And at some point you’ll learn kana anyway.
So that’s how I chose to layer it. Maybe it’s a bit unholy but I like it. It is definitely self-consistent.
1. It relies on people not understanding certain things. In general, you cannot expect people to have exactly the right misunderstanding necessary for a lesson.
2. Spending extra time with Hepburn reinforces it, and it shouldn’t be reinforced.
I am in general extremely skeptical of lessons which try to engineer a way for the students to make mistakes. What I have seen in real classrooms and in informal teaching is that the mistakes are habit-forming and the outcomes of this kind of engineering are unpredictable.
Mistakes are appealing to the developers on HN because we understand things more by seeing them fail. But this does not mean that you can engineer somebody to experience the same moment of enlightenment that you did, because it requires constructing the same (incorrect) mental model that you had when you made that mistake that led to useful insight, and it both difficult and counterproductive to try and make that happen to students. Give people the best chances to learn by giving them the best chances to avoid mistakes, and the mistakes and insight will happen organically on their own, in unique ways for each student.
I also genuinely think it’s not that deep and that there’s no complex mistake being engineered here. I don’t believe you that the “mistake” of “sa with a replaced by i must be si” is an an unusual one for someone who hasn’t yet internalized kana. If we test this on random people on the street, I’m highly confident an overwhelming majority will make this exact mistake.
I agree with your broader point that “teaching via mistakes” is a risky path not worth it when the mistakes start getting combinatorial. I also think it’s absolutely fine when everyone does the same exact mistake, and there’s exactly one way to avoid it.
People off the street are obviously not learning Japanese verb conjugations in isolation. If they are learning it at all, they probably have some broader goals involving spoken or written fluency, and these people are gonna fire up DuoLingo or sign up for a class or something. Japanese verb conjugations are simple and easy to learn but they are usually not taught day 1, and you are not expected to learn the whole table at once, but one or two conjugations at a time along with practice using that conjugation.
So if the pitch is, “this system works for teaching English speakers off the street how to conjugate verbs in Japanese” it seems to me like the goal is a little artificial and maybe not representative.
I think the call to “engineering mindset” may be illuminative, because engineers are likely to have unwarranted confidence in fields outside of their expertise. Engineers in practice often think that they can use engineering skills (broadly speaking) to solve education problems, learn foreign languages, or solve social problems. The phenomenon is sometimes called “engineer’s disease” or “engineer’s syndrome”. What I wonder is whether there is something about engineering mindset that is counterproductive outside of engineering fields—this seems plausible, because it explains why we don’t just teach everyone to use an engineering mindset.
I never claimed it's representative of anything. I said this is the explanation I wish I (me, personally!) were given, and I wrote it for people like me. I appreciate unorthodox explanations as a genre. As long as they're rigorously correct (again, you're welcome to point out factual mistakes), I like experiments like "learn a non-trivial part of the language with no prerequisites as a syntactic transformation in one evening". For many languages, including my native language, this is literally impossible! But for Japanese, it works. Maybe that sort of explanation is not to your taste, but it doesn't mean that it doesn't deserve being written. The rest of your comment reads kinda ad hominem.
From what I can tell, my engineering-brained explanation is consistent with how a linguist would explain it (aside from choices in presentation like romaji). That's good enough for me.
When I write “it’s not representative”, I am hoping to communicate an opinion and not hoping to refute a specific claim you made.
I appreciate unorthodox explanations, and I like to collect them—but sometimes the explanation just doesn’t “land” and in this case the explanation landed especially poorly for me, and I also identified some errors, like the claim that “si” is not in the kana, or that hanasimasu is incorrect—I know that you don’t accept my viewpoint that this is incorrect—sometimes it happens that explaining your point of view or reasoning in more detail doesn’t result in agreement.
What is certainly true about linguists is that they do not present explanations in ways that are consistent with each other—and sometimes not even self-consistent, but they are upfront about the tradeoffs (they present theories and acknowledge that the theories contain errors) and they decompose what they present into different topics like phonology and morphology.
I have read some linguistic texts on Japanese (not very well! It’s a difficult subject). What I saw is a lot of variation.
I think it’s fine that your explanation is good enough for you—that’s exactly what you expect when people make individual breakthroughs in understanding when learning a subject. Sometimes those breakthroughs do not translate well to other people or translate well to lessons and that is the main gist of what I am trying to articulate.
"Romaji" does not (in English) mean "romanisation", as most people who've studied Japanese to at least beginner level know.
> This method of writing is sometimes referred to in Japanese as rōmaji
See how there's not actually a contradiction there?
If you’re nitpicking on this sentence:
> I thought it’s acceptable to refer to Hepburn as just “romaji”
I’ll expand it:
> I thought it’s acceptable to refer to Hepburn [-flavored romaji] as just “romaji”
I think it’s clear by context what I was saying.
Hiragana also has its problems, because the hiragana used before WWII corresponded with an ancient pronunciation of Japanese, from many centuries ago, which no longer matched the modern Japanese pronunciation.
After WWII, under the American occupation, there was a reform of the writing system, which replaced many kanji used before WWII and it also changed the spelling in hiragana of many words.
In general the modern hiragana spelling has been changed to match the modern pronunciation, but there are a few survivals of the older spelling that lead to inconsistencies.
As an example, the hiragana syllable now romanized as "ha" was pronounced for some time several centuries ago as "fa-" in initial syllables and as "-va-" in internal syllables. Then the pronunciation shifted to "ha-" in initial syllables and to "-wa-" in internal syllables. After WWII the "ha" hiragana character was replaced by the "wa" character in most internal syllables, to match the new pronunciation, except in the "-wa" postposed particle, where the "ha" hiragana character was retained, despite the pronunciation. The particle is now romanized as "wa", so going backwards to hiragana would produce the wrong hiragana character, another example of non-bijectivity, besides "zu" and "ji". Yet another non-bijectivity example is that the postposed particle normally romanized as "-o" actually uses the hiragana character "wo".
The changes in hiragana spelling after WWII are also responsible for the fact that many Japanese words reproduced in old books written in English, e.g. from the 19th century, appear quite different from how they are written today in the modern Hepburn romanization.
A spoken language is described by decomposing the spoken words into phonemes, where phonemes are sounds that distinguish words, in the sense that replacing one phoneme in a word with another phoneme will produce a different word.
While ideally each phoneme should be recognized by a distinct pronunciation, in the majority of the languages of the world a phoneme does not have a single pronunciation, but it is pronounced in different ways, depending on the context.
It does not matter at all how one chooses to write a Japanese word, with hiragana or with one of the various methods of romanization. For any writing system, you must know the correspondence between phonemes and how they are written. For very few writing systems there is a one-to-one mapping between phonemes and letters.
The Hepburn romanization does not attempt to be a phonemic writing system, but it attempts to be close to a phonetic writing system from the point of view of an English speaker. The Kunrei-shiki romanization attempts to be closer to a phonemic writing system than to a phonetic writing system. I my opinion a phonemic writing system is superior to a phonetic writing system, but it appears that for most English speakers it has been too difficult to understand the difference between such writing systems, so the Japanese government eventually gave up and they switched to Hepburn, to please the less sharp-witted English-speaking visitors.
Japanese has an "s" phoneme, which happens to be pronounced differently before the vowel "i" than before the other vowels, and before "i" it is pronounced similarly to an English "sh".
In the same way, the Japanese phoneme "t" is pronounced before "i" similarly to an English "ch".
Once you know these two rules, and the few other rules about the other Japanese phonemes whose pronunciation depends on the context, like "n" becoming "m" before "b", there is no point in mentioning them again.
In your discussion about conjugation there is nothing exceptional about the variations in pronunciation that are reflected in the Hepburn Romanization. They are just the general rules of Japanese pronunciation, like for any other words.
So any discussion about these spelling variations is misplaced in the discussion about conjugation, where it occupies a space without contributing anything to the understanding of the conjugation rules.
Otherwise, I think that your article is fine.
I strongly suspect that if I were using Kunrei-shiki, there would be just as many comments here saying my article is wrong because “si” is pronounced closer to English “shi”, but my article makes it seem like it doesn’t — so this is why you should learn kana bla bla bla.
I assume my reader (1) has zero prerequisites and (2) wants words to sound correctly while seeing them the first time. Those are the constraints that motivated my approach. You could argue that it’s a strange set of constraints to pick when teaching but I wanted it to be fun.
Obsessing over romanization, something that a student ought to outgrow, is a sure fire way for a student to get overwhelmed by irrelevant details that discourage learning. The hard part is putting in the work, not learning less than a dozen exceptions.
> (note i could also have used a different romanization that renders し as "si", つ as "tu", and ち as "ti" for this article. i decided to not because everyone else uses romaji, and once you understand this point once, you shouldn't have a difficulty doing this in your head.)
I think the choice is not a good one, whether it is deliberate or by accident, it is not a good choice either way. The main caveat to Hepburn is that it’s unsuitable for explaining how Japanese works and it’s unsuitable for learning Japanese—so before you start working on verb conjugations, you pick up kana or one of the romanizations which is more aligned with Japanese.
The idea that you “shouldn’t think in romaji” is really “you shouldn’t think in Hepburn”. This is an important distinction! Japanese has a relatively small inventory of phonemes, somewhere around 20 or 22 of them, and they map very neatly to the latin alphabet.
But the article doesn’t make this distinction, and seems to rely on confusion induced by the Hepburn romanization in order to make its points.
IMO, this is kind of like seeing an article about how monads are burritos. Thinking that a monad is a burrito does not help me understand monads.
Nomu -> noma-nai / nomi-masu / nome-ru / nomo-u
Miru -> mi-nai / mi-masu / mire-ru / miyo-u
The ichidan and godan verbs are not assigned different categories because existing scholars of Japanese are just bad at explaining how they work, and you can still understand them just fine in romaji. I put the hyphens above to mark a place where you could think that the verb ends and the common conjugation forms end, and you can see that the part on the left has somewhat different rules for ichidan and godan verbs, even when you apply the “tricks”—but some of these forms may be unfamiliar if you are are starting out (are you familiar with miru -> miyou conjugation, or miru -> mirareru?)
> But the article doesn’t make this distinction, and seems to rely on confusion induced by the Hepburn romanization in order to make its points.
Not at all. I give it two sections and then we move on. It doesn’t affect literally anything else on the page. You just learn to shift rows and move on. To make what points?
> you can see that the part on the left has somewhat different rules for ichidan and godan verbs, even when you apply the “tricks”—but some of these forms may be unfamiliar if you are are starting out
I’m not quite sure what you mean to say in this part. I do cover -[r]eba and -[y]ou in the final section (“one more thing”) which extends the model to clearly handle that disappearing consonant. I think -[r]eru fits in there the same way, just as -[r]u itself.
I think explaining it as mi + [y]ou = miyou, but nom_ + [y]ou = nomou is a clearer way to think about this. The rule is that the hole burns down the leading consonant (but takes the vowel).
Oh and I forgot, you have to actually learn how to listen, pronounce and speak them, not just learn a useless romanization mapping. I've heard way too many English speakers just say the romanization with English pronunciation. At that point their learning efforts turn into self sabotage.
In total that's definitively a month of effort, albeit spread out over the first year of learning.
I also drilled on a drag-n-drop kana table [1] in a few ways -- sometimes I'd start from the kana and try to figure out where they should go in the table, and sometimes I'd go along rows or columns in the table and try to find the kana that belong there. These two directions drill both recognition and recollection.
Proper pronunciation is a cross-cutting concern. As a whole, it's not something you can reasonably learn solely from kana, but the aspects that are relevant are not difficult to pick up. Every kana breaks into one (vowels and N) or two (the rest) phonemes, and for the most part, the way you pronounce those phonemes is consistent across rows and columns of the table (admitting exceptions like "shi" and "tsu"). If you are taught those basics, learning how to pronounce kana is not hard. Training your ear to "hear" distinctions among English allophones, and to distinguish pitch accent from the more familiar stress accent, is much harder, and really has to come from experience, not just kana.
[0]: https://realkana.com/hiragana, wow it's improved since I last used it
It takes two weeks to know know the existence of all the kana characters (including katakana), to memorize the sound of enough of them to read some words, and to write some of them.
After a month you should have easily memorized the sound of all of them (maybe a rare one like ム slips by occasionally), be able to write most of them, and be able to read (albeit slowly) anything written in kana.