Great Vowel Shift
en.wikipedia.org
en.wikipedia.org
The classic English author for whom the Great Vowel Shift is most relevant, in terms of audiences today not pronouncing the text anywhere near as society then would have, is Chaucer.
(Imagine other examples, like pronouncing 'couch' a bit more like 'cooch')
Is bikeshedding a habit or a temperament?
Glad to see someone managed to get the joke, anyway.
Aside: I see "term of art" everywhere on HN lately. Why?! What does it add to saying "Vowel shift is a linguistics term" or "a term in linguistics"? Why do people want to sound like patent lawyers? Am I missing something?
Idiomatic:
A speech form or an expression of a given language that is peculiar to itself grammatically or cannot be understood from the individual meanings of its elements, as in keep tabs on.
A specialized vocabulary used by a group of people; jargon.
Since idiom has both the structural denotation of a phrase being semantically 'atomic' (you lose the meaning if you break it into pieces), plus its synonymity with 'jargon', there appears to be nothing added in "term of art".
—but maybe this is a lesser known meaning of 'idiom'.
These words are not synonyms, and should not be substituted.
I think you’re overstating your case though when you say “It does not have the sense of a specific technical meaning for a word, distinct from the ordinary definition”—since jargon is a synonym for one sense of idiom and idiomatic is less specific in the group/individual distinction (and yes it’s an adjective but you can trivially employ it to construct an equivalent noun phrase so this matters little). That said my own case is clearly a stretch here lol.
Maybe more relevant is the fact one could just say “technical term” and they’d be understood perfectly—I’m pretty sure that phrase is the reason I’ve also never come across “term of art”.
I’m curious whether you would consider “domestic partnership” to be both jargon and a term of art (I would).
It’s interesting reviewing definitions of ‘jargon’: while you do find things like e.g. “ specialized technical terminology characteristic of a particular subject” —mostly you find references to incoherent/nonsensical speech.
None of these replies justifying "term of art" seem at all convincing to me. "Term doesn't imply a specialized meaning" – but I think it does here: if "clobber" didn't mean something different in programming, you wouldn't (need to) say "it's a programming term". "It's a programming term" precisely means "the term has a meaning in programming different to what it commonly means". The listener already knows it's a word in everyday English, which is apparently all "term of art" adds.
[edit] I am pleased to see that HN readers as a group love to define things, and also that I need to learn to refresh tabs I've had open for a while before responding
Term of art doesn't really have a clear meaning in general (native English speaker, and this is the first time I've heard this), and it's unclear/ambiguous as to what it actually indicates.
"Domain-specific" is nonsense gibberish to people outside ~computing, more or less.
"Term of art" only implies that it is the specific phrasing used by practitioners to describe the matter at hand.
Use "term of art" if you'd like. I think "jargon" is a commonplace alternative and I like it.
However, I was already familiar with "start of the art" so it wasn't that great a leap.
Term doesn't imply a specialized meaning. Term of art does. And it doesn't have a negative secondary meaning like jargon.
In Math, lemma is a term of the art.
That's every word in existence
Could a computational linguist chime in please? How many words in how many human languages will never be spoken again in 2021?
> The original method presumed that the core vocabulary of a language is replaced at a constant (or constant average) rate across all languages and cultures …
Unfortunately, the only major discovery glottochronology has revealed is that rates of change vary too much to be of any use:
> in Bergsland & Vogt (1962), the authors make an impressive demonstration, on the basis of actual language data verifiable by extralinguistic sources, that the "rate of change" for Icelandic constituted around 4% per millennium, but for closely connected Riksmal (Literary Norwegian), it would amount to as much as 20%.
And there are other factors affecting replacement rate as well. For instance, a curious trait about the non-Austronesian languages of New Guinea is that the word ‘louse’ is practically never replaced: it evolves according to normal sound change, but never gets thrown out entirely. By contrast, many languages have a taboo against mentioning the names of the deceased; this can speed up lexical replacement if those people have names homophonous with common words.
For these reasons, I suspect that a general answer to your question will be extremely difficult — if not impossible — to find.
Probably a combination of the Baader-Meinhof phenomenon and virality (meme).
https://www.macleans.ca/society/life/in-the-midst-of-the-can...
I’m an American, recently back from 2 years in Toronto, and the Toronto accent is just not the stereotypical Canadian accent that I imagined I’d hear. Far less Minnesota, a bit more New Jersey.
Maybe related to what’s happening in Canadian English?
Historical linguistics is a really cool intersection between anthropology and linguistics - in short the idea is we can look at languages around today and use certain well-founded assumptions about how languages change to understand the languages of the past; essentially the words we speak and sign today are the fossils of our linguistic history!
Suggested reading is Historical Linguistics by Lyle Campbell, and Language Files for a more general Linguistics textbook.
This is why folk etymology usually misses the mark—since it focuses on single similar words that in fact often turn out to be unrelated. Also afaik spelling is irrelevant for etymology, at least until the very recent times.
There were a lot of sound shifts.
I'm 29 and grew up in the Midwest, so in my accent, when I say "button," the "tt" is nearly silent. however, at my last job doing remote web development, one guy who was a bit younger than me (and also American) would pronounce it "BUH 'in," with noticeable "stop" between the two syllables. I have since noticed this in other younger Americans as well, and for other words I cannot recall right now but with the same general pattern.
If you are "educated", then you are "supposed to" pronounce the middle constantans, as skipping them is considered "lazy". I'm not making a value judgement, but reporting that this is what parents often tell their kids in private. All things being equal, parents want their children to sound wealthy and well educated. Similar for certain rural or "redneck" talk in the US. Example: "Murica" instead of "America". My mother used to lecture me against certain verbal shortcuts so that I didn't "sound ignorant". Her words, not mine.
I agree. The downside of the loosening of the old standard is heard in the number of actors who mumble through their lines. I wouldn't want to go back to a time when RADA enforced a kind of RP but I would like them to focus just as hard on diction.
[0]: https://books.google.co.uk/books?id=YMS3AwAAQBAJ&pg=PT26&lpg...
You might feel a lot more comfortable talking about this subject and find useful alternatives to your scare-quoted words if you skim https://en.wikipedia.org/wiki/Sociolinguistics
Also due to https://en.wikipedia.org/wiki/Hypercorrection ,I think https://en.wikipedia.org/wiki/Rhoticity_in_English is a fascinating and accessible example. A terrible summary: dropping the R is lazy (dropping anything is "lazy"), but at the same time sounds British and therefore fancy to some American ears. But other people, to avoid laziness, add Rs. But then their speech can sound low-status in some cases too.
I correct my childrens' pronunciations. I've honestly never thought of it as a class thing. In part it's "inherited", so it could have consciously been a class thing for one of my ancestors. But really to me it just seems necessary for preliterate speakers to understand the proper way to say a word based on its spelling, which they don't know. I find myself correcting USA-ian word use ("garbage") and pronunciation far more with my youngest than I had to at that age with my oldest ... we probably let him watch too much TV (several shows, like Paw Patrol have shifted to USA versions from British English versions).
It's hard to discern every phoneme from a new word, and hard to say some of them.
I do shift my accent, and vocabulary, to mark myself as a local I guess (when I'm back in that area of the country). If I get to use certain words in their localised (to about a 20mi long area) meaning it makes me happy for some reason. But I've never consciously worked on my accent to sound more/less posh; but have modified it to be more understood.
My wife is what I call a 'sympathetic speaker', she very quickly adopts the accent of those she's speaking with. An interesting phenomenon.
Politicians are often known to do the same. It's usually painted as "pandering", but I'm not making a value judgement here. There's arguments both for and against.
My anecdata says we covered this dialect in intro linguistics back in the 80s. I have been hearing it from New Jersey natives for a long time.
The typical American "nearly silent" one you are describing tends to be more of a flapped /ɾ/, by the way. <d> is often the same.
yes, exactly! also out here in South Dakota we specifically pronounce the "t" in "Dakota" as a "d," I've noticed.
Historic shifts like that are loafes -> loaves.
One of those "I was there" -no, you weren't things. Like sea level rise. You see a bit of it. The totality is a story which spans generations. Nobody owns all of it.
One thing though, take older movies with a grain of salt; in old movies, actors were taught a specific accent (called the Mid-Atlantic accent, https://en.wikipedia.org/wiki/Mid-Atlantic_accent).
Modern example. Router. Two forms, the British one rooter is being deprecated strongly for rowter.
Etc is etcetera. Now, its etsy.
I can't say I've ever heard that, anywhere.
Definitely have heard a lot of "esetra" / "esedra", though.
/tmp is pretty much temp to everyone. /var vahr not vair. and luckly, we all think soodoo is pronounced the same. Oh wait, soodough. Damn.
[0] https://historyofenglishpodcast.com/2019/09/25/episode-129-c...
I love that history is able break your notion of what is or isn't possible and this podcast is great at that; it repeatedly shows how no language is set in stone, it is a human construct, and how languages are interrelated.
There's a fantastic dialogue illustrating the similarity of Old Danish / Norse vs Old English by Jackson Crawford and Simon Roper: https://www.youtube.com/watch?v=DKzJEIUSWtc
It’s not really surprising considering that region of England used to be colonised by “Danes” and was called The Danelaw before the Norman conquest, as against the south Germanic tribes that settled southern England. I suppose the Norman french influence peters out the further north you go.
This must be the world's largest tech debt. :)))
Obviously, any artificial language is going to be much simpler. But these have never caught on for a variety of reasons.
Spanish is very willing to accept new words, and as diverse as English in terms of decentralization. Grammar doesn't make or break a lingua franca, number of speakers does, which is where English really shines.
So what does the RAE do you ask? They write grammars and compile dictionaries, describe phonology and answer people's questions on Twitter. Just like what Merriam Webster or Oxford would do, but the RAE has official backing and creates consensus among the hispanophone countries. English is a regulated language, just not officially regulated.
Some other language authorities do not take a usage-evidence-based approach to defining their dictionaries, and take into account cultural or historical concerns.
The role of regulation is mostly a legal role employed by countries that use a civil legal system. Most English speaking countries use a common law legal system and so interpretation of words is fairly fluid and subject to interpretation by the courts. Civil law does not have this kind of flexibility and room for interpretation, so courts will make use of regulatory bodies to uniformly interpret language.
This is why almost all countries with a civil law system have a regulatory body, while countries with a common law system do not.
This doesn't mean that learning the language is easy--no language really is. And English has some things that make it harder to master, especially its very large vocabulary.
Some of it is covered in this recent episode: https://slate.com/podcasts/lexicon-valley/2021/03/english-la...
Vocabulary isn't a problem IMHO.
The fact that I have to learn each word twice (how to write it and how to say it) - is.
I was learning German for 3 years at school. After the first month I had no problems with pronunciation. Now after almost 2 decades of not using it I can still pronounce any German word I see.
I've been learning English since I was 10 or so. I'm 36 now. I still have many English words I know (and use correctly in writing) that I'm not sure how to pronounce.
Do you mind sharing some examples? As a native English speaker, I'm so curious! Do you think that if you heard them without seeing the word you'd realize what the written form was? Or might there be words where you know the written form and the spoken form and don't realize it's the same word?
There are quite a few others that tripped me up over the years like "cleanliness".
In general, "Chaos" aka "Dearest creature in creation" shows this problem (I would still struggle to read it even if I know every word there): https://pages.hep.wisc.edu/~jnb/charivarius.html
Parallel - for some reason the second a is eh not ah. Can't remember that, have to check it every time.
I play a lot of D&D over the internet in English and even as common word as "sword" is for some reason hard to remember. Every time I have to guess if the "w" is pronounced or not.
been == bin ? - the rules for that are just evil
> Do you think that if you heard them without seeing the word you'd realize what the written form was?
Sure, from the context if not instantly. I listen to a lot of English media with different accents (I watched the whole Big Bang and IT Crowd and I listen to Critical Role when I'm commuting).
> might there be words where you know the written form and the spoken form and don't realize it's the same word?
Leicester and queue. But these are famous enough that I remember them now. I obviously won't be able to give you examples that I still haven't realized ;)
> for some reason the second a is eh not ah
> been == bin ? - the rules for that are just evil
Actually, the rules are rather simple. They all have to do with unstressed syllables in English: unstressed vowels are reduced to /ə/ or /ɪ/ (the latter is what comes to play in your been -> bin use). Stress rules in English are not simple compared to other languages, and I can definitely see where non-native speakers might get confused.
One downside of the schwa reduction rules is that it can trip you up when you realize that you need to spell a word with a reduced vowel and you're not sure how it's actually written, because every vowel can be reduced to /ə/.
But the best examples are in the poem _The Chaos_ (http://ncf.idallen.com/english.html).
Through, though, throw, tough... I know a trough exists but I have no idea how it's pronounced.
Trough is /tɹɔf/ and rhymes with "cough".
For example I'd seen adenovirus written down, but never heard it said out loud, I was describing the vaccine I'd had to friends, one of whom works in medicine (a doctor, but not of medicine) and she corrected my pronunciation because she's used that word plenty of times so she (presumably) knows how to say it correctly.
Even for a completely native immersed speaker, there's just no clue in English how to correctly say a completely new word you've only seen written down, so you're at no disadvantage there. For "real" words there may be an etymological clue, but those aren't reliable. In fiction it's anything goes. Hearing fictional words I've read pronounced out loud in movies is as weird for me as seeing the (inevitable) transformation of a woman described as plain in the books into a beautiful Holywood actress...
It's obviously a bigger problem with some common English words - either where they are actually two separate words with different pronunciations but the same spelling, or worse, one word but with different stress patterns. But once you've got a fair-sized vocab the new words you're learning won't have that sort of weirdness.
It's definitely true that if you're not confident pronunciation can really be an obstacle, fortunately the huge vocab helps again - a (non-English native but UK citizen) friend of mine will carefully choose to talk about liking the "seaside" never the "beach" because she's concerned she'll manage to make people think she said "bitch". She has a few other words like that, in each case English provides convenient alternatives.
Is her native language Spanish?
It's really cool how you can have "blind spots" depending on your native language. To me, the difference between beach and bitch is huge, because my native language uses short and long vowels extensively, and there are tons of words that only differ in a single vowel length.
But at the same time, I have other blind spots in English. For example, I have to make an effort to remember to use sounding "s" and "j" where appropriate, and the lack of those is a dead give-away for identifying Swedish English speakers.
Thanks for that anecdote, I somehow thought that I am alone living with fear of that happening :)
I don't know why but a lot of people my age do worry about it and get very uncomfortable guessing at a pronunciation of a new word. Doubly so for names. I guess it's an insecurity thing? I have no problem just going for it, with a little question tagged on or just an upward tone if I'm really unsure.
See https://pl.wikipedia.org/wiki/Og%C3%B3lnopolskie_Dyktando
But the mapping sounds->letters isn't as obvious, because there's some tech debt there (ch == h, ó == u, rz == ż or sz depending on the preceding letter).
So if you know how the word sounds you're not always sure how it's spelled, but if you know how it's spelled you always know how it sounds.
In English the mapping is non-obvious both ways.
Adults are really good at learning new words, but struggle with morphological and gender systems. Complex morphology seems to have some benefits that push languages towards including them, but are only sustainable when those languages are predominantly learned by children, who can pick up morphological complexity easily, not adults, who can't.
You have no idea :)
Having spoken English for almost 30 years by now, I am still not sure if "X has died" or "X died" is correct, or in which context.
But the other direction would be worse. Languages like Czech, with its 20+ classes of declension, must be a true nightmare for any English native speaker to learn.
Okay, strictly speaking, this is a distinction in aspect, not tense. But colloquially, tense, aspect, and mood are all referred to as "tense", especially since the conflation is present in most Indo-European conjugation patterns.
"X has died" is the present perfect. The perfect aspect is kind of like the past tense in that it is referring to something that has happened. Indeed, the past perfect ("X had died") is usually described as "the past of the past". But in keeping the tense in the present, the present perfect means that the past occurrence has relevance to the present. This can carry a few connotations. It can be a recent past, especially if you use "just" as an infix (c.f., "X has just died"). Or it can highlight the consequences of the event having occurred (e.g., "Our lord has died. What will become of us now?"). In any case, the speaker is drawing the listener's attention to the connection between past and present when they use the perfect aspect.
So which is correct? "X died" you would expect to find more in a biographical context or maybe a novel. "X has died" would be common in a news report, or someone informing you that a loved one died not too long ago. Which is more correct in a given scenario can usually be informed by the dominant tense in surrounding text; after all "X has died" is the present tense, despite conveying an action that happened in the past. If there's not enough text to dictate a tense, then it's often the case that either form will end up being acceptable--it just sets the tense that will be used.
"Have you [ever] been to New York?" "Did you go to New York last week?"
"Have you seen Star Wars?" "Did you see Star Wars this afternoon?"
Might not work in every circumstance but a good rule of thumb.
An example of a language that has it worse than English is Tibetan, which hasn't had a script reform since 800: https://www.youtube.com/watch?v=btn0-Vce5ug
I think the fact that it is a lingua franca is one of the main reasons keeping any spelling reform from occurring actually. It's in far too wide-spread use and there isn't any centralize authority that would do the spelling reform. Maybe take solace in the likely fact that it written English and spoken English will probably only be _more_ different as time goes forward. In other words, you have it easier than all future generations.
Also as much as spelling matching pronunciation is a convenience, it isn't really necessary. The variety of spoken Chinese languages using the same characters is greater than the spoken romance languages. Maybe English really slowly becoming more character-like over time. There are languages that can undergo spelling reforms, and there are languages people actually use.
There there is a very real cost to not mapping spoken and written languages: kids need to spend more time in school learning basic reading skill - time that could be spent either learning something else, or playing.
As I recall it speaks about the desire by the Japanese leadership of the time to build AI translation and related it to the challenge of full literacy in Japanese due to the requirement of learning 4 alphabets. Hiragana, Katakana, Kanji and Romanji.
Written Finnish is mainly a written thing, and while there is a close correspondence between the orthography and how it would be read aloud, when Finns actually speak they use spoken Finnish (puhekieli), which isn't standardized and varies from region to region.
When reading Finnish, one pronounces the words very closely to how the word is spelled, regardless of whether when they speak they do so in an altogether different manner.
Of course, it could have been worse. We could have ended up with French as the lingua franca (yes, I know what franca means) where there is almost no correlation between written and spoken language.
Going from spelling to pronunciation in French follows (admittedly complex) rules that are rarely broken except for common words (or endings such as -ent). Vowel pronunciations for a given spelling are far more variable in English, and often depend on the etymology of the word. Plus, English has word-level stress that is not marked in writing (French has none, and it's marked in Spanish), and moving the stress will usually make a word unintelligible! That alone makes writing => pronunciation very difficult.
Unsurprisingly, we can vaguely quantify this by looking at dyslexia amongst languages. English and various Southeast Asian languages that rely on Chinese ideographs are by far the worst, followed by things like Arabic, French, Hebrew, and German that have fewer exceptions but less guidance, and then followed last by things like Spanish, Cherokee, and so on that are truly one-to-one.
There are a number of ways currently used, but I have a new one to propose: compare the size of two G2P models (1 for each language), which have similar RMS errors. Assuming they are generated using similar techniques, the one which requires the bigger model probably has a less clean phoneme-to-grapheme correspondence.
Even ignoring all of these, its clearly not bijective. For example:
C --> /k/, /θ/
Z --> /θ/ [0]
K --> /k/
Q --> /k/
G --> /ɡ/, /x/
J --> /x/
N --> /n/ (with several distinct secondary articulations), /m/ (rarely)
M --> /m/
R --> Can be tapped or trilled.
Etc. You can go here and see many bijection-failures here: [1]
I am being intentionally unfair to Spanish (which truly does have a much, much better phoneme-grapheme correspondence than English[2]), mostly to illustrate the point that there aren't really any languages which have a 1:1 mapping between spellings and pronunciations. Even if you decide to use the IPA to write your language, non-standard dialects end up needing to read words that don't match their pronunciations. What happens when inevitably the language undergoes change - do we update all of the books to use the 'new' spellings of words?
The ideal orthography shouldn't be completely 1:1, but it should be relatively shallow. From that perspective, Spanish orthography is a fairly attractive option.
[0] The non-1:1 situation with /θ/ gets much worse in most dialects of Spanish, where it is not distinguished from /s/. See: https://en.wikipedia.org/wiki/Phonological_history_of_Spanis...
[1] https://en.wikipedia.org/wiki/Spanish_orthography#Alphabet_i...
[2] Look at how effective Spanish-speakers are at reading without "decoding" compared with Portuguese, which also has a good p-g correspondence. In particular, look how much faster the Spanish students are at pseudowords, on page 141: https://www.academia.edu/17872463/Differences_in_reading_acq...
The function from written Spanish to spoken Spanish (provided we are talking about a single dialect) is surjective, but darn close to bijective, especially if we exclude words of recent foreign origin.
> Many people expect … to predict the spelling from the pronunciations-- not realizing that few orthographies meet this goal. It's far from true of Spanish, for instance, which is often held up as an example of a good orthography. I stopped fervently admiring Spanish orthography when I saw a sign in a Mexican bakery with about one spelling mistake every third word.
So, no, hardly bijective!
I do, however, wish natives English speakers were more aware of this for their own sake. Most seem to default pronouncing vowels in foreign words as if it was English, whereas they'd be much closer to the correct pronunciation of they defaulted to pronouncing it like words for any language other than English they might know even a little of. To me this should be one of the first things you learn when learning your second language as a natives English speaker. It even holds true for romanizations of Asian languages like pinyin for Chinese or Hepburn for Japanese
For example, in linguistics studies AAVE (or “Ebonics”) is considered like other dialects to basically have a grammar of its own that is internally consistent, and it is the first English of a fairly large population.
That’s before we get into which English is the “correct” one; there’s British English, American English, etc. for the US and CANZUK, but there’s also Indian English, Singaporean English, Euro English, etc.
There's not just descriptive and proscriptive approaches: utilitarian and apathetically approaches seem distinct (perhaps they're a different axis).
Surely tons of existing movies and songs and other media should give a clear idea of a speaking norm so spoken language doesn't deviate much from it. Or it doesn't?
> Uisce beatha, literally "water of life", is the name for whiskey in Irish. It is derived from the Old Irish uisce ("water") and bethu ("life").
In France, many liquors used to be simply called "eau de vie". They now tend to each have their own name (origin and/or brand), but most people would still understand that you're not looking for water if you ask for some eau de vie.
https://www.tiktok.com/@therealivancohen/video/6921416585412...
If you are interested to know more, take a look at the way native speakers of other languages utter "a" and "o" and their variations. Fascinating, to say the least.
Also, vowel chain shifts are exceedingly common across the languages of the world: compare Tatar and Kazakh to the other languages of their subgroup (Kipchak) within the Turkic family, for instance. Sometimes a vowel chain shift can even be observed in progress, as in the case of the Northern Cities Shift in the USA. [0]
[0] https://en.wikipedia.org/wiki/Inland_Northern_American_Engli...
* the pronunciation of descendent languages, especially when there are lots of them, or lots of different dialects
* comparison with languages that have a common ancestor
* how words were transcribed from one language into another
* how words were changed when adopted from one language into another
* what spelling mistakes were made, particularly in cases where the writer was less educated or being less careful, such as graffiti
* rhyme and metre in poetry
* in literature, cases where someone is mocked for their pronunciation
* puns and wordplay in literature
* and of course cases where ancient authors have written more or less explicitly about pronunciation, either descriptively or prescriptively
It's worth calling this out as being one of the key elements of historical linguistics. Establishing genetic relations between languages requires proposing systematic sound shift laws that can explain why cognates sound the way they do, and this means that cognates in modern languages may not bear much resemblance to each other (English five and Sanscrit pankan are cognate yet share 0 sounds!). For example, there's a rule in the Germanic languages that shifts /k/ to /h/, so words like Latin "centum" instead become English "hundred" or "canis" to "hound" [1].
Now these pronunciation shifts often have caveats in them, such as shifting only before certain kinds of vowels or consonants. These restrictions can give you some clues as to why certain words seem to undergo a change while other words with seemingly similar pronunciations didn't. In Proto-Indo-European, this leads to the notion that there are several consonants (specifically, laryngeals) which are no longer present in any modern Indo-European language but whose existence in the original is responsible for sometimes shifting vowels that otherwise appear somewhat anomalous in descendant languages.
[1] To be clear, Latin is not the initial word and English is not the final word. I'm just using Latin to illustrate a word closer to the original Proto-Indo-European pronunciation and English to illustrate what the Proto-Germanic pronunciation shifts towards.
For instance, "battle-sweat" to mean blood.
* poetry that has a rhyme scheme, so you know that two words rhymed for the author
* puns and wordplay (Shakespeare is overflowing with this kind of thing) that only work if words were pronounced a certain way
* misspellings or variant spellings, which before the typewriter era would be almost entirely sound based (i.e. unlike typoes in general)
* borrowings into other languages, which generally are "frozen" at the time of their borrowing and from then on start mutating by the rules of the other language
* related languages with cognate words, if you can devise a coherent set of sound-shift rules to reconstruct the shared ancestor-word and the intermediate forms along the way
* complaints from the older generation about how kids these days are mispronouncing words (these are always fun, and usually come with specific, if informal, descriptions of both the old and the new pronunciation)
https://github.com/yogurt-cultures/kefir/blob/master/kefir/p...
Also I shifted one Const sound in the wikipedia article.
"The European Commission has just announced an agreement whereby English will be the official language of the European Union rather than German, which was the other possibility.
As part of the negotiations, the British Government conceded that English spelling had some room for improvement and has accepted a 5- year phase-in plan that would become known as "Euro-English".
In the first year, "s" will replace the soft "c". Sertainly, this will make the sivil servants jump with joy. The hard "c" will be dropped in favour of "k". This should klear up konfusion, and keyboards kan have one less letter.
There will be growing publik enthusiasm in the sekond year when the troublesome "ph" will be replaced with "f". This will make words like fotograf 20% shorter.
In the 3rd year, publik akseptanse of the new spelling kan be expekted to reach the stage where more komplikated changes are possible.
Governments will enkourage the removal of double letters which have always ben a deterent to akurate speling.
Also, al wil agre that the horibl mes of the silent "e" in the languag is disgrasful and it should go away.
By the 4th yer peopl wil be reseptiv to steps such as replasing "th" with "z" and "w" with "v".
During ze fifz yer, ze unesesary "o" kan be dropd from vords kontaining "ou" and after ziz fifz yer, ve vil hav a reil sensi bl riten styl.
Zer vil be no mor trubl or difikultis and evrivun vil find it ezi TU understand ech oza. Ze drem of a united urop vil finali kum tru.
Und efter ze fifz yer, ve vil al be speking German like zey vunted in ze forst plas."
Mai vot gos tu som nordik languag like Svedish.
German has significantly fewer irregular verbs than English.
It's ~200 to ~300. (French is double that?)
There's enough to be moderately annoying, but not that bad. Also (in my personal opinion), German irregular verbs tend to be not-as-irregular as English.
For example in Czech:
Jan zabil Petra
Jan Petra zabil
Petra Jan zabil
Petra zabil Jan
Zabil Petra Jan
and
Zabil Jan Petra
are all equivalent to the English "John killed Peter." Change Jan to Jana and Petra to Petr and all 6 of those become "Peter killed John." Even more confusing to a foreigner learning it is that Petra and Jana are the feminine forms of those names in the nominative case.
Start by purging English of French influence, starting with words where words of Germanic origin exists with a similar meaning. Simplify German grammar. Undo some consonant shifts. E.g (the German and Dutch I had to check/adjust w/Google translate; no guarantees for accuracy):
Swedish: En dag kan vi alla tala samma språk
Norwegian: En dag kan vi alle snakke samme språk (or "det samme språket")
Danish: En dag kan vi alle tale det samme sprog
German: Eines Tages können wir alle dieselbe Sprache sprechen
Dutch: Op een dag kunnen we allemaal dezelfde taal spreken
English: One day we can all speak the same language
Now consider "speech" as an alternative to "language" in English (alternatively: "tale" is valid but archaic in Norwegian in this context and we have the English cognate "talk"), and undo that D->T consonant shift in German (e.g. compare Tag to Low German "Dag"), and replace "the" (compare det/de/die/das/der etc.).
There are a whole lot of simple spelling and sound changes that'd bring the above languages a lot closer together very easily.
Of course it's easy in theory - in practice I've lived through multiple Norwegian language reforms and know how excruciatingly slow it can be to get people to adapt (e.g. Norway changed the spoken form of numbers above 20 in 1952 from the equivalent of "four and fifty" to "fifty-four"; my parents learned the new forms in primary school, yet I still picked up the old forms from them in the late 70's and still switch back and forth between the old and new forms now)
But we all know that it should be Esperanto.
Romania also had some spelling reform, albeit it was more motivated by a desire to distance itself from a communist past and not cleanup of tech debt.
The only power they have is influencing the way kids are taught at school. Everything else changes by social pressure and exposure - most media choose to follow the new convention and people get used to it over time.
The changes are very gradual - the only big one I remember was in 90s - changing how "not" was written with adjectives and adverbs. The rules got much simpler so few people complained.
To pick a particular pronunciation-based spelling of words would therefore be to prefer one region over another. This would at least trigger a monumental North vs. South argument, assuming that the plans survived the inevitable knee-jerk reactions / incredulity of the usual media suspects.
Probably best if we stick to arguing about daylight saving time, or changes to the format of cricket matches.
But then again, those rules are pretty standard (c/g before e/i [the "skinny vowels"]) are almost always soft and they are hard otherwise. If you need a soft "c" sound before an "a" then you just use the letter "s". So maybe it's not even worth the effort.
Um, yes, I can't think of regular words where a "c" changes.
<thinks a bit more>
Place names. Place names - at least in England - can have significant differences between spelling and pronunciation, and the locals will often use or be aware of a local pronunciation that isn't obvious to outsiders. Examples include Bicester ('bister'), Leicester ('lester'), Salisbury ('solsbry'), Tottenham ('totnam') and many others. It's not quite the same effect as with dialects but it certainly complicates spelling reform.
I had a 85 year old (yes) technical writing professor in college who insisted that we do our papers using the "Queen's English" that he learned growing up in Catholic school. He even went so far as to write a damn book outlining his rules for English writing because none of the existing style guides out there matched his view of the language.
You'll never get people like that to adopt any sort of change to the language.
Oh, and Indonesian is my favorite lingua franca. Super easy to learn for anyone and also simple, phonetic spelling.
I had only ever learned Indo-European languages (ie English, Spanish, French) and a bit of Japanese (also unrelated to Indonesian), but I was able to pick up a useful amount of conversational Indonesian in about 3-4 weeks. Indonesian is an Austronesian language (actually a standardized variant of Malay) and totally unrelated to my mother tongue (American English), yet it was the easiest thing to pick up. Sounding out new words in Indonesian is actually easier than English to me.
The e in many words isn't silent, it's a modifier. Kit and kite are different words. "th", "z", "w", and "v" are all really different sounds. While certain accents (or children) do occasionally conflate them, in basically every case, it's technically incorrect. Zebra, Webra, Vebra, and Thebra are all totally different words to a native English speaker.
I'm absolutely behind the idea of simplified English where spelling and pronunciation match. But that's a lofty goal, as first one would have to canonize English, which is basically impossible at this point. Then they'd have to tackle homonyms, like cot and caught (assuming canonized English has these pronounced the same).
The real low-hanging fruit is getting the British to give up on those spellings that follow a dead branch of French ;-)
Another easy example is just fixing "island". The <s> was never pronounced. Medieval scribes put it there because they incorrectly guessed that the word was related to Latin "insula".
For most non-native speakers, the backtracking of silent `e` is more confusing. It does not help in case of `sake` vs `saké` where most people do not add the acute mark and use context for disambiguation.
But really, it should be kite because that's what it is :)
x can then be used as "sh", as in Portuguese.
This thread finally made me remember to go look it up and it seems like the "ç" used to be a different sound (/dz/). I guess it evolved to the "s" sound we hear today sometime by the 1700s.
I wonder if that means only words older than the 1700s have the cedilha and newer words would just be spelled with an "s"?
Sometimes when I'm in parts of the world where I'm surrounded by other languages I don't speak, I find I'm almost automatically tuning out people speaking them around me. That doesn't happen at all with Dutch.
z, c (as in "ce", "ci"): use "s" (non european spanish speakers do not distinguish these sounds anyway)
v: always use "b"
c (as in "ca", "co", "cu"), q(u) (as in "que", "quiso"): replaced with "k"
w: why do we have this letter?! use "u"
y (as vowel): use "i" (basically only used as "and" in Spanish)
y (as consonant): stays like it is now (important in some variants where it sounds pretty much as "sh" in English)
ll as in "lluvia": replaced with "y"
h (mute as in "hueso", "humano"): Just remove it (ueso, umano)
ch (as in "chorizo"): replaced with "c"
r, rr: Couldn't yet find a good replacement that's not ambiguous for the soft and vibrant sounds in all the use-cases...
ñ: this stays. it gives the language personality!
I've got not much traction with my friend, though!!!!!
We can remove it and call the entire transition the Convergencia año-ano.
I stand by the Ñ!
As far as I know, when properly pronounced, the V in Villa doesn't sound the same as the B in Billete.
Sure, sometimes they blend into each other, but not always.
https://web.archive.org/web/19991006200917/http://users.ox.a...
(We detached this subthread from https://news.ycombinator.com/item?id=26809658.)