This almost looks like something a speaker of a romance language would put together.
My data point is Bulgarian where you generally can't omit "to be". This may have to do with the fact that Bulgarian has mostly lost its case system?
Slovenian:
Jaz sem Rus
Ti si Rus
Croatian: Ja sam Rus
Ti si Rus
Slovak: Som Rus
Ty si Rus
Whereas in Polish is a joint word (but the to be is still merged in there): Jestem z Rosji
Jesteś Rosjaninem
I also thought that using the Latin alphabet feels a bit of a hack and Cyrillic would be a more "natural" way to put things. However, using Cyrillic would impose a higher learning curve for those who do not it.
This is mostly for phonemes that don't have a 1-to-1 representation in the Latin alphabet, like ш, щ, ж and what not. For these cases, languages like Croatian, Slovenian and others have come up with, what is my personal opinion, are quite acceptable workarounds (I just found out that this is called Gaj's Latin alphabet [1])
č, ć, dž, đ, š, ž (I skipped a few).
Without knowing initially how they are pronounced, read in context you do get an idea of what the target sound is.I also think that on-boarding Cyrillic-first cultures would be easier because normally countries that use Cyrillic as their main alphabet also teach the Latin alphabet at school (Russia for example does and they also teach you how to write "Russian" using Latin characters because they understand the importance of this). Whereas it's not the same way around: a lot of the Slavic countries do not teach the Cyrillic alphabet at school.
> I've always had trouble trying to parse out Russian transliterated into Latin script. I also think this is because when you're learning Russian as a non-native speaker, the courses focus mainly on Cyrillic (naturally so) so you never really learn that ж -> zsh
ш -> sh
щ -> sch
х -> kh
ь -> '
etc.
[1]: https://en.wikipedia.org/wiki/Gaj's_Latin_alphabet
Edit: answer question about Russian transliteration, add missing link and try to fix formatting.
There’s a number of romanization standards for Cyrillic, and which one is the most intuitive might be language-dependent (e.g. what Russian writes as a stressed ‹и› Ukrainian writes as ‹і›, which post-1918 Russian doesn’t use; what Ukrainian writes as ‹и› is closer but not identical to what Russian denotes ‹ы›, which Ukrainian doesn’t use).
Wikipedia has a good summary table[1] of some of the formal standards, but these don’t cover some of the vernacular usage, so to say. The most frequent quirk is probably writing ‹щ› as ‹sch›, as you do, when standards insist on ‹shch› (my guess is because the Belarusian counterpart of ‹щ› is ‹шч›, which would also be written ‹shch›); most confusing is perhaps writing the masculine adjectival ending ‹-ий›, ‹-ый› as ‹-y› or rarely ‹-yy› (writing ‹Navalny› for the surname ‹Навальный› or ‹Zelenskyy› for the surname ‹Зеленський›, maybe because old 19th-century German-inspired romanizations usually wrote it as ‹-i›, merging the ending with its Polish counterpart). I have to say I’ve never seen ‹zsh› instead for ‹zh› for ‹ж›, though.
Empirically, I’ve also found that French people struggle with pronouncing an English-inspired transliteration (‹sh› for ‹ш›, ‹ch› for ‹ч›), but have no problem with finding at least a pronounceable fallback using a Czech/Serbo-Croatian-inspired one (‹š› for ‹ш›, ‹č› for ‹ч›), so diacritics might indeed be underrated here.
[1] https://en.wikipedia.org/wiki/Romanization_of_Russian#Transl...
⸻
1. I suspect the other reason was that implementing the tie was technically simple since, by allowing it to extend pass its spacing width, it can be treated like any other above-character accent. Ogonek, on the other hand, unlike a cedilla or the dot-under diacritic, requires positioning based on the letter that it's attached to so can't be programmed as easily as a floating diacritic mark.
Note that in modern times, you’re supposed to use the tie to spell affricates and such in IPA as well, like in ‹t͡ʃ› and ‹d͡ʒ› and, yes, ‹t͡s›, even though nobody does as far as I’ve seen.
[1] https://en.wikipedia.org/wiki/ALA-LC_romanization_for_Russia...
Jestem z Polski == I am from Poland
Jesteś z Polski == You are from Poland
Jestem Polakiem == I am Polish
Jesteś Polakiem == You are Polish
Both have very close meaning but 3&4 seem to be equivalent of examples in other languages given above.Source: I am Polish
(South) Slavic speaker here. Cyrillic would not be much more natural, we haven't used Cyrillic in over 30 years. Latin would be something all of us Slavenes would be able to read.
I appreciate the correction. Definitely didn't expect so many replies, but I'm learning a lot.
i am not sure what you mean but russian is by far the biggest slavic language in terms of number of speakers. i also think it is the biggest in terms of number of litterary and academic works
Something like Slovakian would be much better as a base.
From my west-slavic (Polish) perspective I can understand Czech, Slovakian (a bit harder) and Ukrainian semi-ok. But Russian is way, way harder (there are some similar words but most are alien to me).
I didn't have chance to listen to south slavic languages so I don't know how similar they are to west slavic ones.
If you mean Serbian/Croatian/Bosnian/Montenegrin - they are only recognized as languages for political reasons. They actually are a particularly rich continuum of dialects but it is impractical to recognize this because it would imply there should be one literary norm.
in south slavic family there is a one-to-one mapping between cyrilic and latin scripts. in serbia, for example, both scripts are official
Also I do not believe that there can be a better choice than "es", because I suppose that in order to simplify the grammar they have removed the personal terminations from the verbs, in which case "est" of the 3rd person singular becomes "es", just the simple stem of the verb.
Also, even if their examples show an explicit "es", optionally omitting redundant words is a trivial rule to add to any language grammar.
It is likely that there are other aspects of the language that can be criticized, but these 3 points seem OK.