Which character should represent the apostrophe? The Unicode committee is wrong
tedclancy.wordpress.com
tedclancy.wordpress.com
It's enough to drive you crazy.
The author's point makes sense to me. I agree and my initial thought is that the correct thing to do is recommend U+02BC be the apostrophe.
This is something for the future. It won't be faithfully reflected in text for years or decades.
On the other hand, I've had to track down a ton of mysterious bugs caused by software "helpfully" converting characters out of the standard ASCII set, unbeknownst to the user (and they look basically the same)... only a hexdump shows the truth.
I for one was quite surprised about that. When doing word segmentation for NLP it is common to split these contractions.
Plenty of people argue that it is a single word because there are spaces around it or simply because it is a contraction. Other argue it is 2 words since they are of different word classes.
This article is about the argument that Unicode Committee recommendation that U+2019 (RIGHT SINGLE QUOTATION MARK) is the preferred character for apostrophe in English text is wrong (since, among other things, it breaks detecting matched pairs of quotation marks), and that the preferred character for that use should be U+02BC (APOSTROPHE, MODIFIER LETTER).
That's not at all true; while "guess from context" is one option, using modifier keys is also an option; there's a lot of software that does this for things which don't have distinct keys. Certainly seen this used for different widths of spaces, and different widths of dashes, even though the keyboard has only the space bar and the hyphen.
Fortunately, it does seem to have changed over the years (although I'm not sure to what extent things like auto-correct/auto-capitalisation contribute), so perhaps in the far future when keyboard layouts evolve enough, everyone will be using Unicode apostrophes and single quotes...