Isn't it possible to detect the mixture of cyrillic vs english and so on?
https://en.wikipedia.org/wiki/Cyrillic_(Unicode_block)
From that, you can see domains composed exclusively of those characters with are homoglyphic or nearly so with the ASCII equivalents would be susceptible to this:
AaBCcEeFHhIiJjKMOoPpSsTXxY
That's not all the letters, but there are plenty of domains in English which can be written to look the same with entirely Cyrillic.