Homograph detection isn't about prohibiting mixed characters; it's about considering letters that look the same to be the same when checking for existing duplicates.
It would absolutely have to be a pretty labor-intensive, manually-maintained database.