What this checker finds
A mixed-script string can be legitimate, but an unexpected Cyrillic or Greek character inside Latin text is a common sign of a look-alike identifier. The checker reports the script for every Unicode code point and flags a small, practical set of characters that visually resemble Latin letters.
- Scripts: Latin, Cyrillic, Greek, Arabic, Hebrew, Han, Japanese, Hangul and other broad ranges are labeled separately. Punctuation, spaces and emoji are treated as common characters.
- Confusables: common Greek, Cyrillic, fullwidth and mathematical look-alikes are called out; this is a helpful review aid, not a complete Unicode Security Mechanism implementation.
- Normalization: NFC preserves canonical equivalence, while NFKC also folds compatibility forms such as fullwidth characters. Review any changed text before using it in an identifier.
No text is uploaded. For high-assurance spoofing review, also apply the Unicode Consortium confusables data and your platform's identifier policy.