Unicode Character Inspector
See each character's Unicode code point, UTF-8 bytes, UTF-16 units and HTML entity.
What's really in your text?
Sometimes two words look identical but don't match, a form throws a mysterious error or code won't compile. The cause is often invisible characters (zero-width space, non-breaking space), look-alike letters (Latin 'a' vs Cyrillic 'а') or encoding issues.
Shown for every character
- The character and its category (letter, number, punctuation, symbol, space, control)
- U+ code point, decimal and hex value
- UTF-8 byte sequence and UTF-16 code units
- HTML entity
Ideal for inspecting emoji, debugging encoding errors and analysing suspicious text. Analysis happens in your browser.
How to use Unicode Character Inspector
- Paste the text to inspect.
- Each character is listed on its own row.
- Review the code point, category and encodings.
- Find suspicious characters and fix them at the source.
Why use this tool?
Reveals invisible character problems in seconds; invaluable for developers and data analysts.
FAQ
Why does an emoji show two UTF-16 units?
Characters above U+FFFF are stored as surrogate pairs in UTF-16, which is why JavaScript's length counts them as 2.
What is a zero-width space?
An invisible character (U+200B) that takes up a position in text and can be carried along unnoticed by copy and paste.
Can it detect look-alike letters?
Yes. Each character's code point shows whether a letter is, for example, Latin or Cyrillic.
Related tools
Unicode Escape / Unescape
Convert text to \uXXXX, \u{…}, Python, CSS, HTML or U+ escape sequences and back.
Open toolASCII Converter and Table
Convert text to decimal, hex, binary or octal ASCII codes and back, with a searchable ASCII table.
Open toolCharacter Frequency Counter
Count how often each character appears in a text, with percentages, a bar chart and CSV export.
Open toolHTML Entity Encoder / Decoder
Encode special characters as HTML entities (< → <) or decode entities back to text.
Open tool