Unicode Character Inspector

See each character's Unicode code point, UTF-8 bytes, UTF-16 units and HTML entity.

Runs in your browser Free

What's really in your text?

Sometimes two words look identical but don't match, a form throws a mysterious error or code won't compile. The cause is often invisible characters (zero-width space, non-breaking space), look-alike letters (Latin 'a' vs Cyrillic 'а') or encoding issues.

Shown for every character

  • The character and its category (letter, number, punctuation, symbol, space, control)
  • U+ code point, decimal and hex value
  • UTF-8 byte sequence and UTF-16 code units
  • HTML entity

Ideal for inspecting emoji, debugging encoding errors and analysing suspicious text. Analysis happens in your browser.

How to use Unicode Character Inspector

  1. Paste the text to inspect.
  2. Each character is listed on its own row.
  3. Review the code point, category and encodings.
  4. Find suspicious characters and fix them at the source.

Why use this tool?

Reveals invisible character problems in seconds; invaluable for developers and data analysts.

FAQ

Why does an emoji show two UTF-16 units?

Characters above U+FFFF are stored as surrogate pairs in UTF-16, which is why JavaScript's length counts them as 2.

What is a zero-width space?

An invisible character (U+200B) that takes up a position in text and can be carried along unnoticed by copy and paste.

Can it detect look-alike letters?

Yes. Each character's code point shows whether a letter is, for example, Latin or Cyrillic.

Related tools