Home / Typography & Unicode Tools / UTF-16 Code Unit Inspector
Typography & Unicode Tools

UTF-16 Code Unit Inspector

See code-unit offsets, hexadecimal values, and high/low surrogate roles for supplementary Unicode characters.

Unicode 17 referenceGrapheme ≠ code pointUTF-16 / UTF-8 offsetsLocal inspection
—UTF-16 units
—code points
JavaScript string storage viewSupplementary code points use a high-surrogate + low-surrogate pair in UTF-16.

Code units are storage, not characters

Surrogate pairs are grouped conceptually while every 16-bit unit remains inspectable. Unpaired surrogates are flagged instead of being called valid Unicode scalar values.

Unicode precision boundary

WebToolArc keeps grapheme clusters, Unicode code points, UTF-16 code units and UTF-8 bytes as separate measurements. Normalization follows the browser Unicode implementation; grapheme segmentation uses Intl.Segmenter when available. Property names shown by this tool are limited to the explicit local tables and JavaScript Unicode property escapes used by the page rather than claiming a complete Unicode Character Database name lookup.

Search by task, tool name, or category. Press Esc to close.
Start typing to find a tool.