Home / Typography & Unicode Tools / Unicode Combining Mark Inspector
Typography & Unicode Tools

Unicode Combining Mark Inspector

Inspect combining marks by code point and UTF-16 position without assuming every mark is a removable accent.

Unicode 17 referenceGrapheme ≠ code pointUTF-16 / UTF-8 offsetsLocal inspection
0Matches Unicode Mark-category code points. A mark is not automatically a disposable accent.
Code-point indexUTF-16 offsetCode pointType

NFC/NFD comparison with combining-mark trace

Compare normalized forms, locate combining marks and inspect each grapheme's code-point sequence before deciding which normalization to copy.

—Combining marks
—NFC code points
—NFD code points
Ready.
ClusterCode pointsMarks

One visible character can be base plus marks

Accents and other combining marks may follow a base letter as separate code points. Normalized text can represent an equivalent-looking character differently, which matters for byte comparison and identifiers. Inspect the sequence before assuming duplicate-looking text has identical underlying characters.

Mark category is not the same as removable accent

Combining marks are reported with offsets and whether they have the Unicode Diacritic property. A Mark-category character is never automatically treated as disposable.

Unicode precision boundary

WebToolArc keeps grapheme clusters, Unicode code points, UTF-16 code units and UTF-8 bytes as separate measurements. Normalization follows the browser Unicode implementation; grapheme segmentation uses Intl.Segmenter when available. Property names shown by this tool are limited to the explicit local tables and JavaScript Unicode property escapes used by the page rather than claiming a complete Unicode Character Database name lookup.

Search by task, tool name, or category. Press Esc to close.
Start typing to find a tool.