One visible character can be base plus marks
Accents and other combining marks may follow a base letter as separate code points. Normalized text can represent an equivalent-looking character differently, which matters for byte comparison and identifiers. Inspect the sequence before assuming duplicate-looking text has identical underlying characters.
Mark category is not the same as removable accent
Combining marks are reported with offsets and whether they have the Unicode Diacritic property. A Mark-category character is never automatically treated as disposable.
Unicode precision boundary
WebToolArc keeps grapheme clusters, Unicode code points, UTF-16 code units and UTF-8 bytes as separate measurements. Normalization follows the browser Unicode implementation; grapheme segmentation uses Intl.Segmenter when available. Property names shown by this tool are limited to the explicit local tables and JavaScript Unicode property escapes used by the page rather than claiming a complete Unicode Character Database name lookup.