Back to tools
Encodingpublic tool

Unicode Normalizer

Apply NFC, NFD, NFKC, and NFKD to a string, then inspect exactly which code points it contains.

Code points
7
Graphemes
7
UTF-8 size
10 bytes
Forms changed
3 of 4

Input text

7 code point(s), 7 grapheme cluster(s), 10 UTF-8 bytes

NFC

Canonical composition — combines base + combining mark into one code point

Café fin

Changed

No

Code points

7

UTF-8

10 bytes

NFD

Canonical decomposition — splits one code point into base + combining marks

Café fin

Changed

Yes

Code points

8

UTF-8

11 bytes

NFKC

Compatibility composition — also folds fi to fi and ① to 1

Café fin

Changed

Yes

Code points

8

UTF-8

9 bytes

NFKD

Compatibility decomposition — fully expanded, useful for search indexing

Café fin

Changed

Yes

Code points

9

UTF-8

10 bytes

Code point inspector

Every code point is also its own grapheme cluster

Every code point in the input with its category
#CharEscapeCode pointDecimalUTF-8Category guess
0CCU+0043671Basic Multilingual Plane
1aaU+0061971Basic Multilingual Plane
2ffU+00661021Basic Multilingual Plane
3ééU+00E92332Latin-1 Supplement Letter
4␠spaceU+0020321Space
5fifiU+FB01642573Alphabetic Presentation Forms (ligature)
6nnU+006E1101Basic Multilingual Plane

Grapheme vs code point

Why your user sees fewer characters than your string length suggests

Code points

7

Graphemes

7

UTF-16 units

7