Unicode character units
Distinguish visible graphemes, Unicode code points, UTF-16 units and UTF-8 bytes.
Unicode character units: inputs, results and limitations
Distinguish visible graphemes, Unicode code points, UTF-16 units and UTF-8 bytes.
How to use Unicode character units
- Paste your text or use the example to inspect the operation.
- Run the tool once the input is ready.
- Inspect the preview or result, then download a new file.
An illustrative use case
An emoji and Arabic message compared across graphemes, code points and code units.
Sample input
Hello 👨👩👧👦 مرحبا
What to inspect in the output
Use the count that matches the destination system. A visible emoji can use multiple encoded units, so a platform limit may disagree with a visual-character count.
Where this tool stops
Review the output before using it. Your original input is not changed.
Troubleshooting Unicode character units
Check the generated markup in the published HTML head, not merely in a text editor. Replace example URLs with your real canonical pages and publicly reachable images. Use live inspection tools after deployment.
Questions about Unicode character units
Why are the character counts different?
They measure different layers: visible graphemes, Unicode code points and JavaScript UTF-16 units are not interchangeable.
What input does Unicode character units need?
Use the sample text shown below, then replace it with your own input. Review the result for the destination where it will be used.
Which options does Unicode character units provide?
There are no extra settings in this version. Select or paste the appropriate input and review the produced result.
Your original input is preserved. Review the new output before using it in a published page, customer delivery or production workflow.
JavaScript enables the interactive workspace.