How to use the Unicode Inspector.
Inspect code points when text looks identical but behaves differently. Normalize only when the destination's rules allow it.
Make the workflow fit your task.
Inspect the actual characters, listing code points and positions rather than relying on appearance. Identify combining marks, lookalike symbols and invisible characters. Compare normalized forms only under an agreed rule and retain originals when exact identifiers or bytes matter.
- What you provide
- Text with unexpected characters.
- What you get
- Code points, invisible characters and normalization suggestions.
See the input and the result.
Illustrative input and output · a teaching example, not a live WebAct run
Example input
Inspect these two strings: café and café. Compare their code points and their NFC-normalized values.
Completed example
café: U+0063 U+0061 U+0066 U+00E9 (4 code points). café: U+0063 U+0061 U+0066 U+0065 U+0301 (5 code points). Both normalize to café under NFC. Their original sequences differ; preserve originals when exact byte identity matters.
Load this input into the prompt, then copy it to WebAct to try the task. Your result may differ from the illustration.
Decisions and troubleshooting.
Can visually identical Unicode strings have different lengths?
Yes. A visible character can be represented by one code point or several, and byte length adds another distinct measure.
Why did normalization fail to make two lookalike letters equal?
They may be different characters from different scripts rather than alternative representations of the same character. Inspect their code points.
Reference for this workflow.
Try it with your own source.
Replace the example with your material in the task prompt. Keep the requirements you need, then copy the task into WebAct.
Customize and copy the task ↑