convertCASEpro

MODE

Unicode Text Converter

See the code point behind every character, and decode escapes back to text.

INPUT
CHARS: 0WORDS: 0LINES: 0
OUTPUT
CHARS: 0WORDS: 0LINES: 0

How the Unicode Text Converter works

Every character in Unicode has a number called a code point. This tool writes those numbers in the formats developers use: U+0041, JavaScript style backslash u sequences (with surrogate pairs for characters above U+FFFF), HTML hexadecimal and decimal entities and CSS escapes.

Switch to the decode direction to turn those sequences back into text. The keep printable ASCII option leaves normal letters alone, so only special characters are escaped.

When it is useful

  • •Putting special characters in JavaScript, HTML or CSS source.
  • •Finding out what an unfamiliar or invisible character is.
  • •Debugging text that looks identical but is not equal.
  • •Writing source that must contain only ASCII.

Frequently Asked Questions

What is a code point?

The number Unicode assigns to a character, written as U+ followed by hex digits. Letter A is U+0041 and the grinning face emoji is U+1F600.

Why do emoji become two escapes in JavaScript?

JavaScript strings use 16-bit units, so characters above U+FFFF are stored as a surrogate pair of two escapes.

How is this different from UTF-8 encoding?

Code points are the abstract numbers. UTF-8 is one way to store them as bytes. See the UTF-8 tool for bytes.

Related tools

[ Browse all tools ]