Reference & Lookup Tools
100% Client-Side 100% Client-Side. Unicode character inspection executes via browser TextEncoder and String.codePointAt.

Unicode Lookup & Character Inspector

Inspect any character or code point: U+XXXX hex, UTF-8 bytes, UTF-16 surrogates, and block names

Unicode 15.1 Character Inspector & Decomposer

Type or paste any characters, glyphs, or emoji to inspect code points, UTF-8 bytes, surrogate pairs, and HTML decimal/hex escapes.

N
U+004EBasic Latin (ASCII)Dec: 78
UTF-8 Bytes:4E
UTF-16 Surrogates:0x004E
HTML Hex:
JS Escape:
e
U+0065Basic Latin (ASCII)Dec: 101
UTF-8 Bytes:65
UTF-16 Surrogates:0x0065
HTML Hex:
JS Escape:
o
U+006FBasic Latin (ASCII)Dec: 111
UTF-8 Bytes:6F
UTF-16 Surrogates:0x006F
HTML Hex:
JS Escape:
k
U+006BBasic Latin (ASCII)Dec: 107
UTF-8 Bytes:6B
UTF-16 Surrogates:0x006B
HTML Hex:
JS Escape:
y
U+0079Basic Latin (ASCII)Dec: 121
UTF-8 Bytes:79
UTF-16 Surrogates:0x0079
HTML Hex:
JS Escape:
r
U+0072Basic Latin (ASCII)Dec: 114
UTF-8 Bytes:72
UTF-16 Surrogates:0x0072
HTML Hex:
JS Escape:
u
U+0075Basic Latin (ASCII)Dec: 117
UTF-8 Bytes:75
UTF-16 Surrogates:0x0075
HTML Hex:
JS Escape:
U+0020Basic Latin (ASCII)Dec: 32
UTF-8 Bytes:20
UTF-16 Surrogates:0x0020
HTML Hex:
JS Escape:
🚀
U+1F680Transport and Map SymbolsDec: 128640
UTF-8 Bytes:F0 9F 9A 80
UTF-16 Surrogates:0xD83D 0xDE80
HTML Hex:
JS Escape:

Popular Unicode Symbols & Quick Lookup

Search arrows, math symbols, stars, currency marks, or enter U+XXXX code points.

About Unicode Lookup & Character Inspector

The Unicode Standard 15.1 character inspector and symbol lookup table. Paste or type any glyph, emoji, or text to decompose it into its exact hexadecimal code point, UTF-8 byte sequence, UTF-16 surrogate pairs, HTML entities, and language escape sequences (\u{...}). Includes a curated searchable table of popular Unicode symbols.

Key Capabilities & Features

  • Real-time Unicode 15.1 character decomposition for single characters, emojis, and strings
  • Decodes formal code point (U+XXXX), UTF-8 bytes, and UTF-16 surrogate code units
  • Generates HTML decimal, HTML hex, and JavaScript/TypeScript string escape syntax
  • Automated Unicode block category classification (Latin, CJK, Dingbats, Math, Arrows, etc.)
  • Pre-indexed quick symbol palette for popular stars, hearts, checks, and arrows

How to Use Unicode Lookup & Character Inspector

1

Enter Character or String

Paste any symbol, emoji, or string into the inspector input field.

2

Inspect Architecture

Review the hex code point, byte sequences, and character block.

3

Search Symbols

Use the symbol search bar to find mathematical marks, stars, or currency symbols.

4

Copy Code Units

Copy JavaScript escape strings (\u{...}) or HTML entities with one click.

Privacy & In-Browser Execution Guarantee

100% Client-Side. Unicode character inspection executes via browser TextEncoder and String.codePointAt.

Frequently Asked Questions

What is a UTF-16 surrogate pair?

Characters outside the Basic Multilingual Plane (code points above U+FFFF, such as modern emojis) cannot fit into a single 16-bit code unit. UTF-16 represents them using two 16-bit code units called a high and low surrogate.

How do code points relate to UTF-8 bytes?

A Unicode code point is an abstract integer (from 0 to 0x10FFFF). UTF-8 is a variable-length binary encoding that serializes that integer into 1, 2, 3, or 4 bytes.