Skip to main content

Unicode Character Table & Lookup

Search the Unicode standard by character, name, or code point, or explore characters organized by block, script, category, and standard version.

Unicode Planes & Complete Block Directory

346 Blocks Total

The Unicode code space is divided into 17 planes of 65,536 code points each. Explore all allocated blocks below.

Plane 0: Basic Multilingual Plane (BMP)

Contains almost all modern writing systems, basic symbols, and punctuation.

Range: U+0000–U+FFFF 164 Blocks 16,969 Assigned
Basic Latin
U+0000–U+007F 127 / 128
Latin-1 Supplement
U+0080–U+00FF 128 / 128
Latin Extended-A
U+0100–U+017F 128 / 128
Latin Extended-B
U+0180–U+024F 208 / 208
IPA Extensions
U+0250–U+02AF 96 / 96
Spacing Modifier Letters
U+02B0–U+02FF 80 / 80
Combining Diacritical Marks
U+0300–U+036F 112 / 112
Greek and Coptic
U+0370–U+03FF 135 / 144
Cyrillic
U+0400–U+04FF 256 / 256
Cyrillic Supplement
U+0500–U+052F 48 / 48
Armenian
U+0530–U+058F 91 / 96
Hebrew
U+0590–U+05FF 88 / 112
Arabic
U+0600–U+06FF 256 / 256
Syriac
U+0700–U+074F 77 / 80
Arabic Supplement
U+0750–U+077F 48 / 48
Thaana
U+0780–U+07BF 50 / 64
NKo
U+07C0–U+07FF 62 / 64
Samaritan
U+0800–U+083F 61 / 64
Mandaic
U+0840–U+085F 29 / 32
Syriac Supplement
U+0860–U+086F 11 / 16
Arabic Extended-B
U+0870–U+089F 43 / 48
Arabic Extended-A
U+08A0–U+08FF 96 / 96
Devanagari
U+0900–U+097F 128 / 128
Bengali
U+0980–U+09FF 96 / 128
Gurmukhi
U+0A00–U+0A7F 80 / 128
Gujarati
U+0A80–U+0AFF 91 / 128
Oriya
U+0B00–U+0B7F 91 / 128
Tamil
U+0B80–U+0BFF 72 / 128
Telugu
U+0C00–U+0C7F 101 / 128
Kannada
U+0C80–U+0CFF 92 / 128
Malayalam
U+0D00–U+0D7F 118 / 128
Sinhala
U+0D80–U+0DFF 91 / 128
Thai
U+0E00–U+0E7F 87 / 128
Lao
U+0E80–U+0EFF 83 / 128
Tibetan
U+0F00–U+0FFF 211 / 256
Myanmar
U+1000–U+109F 160 / 160
Georgian
U+10A0–U+10FF 88 / 96
Hangul Jamo
U+1100–U+11FF 256 / 256
Ethiopic
U+1200–U+137F 358 / 384
Ethiopic Supplement
U+1380–U+139F 26 / 32
Cherokee
U+13A0–U+13FF 92 / 96
Unified Canadian Aboriginal Syllabics
U+1400–U+167F 640 / 640
Ogham
U+1680–U+169F 29 / 32
Runic
U+16A0–U+16FF 89 / 96
Tagalog
U+1700–U+171F 23 / 32
Hanunoo
U+1720–U+173F 23 / 32
Buhid
U+1740–U+175F 20 / 32
Tagbanwa
U+1760–U+177F 18 / 32
Khmer
U+1780–U+17FF 114 / 128
Mongolian
U+1800–U+18AF 158 / 176
Unified Canadian Aboriginal Syllabics Extended
U+18B0–U+18FF 70 / 80
Limbu
U+1900–U+194F 68 / 80
Tai Le
U+1950–U+197F 35 / 48
New Tai Lue
U+1980–U+19DF 83 / 96
Khmer Symbols
U+19E0–U+19FF 32 / 32
Buginese
U+1A00–U+1A1F 30 / 32
Tai Tham
U+1A20–U+1AAF 127 / 144
Combining Diacritical Marks Extended
U+1AB0–U+1AFF 58 / 80
Balinese
U+1B00–U+1B7F 127 / 128
Sundanese
U+1B80–U+1BBF 64 / 64
Batak
U+1BC0–U+1BFF 56 / 64
Lepcha
U+1C00–U+1C4F 74 / 80
Ol Chiki
U+1C50–U+1C7F 48 / 48
Cyrillic Extended-C
U+1C80–U+1C8F 11 / 16
Georgian Extended
U+1C90–U+1CBF 46 / 48
Sundanese Supplement
U+1CC0–U+1CCF 8 / 16
Vedic Extensions
U+1CD0–U+1CFF 43 / 48
Phonetic Extensions
U+1D00–U+1D7F 128 / 128
Phonetic Extensions Supplement
U+1D80–U+1DBF 64 / 64
Combining Diacritical Marks Supplement
U+1DC0–U+1DFF 64 / 64
Latin Extended Additional
U+1E00–U+1EFF 256 / 256
Greek Extended
U+1F00–U+1FFF 233 / 256
General Punctuation
U+2000–U+206F 111 / 112
Superscripts and Subscripts
U+2070–U+209F 42 / 48
Currency Symbols
U+20A0–U+20CF 34 / 48
Combining Diacritical Marks for Symbols
U+20D0–U+20FF 33 / 48
Letterlike Symbols
U+2100–U+214F 80 / 80
Number Forms
U+2150–U+218F 60 / 64
Arrows
U+2190–U+21FF 112 / 112
Mathematical Operators
U+2200–U+22FF 256 / 256
Miscellaneous Technical
U+2300–U+23FF 256 / 256
Control Pictures
U+2400–U+243F 42 / 64
Optical Character Recognition
U+2440–U+245F 11 / 32
Enclosed Alphanumerics
U+2460–U+24FF 160 / 160
Box Drawing
U+2500–U+257F 128 / 128
Block Elements
U+2580–U+259F 32 / 32
Geometric Shapes
U+25A0–U+25FF 96 / 96
Miscellaneous Symbols
U+2600–U+26FF 256 / 256
Dingbats
U+2700–U+27BF 192 / 192
Miscellaneous Mathematical Symbols-A
U+27C0–U+27EF 48 / 48
Supplemental Arrows-A
U+27F0–U+27FF 16 / 16
Braille Patterns
U+2800–U+28FF 256 / 256
Supplemental Arrows-B
U+2900–U+297F 128 / 128
Miscellaneous Mathematical Symbols-B
U+2980–U+29FF 128 / 128
Supplemental Mathematical Operators
U+2A00–U+2AFF 256 / 256
Miscellaneous Symbols and Arrows
U+2B00–U+2BFF 254 / 256
Glagolitic
U+2C00–U+2C5F 96 / 96
Latin Extended-C
U+2C60–U+2C7F 32 / 32
Coptic
U+2C80–U+2CFF 123 / 128
Georgian Supplement
U+2D00–U+2D2F 40 / 48
Tifinagh
U+2D30–U+2D7F 59 / 80
Ethiopic Extended
U+2D80–U+2DDF 79 / 96
Cyrillic Extended-A
U+2DE0–U+2DFF 32 / 32
Supplemental Punctuation
U+2E00–U+2E7F 94 / 128
CJK Radicals Supplement
U+2E80–U+2EFF 115 / 128
Kangxi Radicals
U+2F00–U+2FDF 214 / 224
Ideographic Description Characters
U+2FF0–U+2FFF 16 / 16
CJK Symbols and Punctuation
U+3000–U+303F 64 / 64
Hiragana
U+3040–U+309F 93 / 96
Katakana
U+30A0–U+30FF 96 / 96
Bopomofo
U+3100–U+312F 43 / 48
Hangul Compatibility Jamo
U+3130–U+318F 94 / 96
Kanbun
U+3190–U+319F 16 / 16
Bopomofo Extended
U+31A0–U+31BF 32 / 32
CJK Strokes
U+31C0–U+31EF 39 / 48
Katakana Phonetic Extensions
U+31F0–U+31FF 16 / 16
Enclosed CJK Letters and Months
U+3200–U+32FF 255 / 256
CJK Compatibility
U+3300–U+33FF 256 / 256
CJK Unified Ideographs Extension A
U+3400–U+4DBF 2 / 6,592
Yijing Hexagram Symbols
U+4DC0–U+4DFF 64 / 64
CJK Unified Ideographs
U+4E00–U+9FFF 2 / 20,992
Yi Syllables
U+A000–U+A48F 1,165 / 1,168
Yi Radicals
U+A490–U+A4CF 55 / 64
Lisu
U+A4D0–U+A4FF 48 / 48
Vai
U+A500–U+A63F 300 / 320
Cyrillic Extended-B
U+A640–U+A69F 96 / 96
Bamum
U+A6A0–U+A6FF 88 / 96
Modifier Tone Letters
U+A700–U+A71F 32 / 32
Latin Extended-D
U+A720–U+A7FF 204 / 224
Syloti Nagri
U+A800–U+A82F 45 / 48
Common Indic Number Forms
U+A830–U+A83F 10 / 16
Phags-pa
U+A840–U+A87F 56 / 64
Saurashtra
U+A880–U+A8DF 82 / 96
Devanagari Extended
U+A8E0–U+A8FF 32 / 32
Kayah Li
U+A900–U+A92F 48 / 48
Rejang
U+A930–U+A95F 37 / 48
Hangul Jamo Extended-A
U+A960–U+A97F 29 / 32
Javanese
U+A980–U+A9DF 91 / 96
Myanmar Extended-B
U+A9E0–U+A9FF 31 / 32
Cham
U+AA00–U+AA5F 83 / 96
Myanmar Extended-A
U+AA60–U+AA7F 32 / 32
Tai Viet
U+AA80–U+AADF 72 / 96
Meetei Mayek Extensions
U+AAE0–U+AAFF 23 / 32
Ethiopic Extended-A
U+AB00–U+AB2F 32 / 48
Latin Extended-E
U+AB30–U+AB6F 60 / 64
Cherokee Supplement
U+AB70–U+ABBF 80 / 80
Meetei Mayek
U+ABC0–U+ABFF 56 / 64
Hangul Syllables
U+AC00–U+D7AF 2 / 11,184
Hangul Jamo Extended-B
U+D7B0–U+D7FF 72 / 80
High Surrogates
U+D800–U+DB7F 0 / 896
High Private Use Surrogates
U+DB80–U+DBFF 0 / 128
Low Surrogates
U+DC00–U+DFFF 0 / 1,024
Private Use Area
U+E000–U+F8FF 2 / 6,400
CJK Compatibility Ideographs
U+F900–U+FAFF 472 / 512
Alphabetic Presentation Forms
U+FB00–U+FB4F 58 / 80
Arabic Presentation Forms-A
U+FB50–U+FDFF 656 / 688
Variation Selectors
U+FE00–U+FE0F 16 / 16
Vertical Forms
U+FE10–U+FE1F 10 / 16
Combining Half Marks
U+FE20–U+FE2F 16 / 16
CJK Compatibility Forms
U+FE30–U+FE4F 32 / 32
Small Form Variants
U+FE50–U+FE6F 26 / 32
Arabic Presentation Forms-B
U+FE70–U+FEFF 141 / 144
Halfwidth and Fullwidth Forms
U+FF00–U+FFEF 225 / 240
Specials
U+FFF0–U+FFFF 5 / 16

Unicode Scripts Directory

174 Writing Systems

Unicode assigns characters to scripts based on the writing system they belong to.

Adlam Ahom Anatolian_Hieroglyphs Arabic Armenian Avestan Balinese Bamum Bassa_Vah Batak Bengali Beria_Erfe Bhaiksuki Bopomofo Brahmi Braille Buginese Buhid Canadian_Aboriginal Carian Caucasian_Albanian Chakma Cham Cherokee Chorasmian Common Coptic Cuneiform Cypriot Cypro_Minoan Cyrillic Deseret Devanagari Dives_Akuru Dogra Duployan Egyptian_Hieroglyphs Elbasan Elymaic Ethiopic Garay Georgian Glagolitic Gothic Grantha Greek Gujarati Gunjala_Gondi Gurmukhi Gurung_Khema Han Hangul Hanifi_Rohingya Hanunoo Hatran Hebrew Hiragana Imperial_Aramaic Inherited Inscriptional_Pahlavi Inscriptional_Parthian Javanese Kaithi Kannada Katakana Kawi Kayah_Li Kharoshthi Khitan_Small_Script Khmer Khojki Khudawadi Kirat_Rai Lao Latin Lepcha Limbu Linear_A Linear_B Lisu Lycian Lydian Mahajani Makasar Malayalam Mandaic Manichaean Marchen Masaram_Gondi Medefaidrin Meetei_Mayek Mende_Kikakui Meroitic_Cursive Meroitic_Hieroglyphs Miao Modi Mongolian Mro Multani Myanmar Nabataean Nag_Mundari Nandinagari New_Tai_Lue Newa Nko Nushu Nyiakeng_Puachue_Hmong Ogham Ol_Chiki Ol_Onal Old_Hungarian Old_Italic Old_North_Arabian Old_Permic Old_Persian Old_Sogdian Old_South_Arabian Old_Turkic Old_Uyghur Oriya Osage Osmanya Pahawh_Hmong Palmyrene Pau_Cin_Hau Phags_Pa Phoenician Psalter_Pahlavi Rejang Runic Samaritan Saurashtra Sharada Shavian Siddham Sidetic SignWriting Sinhala Sogdian Sora_Sompeng Soyombo Sundanese Sunuwar Syloti_Nagri Syriac Tagalog Tagbanwa Tai_Le Tai_Tham Tai_Viet Tai_Yo Takri Tamil Tangsa Tangut Telugu Thaana Thai Tibetan Tifinagh Tirhuta Todhri Tolong_Siki Toto Tulu_Tigalari Ugaritic Vai Vithkuqi Wancho Warang_Citi Yezidi Yi Zanabazar_Square

Unicode General Categories

7 Primary Classes • 33 Subcategories

Every Unicode character has an authoritative General_Category property designating its primary grammatical function.

Official Specification Released September 2025

Unicode Standard Version 17.0

The latest Unicode specification defines characters across modern writing systems, historical scripts, emojis, and mathematical symbols.

📊 40,568 Total Encoded ✨ 463 in Version 17.0
Explore Version 17.0 Characters → View All Specification Versions

Essential Unicode Characters

Quick reference and copy for widely used Unicode characters. Click any character to copy to clipboard.

Understanding Unicode Architecture

What is a Unicode Block?

A Unicode Block is a contiguous range of code points allocated for administrative organization and publishing. Blocks are defined by fixed starting and ending hexadecimal boundaries (e.g. U+0000 to U+007F for Basic Latin).

What is a Unicode Script?

A Unicode Script is an intrinsic linguistic property identifying the writing system a character belongs to. Unlike blocks, scripts are not restricted to contiguous ranges; a single script like Latin spans across dozens of different blocks.

What is a General Category?

General_Category is a fundamental character classification grouping code points into major types: Letters (L), Marks (M), Numbers (N), Punctuation (P), Symbols (S), Separators (Z), and Other (C).

Character properties, names, and allocations are dynamically derived from the official Unicode Character Database (UCD) 17.0 standard.
View Data Sources & Methodology →