Script

Unicode Version 18.0

Script is a character property that assigns each encoded character to a specific writing system, such as Latin, Greek, Cyrillic, Arabic, or Han, based on its historical and linguistic origin. This categorization helps software render text correctly by selecting appropriate fonts and shaping rules, and it enables text processing tasks like language detection, sorting, and searching. Unlike simple block assignments, which group characters by code point ranges, script reflects actual usage, so a single block may contain characters from multiple scripts, and characters like punctuation or digits are often classified as "Common" or "Inherited" since they are shared across systems. The value is a stable identifier that aids in internationalization and accessibility, though a single character can only belong to one script, while text as a whole can mix many scripts.

Adlam
Abbr
Adlm
Count
88

Adlam is a script used for writing the Fulani language, encoded with its own character block for digital text.

Afaka
Abbr
Afak
Count
0

Afaka is a script used for writing the Aukan language, also known as Ndyuka, a creole spoken in Suriname.

Caucasian Albanian
Abbr
Aghb
Count
53

Caucasian Albanian is a script used for an ancient Christian language in the Caucasus, encoded for digital text.

Ahom
Abbr
Ahom
Count
65

Ahom is a historical script used for the Tai language of the Ahom people, encoded for scholarly and digital preservation.

Arabic
Abbr
Arab
Count
1452

Arabic is a script used for writing Arabic, Persian, Urdu, and several other languages, with cursive letterforms and right-to-left direction.

Arabic (Nastaliq variant)
Abbr
Aran
Count
0

Arabic (Nastaliq variant) is a script value used for languages like Urdu and Punjabi, characterized by its flowing, slanted calligraphic style.

Imperial Aramaic
Abbr
Armi
Count
31

Imperial Aramaic is a script value used for writing ancient Aramaic from the 8th to 3rd centuries BCE, including in Persian Empire records.

Armenian
Abbr
Armn
Count
99

Armenian is a script value used to identify text written in the Armenian alphabet, supporting both classical and modern forms.

Avestan
Abbr
Avst
Count
61

Avestan is the script used for writing the sacred texts of Zoroastrianism, originating in ancient Iran.

Balinese
Abbr
Bali
Count
127

Balinese is a script used for writing the Balinese language, with its Unicode value covering consonants, vowels, and diacritics.

Bamum
Abbr
Bamu
Count
657

Bamum is a script used for the Bamum language and kingdom in Cameroon, with historical and modern variants.

Bassa Vah
Abbr
Bass
Count
36

Bassa Vah is a script used for writing the Bassa language, with its Unicode characters encoded in the range U+16AD0 to U+16AFF.

Batak
Abbr
Batk
Count
56

Batak is a script used for writing the Batak languages of Sumatra, with characters that include syllabic signs, vowels, and punctuation marks.

Bengali
Abbr
Beng
Count
98

Bengali is a writing system used for the Assamese and Bengali languages, with distinct letterforms and conjunct ligatures.

Beria Erfe
Abbr
Berf
Count
50

Beria Erfe is a script value used for writing the Beria language, primarily in Chad and Sudan.

Bhaiksuki
Abbr
Bhks
Count
97

Bhaiksuki is a historical script used for writing Buddhist texts, primarily in Sanskrit, from around the 11th to 12th centuries in northeastern India.

Blissymbols
Abbr
Blis
Count
0

Blissymbols is a constructed ideographic writing system used for augmentative and alternative communication, with its own script category in Unicode.

Bopomofo
Abbr
Bopo
Count
77

Bopomofo is a script value representing the phonetic symbols used in Taiwanese Mandarin education, covering characters like ㄅ to ㄦ.

Brahmi
Abbr
Brah
Count
115

Brahmi is an ancient writing system of South Asia, historically used for Sanskrit and Prakrit, and encoded for digital text.

Braille
Abbr
Brai
Count
256

Braille is a script used to encode tactile writing systems, with Unicode codepoints for all six and eight dot patterns.

Buginese
Abbr
Bugi
Count
30

Buginese is an abugida script used for writing the Bugis language, primarily in South Sulawesi, Indonesia.

Buhid
Abbr
Buhd
Count
20

Buhid is a script used for writing the Buhid language, with its characters encoded in Unicode for digital text representation.

Chakma
Abbr
Cakm
Count
71

Chakma is a Brahmic script used for the Chakma language, featuring unique vowel and consonant signs, with a distinct set of digits.

Canadian Aboriginal
Abbr
Cans
Count
726

Canadian Aboriginal is an assigned script value indicating text written in Canadian Indigenous syllabics, used for languages like Cree and Inuktitut.

Carian
Abbr
Cari
Count
49

Carian is a script used for the ancient Carian language of Anatolia, written right to left with a distinct alphabet.

Cham
Abbr
Cham
Count
83

Cham is a script used for writing the Eastern Cham language, primarily in Vietnam and Cambodia.

Cherokee
Abbr
Cher
Count
172

Cherokee is a North American syllabary script used for writing the Cherokee language, with distinct characters for syllables and a unique visual style.

Chisoi
Abbr
Chis
Count
0

Chisoi is a script value denoting the historical Chisoi writing system, used for inscriptions in ancient Central Asia.

Chorasmian
Abbr
Chrs
Count
28

Chorasmian is a script used for the extinct Eastern Iranian language of the same name, encoded for historic texts.

Cirth
Abbr
Cirt
Count
0

Cirth is a fictional runic script from J.R.R. Tolkien's legendarium, used for writing Dwarvish and other Middle-earth languages.

Coptic
Abbr
Copt
Count
137

Coptic is a historical script used for the Egyptian Coptic language, derived from Greek with added letters for native sounds.

Cypro Minoan
Abbr
Cpmn
Count
99

Cypro Minoan is a script value used for the undeciphered writing system of Late Bronze Age Cyprus, encoded for ancient inscriptions.

Cypriot
Abbr
Cprt
Count
55

Cypriot is a script used to write the ancient Cypriot Greek language, featuring syllabic signs for consonants and vowels.

Cyrillic
Abbr
Cyrl
Count
508

Cyrillic is a script value used for writing Slavic and non-Slavic languages across Eurasia, including Russian, Ukrainian, Serbian, and Mongolian.

Cyrillic (Old Church Slavonic variant)
Abbr
Cyrs
Count
0

Cyrillic (Old Church Slavonic variant) is a script value identifying early Slavic liturgical texts written in a distinctive medieval orthography.

Devanagari
Abbr
Deva
Count
165

Devanagari is a script used for writing Sanskrit, Hindi, and other South Asian languages, with each character representing a syllable.

Dives Akuru
Abbr
Diak
Count
72

Dives Akuru is a script historically used for writing the Sinhala language in Sri Lanka, with its Unicode encoding supporting modern digital use.

Dogra
Abbr
Dogr
Count
60

Dogra is a script used for writing the Dogri language, primarily in the Jammu region of India, with historical ties to the Takri family.

Deseret
Abbr
Dsrt
Count
80

Deseret is a script used for a phonetic English orthography, with its Unicode values covering letters for writing Deseret alphabet text.

Duployan
Abbr
Dupl
Count
143

Duployan is a script used for shorthand, with characters encoded for stenographic writing systems and phonetic transcription.

Egyptian demotic
Abbr
Egyd
Count
0

Egyptian demotic is a script value used to identify the writing system of ancient Egyptian cursive texts, distinct from hieroglyphs and hieratic.

Egyptian hieratic
Abbr
Egyh
Count
0

Egyptian hieratic is a script value used to encode cursive hieroglyphic texts, primarily from ancient Egypt’s religious and administrative documents.

Egyptian Hieroglyphs
Abbr
Egyp
Count
5105

Egyptian Hieroglyphs is a script value used for ancient Egyptian writing, covering logographic and alphabetic signs in the Unicode standard.

Elbasan
Abbr
Elba
Count
40

Elbasan is a historical script used for writing the Albanian language in the 18th century, with Unicode encoding for its characters.

Elymaic
Abbr
Elym
Count
23

Elymaic is a script used for writing the Elymaic language in ancient southwestern Iran, with right-to-left directionality and a consonant-based alphabet.

Ethiopic
Abbr
Ethi
Count
523

Ethiopic is a script value used for Geʽez and related languages like Amharic and Tigrinya, covering syllabic and alphabetic writing.

Garay
Abbr
Gara
Count
69

Garay is a script used for writing the Mandé language, with distinct letterforms and a right‑to‑left direction.

Georgian
Abbr
Geok
Count
0

Georgian is an alphabetic script used for writing the Kartvelian languages, primarily Georgian, with distinct forms like Mkhedruli.

Georgian
Abbr
Geor
Count
173

Georgian is a script used for writing the Kartvelian languages, primarily Georgian, with distinct forms like Mkhedruli and Asomtavruli.

Glagolitic
Abbr
Glag
Count
134

Glagolitic is a script value used for the medieval Slavic writing system, encompassing letters and symbols for languages like Old Church Slavonic.

Gunjala Gondi
Abbr
Gong
Count
63

Gunjala Gondi is a script used for writing the Gondi language, primarily in central India.

Masaram Gondi
Abbr
Gonm
Count
75

Masaram Gondi is a Script value representing the writing system used for the Gondi language, encoded for digital text.

Gothic
Abbr
Goth
Count
27

Gothic is a script used for the East Germanic language, with characters encoded for historical and linguistic documentation.

Grantha
Abbr
Gran
Count
85

Grantha is a script value representing the historical Grantha writing system, used primarily for Sanskrit in South India.

Greek
Abbr
Grek
Count
520

Greek is a script value covering the Greek alphabet, used for Ancient and Modern Greek, plus Coptic and some mathematical symbols.

Gujarati
Abbr
Gujr
Count
91

Gujarati is a script used for writing the Gujarati language, primarily spoken in the Indian state of Gujarat.

Gurung Khema
Abbr
Gukh
Count
58

Gurung Khema is a script used for writing the Tamu language, primarily in Nepal, with its own distinct Unicode character block.

Gurmukhi
Abbr
Guru
Count
80

Gurmukhi is a script used primarily for writing the Punjabi language, with its Unicode value covering characters from U+0A00 to U+0A7F.

Han with Bopomofo (alias for Han + Bopomofo)
Abbr
Hanb
Count
0

Han with Bopomofo (alias for Han + Bopomofo) is a script value that groups Han ideographs with Bopomofo phonetic symbols for unified text processing.

Hangul
Abbr
Hang
Count
11739

Hangul is the script value assigned to Korean characters, covering the modern Hangul syllables and related jamo letters.

Han
Abbr
Hani
Count
103352

Han is the script value for Chinese characters, used historically across China, Japan, Korea, and Vietnam, with simplified and traditional forms.

Hanunoo
Abbr
Hano
Count
21

Hanunoo is a script used for writing the Hanunó’o language, featuring an indigenous syllabic system with 48 characters.

Han (Simplified variant)
Abbr
Hans
Count
0

Han (Simplified variant) is a script value for Chinese characters using simplified forms, distinct from traditional variants, with a single sentence summary.

Han (Traditional variant)
Abbr
Hant
Count
0

Han (Traditional variant) is a script value covering Chinese characters in their traditional forms, used primarily in Taiwan, Hong Kong, and Macau.

Hatran
Abbr
Hatr
Count
26

Hatran is a script value representing the Aramaic-derived writing system used for inscriptions in the ancient city of Hatra.

Hebrew
Abbr
Hebr
Count
136

Hebrew is the script used for writing Hebrew, Yiddish, and other Jewish languages, with right to left orientation and shared letterforms.

Hiragana
Abbr
Hira
Count
382

Hiragana is a script used for native Japanese words, grammatical elements, and phonetic readings, distinct from katakana and kanji.

Anatolian Hieroglyphs
Abbr
Hluw
Count
583

Anatolian Hieroglyphs is a script used in ancient Anatolia, encoded for digital text to represent its syllabic and logographic signs.

Pahawh Hmong
Abbr
Hmng
Count
127

Pahawh Hmong is a script used to write the Hmong language, uniquely invented in 1959 by Shong Lue Yang.

Nyiakeng Puachue Hmong
Abbr
Hmnp
Count
71

Nyiakeng Puachue Hmong is a script used for writing the Hmong language, with its characters encoded in Unicode.

Katakana Or Hiragana
Abbr
Hrkt
Count
0

Katakana Or Hiragana is a script value for characters that can be written in either Japanese syllabary, with usage determined by context.

Old Hungarian
Abbr
Hung
Count
108

Old Hungarian is a script used for the ancient Hungarian language, with characters historically carved on wood or stone.

Indus (Harappan)
Abbr
Inds
Count
0

Indus (Harappan) is a script value representing the undeciphered writing system of the ancient Indus Valley civilization.

Old Italic
Abbr
Ital
Count
39

Old Italic is a script used for ancient Italian languages, including Etruscan and Oscan, with letters derived from Greek.

Jamo (alias for Jamo subset of Hangul)
Abbr
Jamo
Count
0

Jamo (alias for Jamo subset of Hangul) is used for Korean phonetic components, marking consonants and vowels as distinct script characters.

Javanese
Abbr
Java
Count
90

Javanese is a script used for writing the Javanese language, primarily on the Indonesian island of Java.

Japanese (alias for Han + Hiragana + Katakana)
Abbr
Jpan
Count
0

Japanese (alias for Han + Hiragana + Katakana) is a script value used to group CJK ideographs with Japanese syllabaries for text processing.

Jurchen
Abbr
Jurc
Count
965

Jurchen is a historical script used for the Jurchen language, encoded for scholarly and digital preservation of medieval Manchurian texts.

Kayah Li
Abbr
Kali
Count
47

Kayah Li is a script used for writing the Karen languages in Myanmar and Thailand, featuring 48 characters including vowels and tone marks.

Katakana
Abbr
Kana
Count
327

Katakana is a Japanese syllabary used primarily for writing foreign loanwords, onomatopoeia, and emphasis, distinct from hiragana and kanji.

Kawi
Abbr
Kawi
Count
87

Kawi is a historic script used across Southeast Asia, primarily for Old Javanese, Balinese, and Malay inscriptions and literary texts.

Kharoshthi
Abbr
Khar
Count
68

Kharoshthi is a script used in ancient Gandhara for writing Prakrit and Sanskrit, with a right to left direction.

Khmer
Abbr
Khmr
Count
146

Khmer is a script used for writing the Khmer language, primarily in Cambodia, with characters ranging from consonants and vowels to digits and diacritics.

Khojki
Abbr
Khoj
Count
65

Khojki is a script used historically for writing Sindhi and other languages, primarily in the Indian subcontinent.

Khitan large script
Abbr
Kitl
Count
0

Khitan large script is a historical writing system used for the Khitan people, encoded for scholarly and digital text preservation.

Khitan Small Script
Abbr
Kits
Count
477

Khitan Small Script is a writing system used for the extinct Khitan language, encoded in Unicode to represent its unique syllabic and logographic characters.

Kannada
Abbr
Knda
Count
92

Kannada is a Unicode script value representing the writing system used for the Kannada language, primarily in Karnataka, India.

Korean (alias for Hangul + Han)
Abbr
Kore
Count
0

Korean (alias for Hangul + Han) is a script value covering both the Korean alphabet Hangul and Chinese characters Han used in Korean writing.

Kpelle
Abbr
Kpel
Count
0

Kpelle is a script used for writing the Kpelle language of West Africa, characterized by syllabic symbols.

Kirat Rai
Abbr
Krai
Count
58

Kirat Rai is a script used for writing the Kirat languages of Nepal, with unique character shapes and historical significance.

Kaithi
Abbr
Kthi
Count
68

Kaithi is a script used historically for writing languages like Bhojpuri, Magahi, and Maithili in northern India.

Tai Tham
Abbr
Lana
Count
127

Tai Tham is a script used for writing Northern Thai, Tai Lue, and Khün languages, primarily in religious and historical texts.

Lao
Abbr
Laoo
Count
83

Lao is a script used for writing the Lao language, primarily in Laos, with distinctive rounded letterforms and no spaces between words.

Latin (Fraktur variant)
Abbr
Latf
Count
0

Latin (Fraktur variant) is a script value denoting blackletter style letterforms, historically used for German texts, now decorative.

Latin (Gaelic variant)
Abbr
Latg
Count
0

Latin (Gaelic variant) is a script identifier for Irish and Scottish Gaelic orthography, covering letters like á, é, í, ó, ú, and ḃ, ċ, ḋ, ḟ, ġ, ṁ, ṗ, ṡ, ṫ.

Latin
Abbr
Latn
Count
1653

Latin is a writing system value used for texts in languages like English, Spanish, and French, covering letters and symbols.

Leke
Abbr
Leke
Count
0

Leke is a script value used in text processing to identify characters written in the Leke script, primarily for digital rendering and language support.

Lepcha
Abbr
Lepc
Count
74

Lepcha is a script used for writing the Lepcha language of Sikkim, India, with its own distinct letters and tonal marks.

Limbu
Abbr
Limb
Count
68

Limbu is a script used for writing the Limbu language, primarily in Nepal and India, with distinct glyph shapes.

Linear A
Abbr
Lina
Count
341

Linear A is a script value used to encode the ancient Minoan writing system, distinct from Linear B and other scripts.

Linear B
Abbr
Linb
Count
211

Linear B is a script value used to encode ancient Mycenaean Greek syllabic writing, covering signs for syllables, ideograms, and numbers.

Lisu
Abbr
Lisu
Count
49

Lisu is a script used for the Lisu language, originating from Myanmar and China, with alphabetic letters written left to right.

Loma
Abbr
Loma
Count
0

Loma is a writing system used for the Loma language, primarily in Guinea and Liberia, with its own distinct script for tonal syllables.

Lycian
Abbr
Lyci
Count
29

Lycian is a script used for the ancient Anatolian language, encoded in Unicode for historical text representation.

Lydian
Abbr
Lydi
Count
27

Lydian is a script value representing the ancient Anatolian writing system used for the Lydian language, encoded in Unicode.

Mahajani
Abbr
Mahj
Count
39

Mahajani is a script used historically for accounting and mercantile records in northern India, encoded separately in Unicode.

Makasar
Abbr
Maka
Count
25

Makasar is a script used for writing the Makassarese language, historically employed in Sulawesi, Indonesia, with modern revival efforts.

Mandaic
Abbr
Mand
Count
29

Mandaic is a script used for the liturgical language of the Mandaean religion, originating from ancient Mesopotamia and written right to left.

Manichaean
Abbr
Mani
Count
51

Manichaean is a script value used for writing the ancient Manichaean religion’s texts, including liturgical and doctrinal works.

Marchen
Abbr
Marc
Count
68

Marchen is a script used for the historical Tibetan-related Marchen language, with Unicode codepoints assigned for its letters and digits.

Mayan hieroglyphs
Abbr
Maya
Count
0

Mayan hieroglyphs is a script value representing the historical writing system of the Maya civilization, used for inscriptions and codices.

Medefaidrin
Abbr
Medf
Count
91

Medefaidrin is a script used for the Medefaidrin language, with its Unicode characters supporting this indigenous Nigerian writing system.

Mende Kikakui
Abbr
Mend
Count
213

Mende Kikakui is a script used for the Mende language of Sierra Leone, with its own unique syllabic writing system.

Meroitic Cursive
Abbr
Merc
Count
90

Meroitic Cursive is a script value used for writing the ancient Meroitic language, encoded for digital text representation.

Meroitic Hieroglyphs
Abbr
Mero
Count
32

Meroitic Hieroglyphs is a script value used for the ancient cursive and hieroglyphic writing system of the Kingdom of Kush, encoded for digital text.

Malayalam
Abbr
Mlym
Count
118

Malayalam is a script used to write the Malayalam language, primarily in Kerala, India, with its own distinct letterforms and conjuncts.

Modi
Abbr
Modi
Count
79

Modi is a script value used to denote text written in the historical Modi script, primarily for the Marathi language.

Mongolian
Abbr
Mong
Count
168

Mongolian is a script value used for the traditional vertical writing system of the Mongolian language, including its historic forms like Manchu.

Moon (Moon code, Moon script, Moon type)
Abbr
Moon
Count
0

Moon (Moon code, Moon script, Moon type) is a tactile writing system for the blind, using raised curved lines and shapes rather than embossed dots.

Mro
Abbr
Mroo
Count
43

Mro is a script value for the Mru language, used in Myanmar and Bangladesh, with about 50,000 speakers.

Meetei Mayek
Abbr
Mtei
Count
79

Meetei Mayek is a script value used for writing the Manipuri language, primarily in northeastern India.

Multani
Abbr
Mult
Count
38

Multani is a script used for writing the Saraiki language, historically in the Multan region of Pakistan.

Myanmar
Abbr
Mymr
Count
243

Myanmar is a script used for writing Burmese and related languages, featuring circular letters and no spaces between words.

Nag Mundari
Abbr
Nagm
Count
42

Nag Mundari is a script used for writing the Mundari language, primarily in Jharkhand, India, with a distinct character set.

Nandinagari
Abbr
Nand
Count
65

Nandinagari is a historical script used primarily for writing Sanskrit and Kannada manuscripts, especially in southern India.

Old North Arabian
Abbr
Narb
Count
32

Old North Arabian is a script used for ancient inscriptions in the Arabian Peninsula, encoded for historical text representation.

Nabataean
Abbr
Nbat
Count
40

Nabataean is a historical script used for writing the Nabataean language and Aramaic, featuring cursive, right-to-left letterforms.

Newa
Abbr
Newa
Count
97

Newa is a script used for writing the Nepal Bhasa language, primarily in the Kathmandu Valley, with historical and modern applications.

Naxi Dongba (na²¹ɕi³³ to³³ba²¹, Nakhi Tomba)
Abbr
Nkdb
Count
0

Naxi Dongba (na²¹ɕi³³ to³³ba²¹, Nakhi Tomba) is a script value identifying pictographic writing used for ritual texts and everyday records.

Naxi Geba (na²¹ɕi³³ gʌ²¹ba²¹, 'Na-'Khi ²Ggŏ-¹baw, Nakhi Geba)
Abbr
Nkgb
Count
0

Naxi Geba (na²¹ɕi³³ gʌ²¹ba²¹, 'Na-'Khi ²Ggŏ-¹baw, Nakhi Geba) is a script used for writing the Naxi language, with historical syllabic and pictographic forms.

Nko
Abbr
Nkoo
Count
62

Nko is a script used for writing the Manding languages of West Africa, with characters encoded in the U+07C0 to U+07FF range.

Nushu
Abbr
Nshu
Count
397

Nushu is a script used exclusively by women in Hunan, China, encoding a syllabic writing system for the local Xiangnan Tuhua dialect.

Ogham
Abbr
Ogam
Count
29

Ogham is a script used for writing the early Irish language, typified by linear strokes along stone edges, and encoded in Unicode.

Ol Chiki
Abbr
Olck
Count
48

Ol Chiki is a script used for writing the Santali language, with its own distinct Unicode character assignments.

Ol Onal
Abbr
Onao
Count
44

Ol Onal is a script used for writing the Ho language, primarily in eastern India, with an alphabetic system distinct from other regional scripts.

Old Turkic
Abbr
Orkh
Count
73

Old Turkic is a script used for the Orkhon and Yenisei inscriptions, written right to left, with distinct runic-like letterforms.

Oriya
Abbr
Orya
Count
93

Oriya is a script used for writing the Odia language, primarily in the Indian state of Odisha.

Osage
Abbr
Osge
Count
72

Osage is a script value used for writing the Osage language, with characters encoded in the Unicode standard.

Osmanya
Abbr
Osma
Count
40

Osmanya is a script used for writing the Somali language, encoded for digital text support.

Old Uyghur
Abbr
Ougr
Count
26

Old Uyghur is a script used for the Turkic language, written right‑to‑left with distinct letter forms.

Palmyrene
Abbr
Palm
Count
32

Palmyrene is a script used for the Aramaic dialect of the ancient city of Palmyra, written right to left.

Pau Cin Hau
Abbr
Pauc
Count
57

Pau Cin Hau is a script used for writing the Zomi language, with its characters encoded in Unicode.

Proto-Cuneiform
Abbr
Pcun
Count
164

Proto-Cuneiform is a script value used for the earliest Mesopotamian writing system, predating cuneiform and dating to around 3300 BCE.

Proto-Elamite
Abbr
Pelm
Count
0

Proto-Elamite is a script value used for the ancient writing system of Iran, distinct from Linear Elamite, and encoded for historical text representation.

Old Permic
Abbr
Perm
Count
43

Old Permic is a script used for writing the Komi language, primarily in medieval northeastern Europe.

Phags Pa
Abbr
Phag
Count
56

Phags Pa is a script used historically for Mongolian and Chinese, now encoded for scholarly and digital text preservation.

Inscriptional Pahlavi
Abbr
Phli
Count
27

Inscriptional Pahlavi is a script used for writing Middle Persian, primarily in inscriptions from the Parthian and early Sasanian periods.

Psalter Pahlavi
Abbr
Phlp
Count
29

Psalter Pahlavi is a script used for Middle Persian texts, written right to left, with 29 letters.

Book Pahlavi
Abbr
Phlv
Count
0

Book Pahlavi is a script used for writing Middle Persian, primarily in Zoroastrian religious texts from the 3rd to 10th centuries.

Phoenician
Abbr
Phnx
Count
29

Phoenician is a script used for ancient Semitic inscriptions, with letters read right to left and encoded separately from similar alphabets.

Miao
Abbr
Plrd
Count
149

Miao is a script used for writing the Hmong language, primarily in China and Southeast Asia, with distinct tonal and consonant representations.

Klingon (KLI pIqaD)
Abbr
Piqd
Count
0

Klingon (KLI pIqaD) is a constructed script used for writing the Klingon language, featuring distinct glyphs and a left-to-right orientation.

Inscriptional Parthian
Abbr
Prti
Count
30

Inscriptional Parthian is a script used for writing the Parthian language, primarily on stone inscriptions from ancient Iran and Mesopotamia.

Proto-Sinaitic
Abbr
Psin
Count
0

Proto-Sinaitic is an ancient script value marking inscriptions from the Sinai Peninsula, dating to roughly the 19th to 16th century BCE.

Reserved for private use (start)
Abbr
Qaaa
Count
0

Reserved for private use (start) is the initial code point in a range that applications can assign custom characters to without standardized meaning.

Reserved for private use (end)
Abbr
Qabx
Count
0

Reserved for private use (end) is a sentinel value marking the upper boundary of the private use script range.

Ranjana
Abbr
Ranj
Count
0

Ranjana is a script used historically for writing Nepali and Maithili, characterized by curved, ornate letterforms often seen in Buddhist manuscripts.

Rejang
Abbr
Rjng
Count
37

Rejang is a script used historically for writing the Rejang language of Sumatra, characterized by its distinctive abugida syllabic system.

Hanifi Rohingya
Abbr
Rohg
Count
50

Hanifi Rohingya is a script used for writing the Rohingya language, with letters encoded for digital text.

Rongorongo
Abbr
Roro
Count
0

Rongorongo is a script value used for the undeciphered glyphs of Easter Island, encoded in the Unicode standard.

Runic
Abbr
Runr
Count
86

Runic is a script value identifying characters used for ancient Germanic inscriptions, including Elder and Younger Futhark, with regional variants.

Samaritan
Abbr
Samr
Count
61

Samaritan is a script used for the Samaritan Pentateuch, with writing from right to left and roots in ancient Hebrew.

Sarati
Abbr
Sara
Count
0

Sarati is a script value identifying the constructed alphabet used for invented languages, primarily in Tolkien’s fictional works.

Old South Arabian
Abbr
Sarb
Count
32

Old South Arabian is a historical script value used for encoding ancient Yemeni inscriptions, including Sabaean and Minaic texts.

Saurashtra
Abbr
Saur
Count
82

Saurashtra is a script used historically for the Saurashtra language in Gujarat, India, now revived in modern digital text.

(Small) Seal
Abbr
Seal
Count
11328
SignWriting
Abbr
Sgnw
Count
672

SignWriting is a script value indicating a visual notation system for sign languages, using iconic symbols for handshapes, movements, and facial expressions.

Shavian
Abbr
Shaw
Count
48

Shavian is a script value used to encode texts in the phonemic alphabet devised by George Bernard Shaw for English.

Sharada
Abbr
Shrd
Count
104

Sharada is a script used for writing Sanskrit and Kashmiri, primarily in historical inscriptions and manuscripts from northwestern India.

Shuishu
Abbr
Shui
Count
0

Shuishu is a script value used to encode the historical Shuishu writing system, primarily for ritual texts of the Shui people.

Siddham
Abbr
Sidd
Count
92

Siddham is an ancient Indian script used to write Sanskrit, primarily in Buddhist texts and esoteric rituals across East Asia.

Sidetic
Abbr
Sidt
Count
26

Sidetic is an ancient script from Asia Minor, used for inscriptions, encoded in Unicode for digital representation.

Khudawadi
Abbr
Sind
Count
69

Khudawadi is a script used for writing the Sindhi language, primarily in the Khudawadi region of India.

Sinhala
Abbr
Sinh
Count
111

Sinhala is a script used for writing the Sinhalese language, primarily in Sri Lanka, with distinct character shapes and vowel marks.

Sogdian
Abbr
Sogd
Count
42

Sogdian is a script used for the ancient Eastern Iranian language, written right to left with cursive and separated forms.

Old Sogdian
Abbr
Sogo
Count
40

Old Sogdian is a script used for an ancient Eastern Iranian language, written right-to-left, with letters denoting consonants and some vowels.

Sora Sompeng
Abbr
Sora
Count
35

Sora Sompeng is a script used for writing the Sora language, primarily spoken in eastern India.

Soyombo
Abbr
Soyo
Count
83

Soyombo is a historical script from Mongolia and Tibet, used mainly for Buddhist texts and the Mongolian flag’s national emblem.

Sundanese
Abbr
Sund
Count
72

Sundanese is the script value used for writing the Sundanese language, primarily in West Java, Indonesia, with its own distinct characters.

Sunuwar
Abbr
Sunu
Count
44

Sunuwar is a script used for writing the Sunuwar language, primarily in Nepal, with its own distinct glyphs and characters.

Syloti Nagri
Abbr
Sylo
Count
45

Syloti Nagri is a script used for writing the Sylheti language, primarily in northeastern Bangladesh and nearby Indian regions.

Syriac
Abbr
Syrc
Count
88

Syriac is a writing system used for Aramaic dialects, with distinct letter forms for classical, Eastern, and Western variants.

Syriac (Estrangelo variant)
Abbr
Syre
Count
0

Syriac (Estrangelo variant) is a script value used for classical Syriac texts, written right to left, distinct from other Syriac forms.

Syriac (Western variant)
Abbr
Syrj
Count
0

Syriac (Western variant) is a script value identifying text written in the West Syriac tradition, used mainly by the Syriac Orthodox Church.

Syriac (Eastern variant)
Abbr
Syrn
Count
0

Syriac (Eastern variant) is used to write texts in the classical Syriac language, primarily in liturgical and historical contexts.

Tagbanwa
Abbr
Tagb
Count
18

Tagbanwa is a script used for writing the Tagbanwa language, primarily in the Philippines, with its own distinct character shapes and syllabic structure.

Takri
Abbr
Takr
Count
68

Takri is a script used historically for languages like Dogri, Kangri, and Chamba, now encoded for digital text.

Tai Le
Abbr
Tale
Count
35

Tai Le is a script used for writing the Tai Nüa language, primarily in southwestern China and parts of Myanmar.

New Tai Lue
Abbr
Talu
Count
83

New Tai Lue is a script used for writing the Tai Lü language, primarily in southern China and Southeast Asia, with distinct rounded letterforms.

Tamil
Abbr
Taml
Count
123

Tamil is a script used for writing the Tamil language, primarily in southern India and Sri Lanka, with distinct letterforms.

Tangut
Abbr
Tang
Count
7061

Tangut is a script used for the Tangut language, primarily in the Western Xia dynasty, with characters modeled on Chinese but unique in structure.

Tai Viet
Abbr
Tavt
Count
72

Tai Viet is a script used for writing Tai Dam and related languages, with letters arranged in a distinctive syllabic style.

Tai Yo
Abbr
Tayo
Count
55

Tai Yo is a script used for writing the Tai Yo language, primarily in Vietnam and Laos, with distinct glyph forms.

Telugu
Abbr
Telu
Count
101

Telugu is a script used primarily for writing the Telugu language, predominantly spoken in the Indian state of Andhra Pradesh and Telangana.

Tengwar
Abbr
Teng
Count
0

Tengwar is a script used to write invented languages, like Quenya and Sindarin, with values assigned in Unicode for encoding its letters.

Tifinagh
Abbr
Tfng
Count
59

Tifinagh is a script used for writing Berber languages, with Unicode assigning characters for its letters and diacritics across two blocks.

Tagalog
Abbr
Tglg
Count
23

Tagalog is a script used for writing the Philippine language, with its Unicode characters representing syllables and historical Baybayin forms.

Thaana
Abbr
Thaa
Count
50

Thaana is the script used for writing Dhivehi, the official language of the Maldives, with right‑to‑left orientation and derived from Arabic numerals.

Thai
Abbr
Thai
Count
86

Thai is a script used for writing the Thai language, characterized by its distinct, rounded letterforms and no spaces between words.

Tibetan
Abbr
Tibt
Count
207

Tibetan is the script value for characters used to write the Tibetan language, including its historical and religious texts.

Tirhuta
Abbr
Tirh
Count
82

Tirhuta is a script used historically for writing the Maithili language, primarily in the Mithila region of Nepal and India.

Tangsa
Abbr
Tnsa
Count
89

Tangsa is a script used for writing several Tibeto-Burman languages spoken in northeastern India and adjacent Myanmar, with a distinct set of letters.

Todhri
Abbr
Todr
Count
52

Todhri is a script value used for the historical Albanian Todhri alphabet, distinct from Latin and Greek.

Tolong Siki
Abbr
Tols
Count
54

Tolong Siki is a script used for the Bugis language, primarily written in South Sulawesi, Indonesia, with its own distinct character forms.

Toto
Abbr
Toto
Count
31

Toto is a script value identifying a specific writing system, used to classify characters by their linguistic origin and historical usage.

Tulu Tigalari
Abbr
Tutg
Count
80

Tulu Tigalari is a script used historically for writing Tulu and Sanskrit in the Karnataka region of India.

Ugaritic
Abbr
Ugar
Count
31

Ugaritic is a script used for the ancient Semitic language of Ugarit, written in cuneiform signs.

Vai
Abbr
Vaii
Count
300

Vai is a script used for the Vai language of Liberia, featuring a unique syllabary with 300 characters.

Visible Speech
Abbr
Visp
Count
0

Visible Speech is a script value identifying characters used for phonetic notation, distinct from Latin and other writing systems.

Vithkuqi
Abbr
Vith
Count
70

Vithkuqi is a script used historically for writing the Albanian language, featuring distinct letterforms and a right-to-left direction.

Warang Citi
Abbr
Wara
Count
84

Warang Citi is a script used for writing the Ho language, encoded with its own distinct character set.

Wancho
Abbr
Wcho
Count
59

Wancho is a script used for the Wancho language of northeast India, featuring an alphabetic system with 49 letters.

Woleai
Abbr
Wole
Count
0

Woleai is a writing system used historically for the Woleaian language on the Caroline Islands, now encoded for digital text.

Old Persian
Abbr
Xpeo
Count
50

Old Persian is a script used for cuneiform inscriptions of ancient Iran, valued for encoding royal texts from the Achaemenid era.

Cuneiform
Abbr
Xsux
Count
1393

Cuneiform is a script value representing the ancient Sumerian, Akkadian, and other Mesopotamian writing systems, encoded for digital text.

Yezidi
Abbr
Yezi
Count
47

Yezidi is a writing system used for the Kurdish Yezidi language, encoded with its own script in digital text.

Yi
Abbr
Yiii
Count
1220

Yi is a script value used for the Yi syllabary and logograms, primarily for the Nuosu language in southwestern China.

Zanabazar Square
Abbr
Zanb
Count
72

Zanabazar Square is a script value used for the historical Mongolian script, encoded for texts in Buddhist and secular documents.

Inherited
Abbr
Zinh
Count
695

Inherited is used when a character’s script is determined by surrounding characters, like combining marks or punctuation.

Mathematical notation
Abbr
Zmth
Count
0

Mathematical notation is used to identify which mathematical script, like Latin or Fraktur, a character belongs to for semantic rendering.

Symbols (Emoji variant)
Abbr
Zsye
Count
0

Symbols (Emoji variant) is a script value used for characters that represent emoji, pictographs, and related symbolic glyphs without a distinct written script.

Symbols
Abbr
Zsym
Count
0

Symbols is a script value for characters used in symbolic, non-alphabetic writing systems, including math, currency, and technical notation.

Code for unwritten documents
Abbr
Zxxx
Count
0

Code for unwritten documents is a special script value used to mark text in unknown, invented, or undeciphered writing systems.

Common
Abbr
Zyyy
Count
9276

Common is a script value used for characters shared across writing systems, like punctuation, digits, and symbols, lacking a specific script identity.

Unknown
Abbr
Zzzz
Count
137468

Unknown is a fallback for characters not assigned to any specific script, covering unassigned, private use, and non-script codepoints.