Script
Script is a character property that assigns each encoded character to a specific writing system, such as Latin, Greek, Cyrillic, Arabic, or Han, based on its historical and linguistic origin. This categorization helps software render text correctly by selecting appropriate fonts and shaping rules, and it enables text processing tasks like language detection, sorting, and searching. Unlike simple block assignments, which group characters by code point ranges, script reflects actual usage, so a single block may contain characters from multiple scripts, and characters like punctuation or digits are often classified as "Common" or "Inherited" since they are shared across systems. The value is a stable identifier that aids in internationalization and accessibility, though a single character can only belong to one script, while text as a whole can mix many scripts.
- Abbr
- Adlm
- Count
- 88
Adlam is a script used for writing the Fulani language, encoded with its own character block for digital text.
- Abbr
- Afak
- Count
- 0
Afaka is a script used for writing the Aukan language, also known as Ndyuka, a creole spoken in Suriname.
- Abbr
- Aghb
- Count
- 53
Caucasian Albanian is a script used for an ancient Christian language in the Caucasus, encoded for digital text.
- Abbr
- Ahom
- Count
- 65
Ahom is a historical script used for the Tai language of the Ahom people, encoded for scholarly and digital preservation.
- Abbr
- Arab
- Count
- 1452
Arabic is a script used for writing Arabic, Persian, Urdu, and several other languages, with cursive letterforms and right-to-left direction.
- Abbr
- Aran
- Count
- 0
Arabic (Nastaliq variant) is a script value used for languages like Urdu and Punjabi, characterized by its flowing, slanted calligraphic style.
- Abbr
- Armi
- Count
- 31
Imperial Aramaic is a script value used for writing ancient Aramaic from the 8th to 3rd centuries BCE, including in Persian Empire records.
- Abbr
- Armn
- Count
- 99
Armenian is a script value used to identify text written in the Armenian alphabet, supporting both classical and modern forms.
- Abbr
- Avst
- Count
- 61
Avestan is the script used for writing the sacred texts of Zoroastrianism, originating in ancient Iran.
- Abbr
- Bali
- Count
- 127
Balinese is a script used for writing the Balinese language, with its Unicode value covering consonants, vowels, and diacritics.
- Abbr
- Bamu
- Count
- 657
Bamum is a script used for the Bamum language and kingdom in Cameroon, with historical and modern variants.
- Abbr
- Bass
- Count
- 36
Bassa Vah is a script used for writing the Bassa language, with its Unicode characters encoded in the range U+16AD0 to U+16AFF.
- Abbr
- Batk
- Count
- 56
Batak is a script used for writing the Batak languages of Sumatra, with characters that include syllabic signs, vowels, and punctuation marks.
- Abbr
- Beng
- Count
- 98
Bengali is a writing system used for the Assamese and Bengali languages, with distinct letterforms and conjunct ligatures.
- Abbr
- Berf
- Count
- 50
Beria Erfe is a script value used for writing the Beria language, primarily in Chad and Sudan.
- Abbr
- Bhks
- Count
- 97
Bhaiksuki is a historical script used for writing Buddhist texts, primarily in Sanskrit, from around the 11th to 12th centuries in northeastern India.
- Abbr
- Blis
- Count
- 0
Blissymbols is a constructed ideographic writing system used for augmentative and alternative communication, with its own script category in Unicode.
- Abbr
- Bopo
- Count
- 77
Bopomofo is a script value representing the phonetic symbols used in Taiwanese Mandarin education, covering characters like ㄅ to ㄦ.
- Abbr
- Brah
- Count
- 115
Brahmi is an ancient writing system of South Asia, historically used for Sanskrit and Prakrit, and encoded for digital text.
- Abbr
- Brai
- Count
- 256
Braille is a script used to encode tactile writing systems, with Unicode codepoints for all six and eight dot patterns.
- Abbr
- Bugi
- Count
- 30
Buginese is an abugida script used for writing the Bugis language, primarily in South Sulawesi, Indonesia.
- Abbr
- Buhd
- Count
- 20
Buhid is a script used for writing the Buhid language, with its characters encoded in Unicode for digital text representation.
- Abbr
- Cakm
- Count
- 71
Chakma is a Brahmic script used for the Chakma language, featuring unique vowel and consonant signs, with a distinct set of digits.
- Abbr
- Cans
- Count
- 726
Canadian Aboriginal is an assigned script value indicating text written in Canadian Indigenous syllabics, used for languages like Cree and Inuktitut.
- Abbr
- Cari
- Count
- 49
Carian is a script used for the ancient Carian language of Anatolia, written right to left with a distinct alphabet.
- Abbr
- Cham
- Count
- 83
Cham is a script used for writing the Eastern Cham language, primarily in Vietnam and Cambodia.
- Abbr
- Cher
- Count
- 172
Cherokee is a North American syllabary script used for writing the Cherokee language, with distinct characters for syllables and a unique visual style.
- Abbr
- Chis
- Count
- 0
Chisoi is a script value denoting the historical Chisoi writing system, used for inscriptions in ancient Central Asia.
- Abbr
- Chrs
- Count
- 28
Chorasmian is a script used for the extinct Eastern Iranian language of the same name, encoded for historic texts.
- Abbr
- Cirt
- Count
- 0
Cirth is a fictional runic script from J.R.R. Tolkien's legendarium, used for writing Dwarvish and other Middle-earth languages.
- Abbr
- Copt
- Count
- 137
Coptic is a historical script used for the Egyptian Coptic language, derived from Greek with added letters for native sounds.
- Abbr
- Cpmn
- Count
- 99
Cypro Minoan is a script value used for the undeciphered writing system of Late Bronze Age Cyprus, encoded for ancient inscriptions.
- Abbr
- Cprt
- Count
- 55
Cypriot is a script used to write the ancient Cypriot Greek language, featuring syllabic signs for consonants and vowels.
- Abbr
- Cyrl
- Count
- 508
Cyrillic is a script value used for writing Slavic and non-Slavic languages across Eurasia, including Russian, Ukrainian, Serbian, and Mongolian.
- Abbr
- Cyrs
- Count
- 0
Cyrillic (Old Church Slavonic variant) is a script value identifying early Slavic liturgical texts written in a distinctive medieval orthography.
- Abbr
- Deva
- Count
- 165
Devanagari is a script used for writing Sanskrit, Hindi, and other South Asian languages, with each character representing a syllable.
- Abbr
- Diak
- Count
- 72
Dives Akuru is a script historically used for writing the Sinhala language in Sri Lanka, with its Unicode encoding supporting modern digital use.
- Abbr
- Dogr
- Count
- 60
Dogra is a script used for writing the Dogri language, primarily in the Jammu region of India, with historical ties to the Takri family.
- Abbr
- Dsrt
- Count
- 80
Deseret is a script used for a phonetic English orthography, with its Unicode values covering letters for writing Deseret alphabet text.
- Abbr
- Dupl
- Count
- 143
Duployan is a script used for shorthand, with characters encoded for stenographic writing systems and phonetic transcription.
- Abbr
- Egyd
- Count
- 0
Egyptian demotic is a script value used to identify the writing system of ancient Egyptian cursive texts, distinct from hieroglyphs and hieratic.
- Abbr
- Egyh
- Count
- 0
Egyptian hieratic is a script value used to encode cursive hieroglyphic texts, primarily from ancient Egypt’s religious and administrative documents.
- Abbr
- Egyp
- Count
- 5105
Egyptian Hieroglyphs is a script value used for ancient Egyptian writing, covering logographic and alphabetic signs in the Unicode standard.
- Abbr
- Elba
- Count
- 40
Elbasan is a historical script used for writing the Albanian language in the 18th century, with Unicode encoding for its characters.
- Abbr
- Elym
- Count
- 23
Elymaic is a script used for writing the Elymaic language in ancient southwestern Iran, with right-to-left directionality and a consonant-based alphabet.
- Abbr
- Ethi
- Count
- 523
Ethiopic is a script value used for Geʽez and related languages like Amharic and Tigrinya, covering syllabic and alphabetic writing.
- Abbr
- Gara
- Count
- 69
Garay is a script used for writing the Mandé language, with distinct letterforms and a right‑to‑left direction.
- Abbr
- Geok
- Count
- 0
Georgian is an alphabetic script used for writing the Kartvelian languages, primarily Georgian, with distinct forms like Mkhedruli.
- Abbr
- Geor
- Count
- 173
Georgian is a script used for writing the Kartvelian languages, primarily Georgian, with distinct forms like Mkhedruli and Asomtavruli.
- Abbr
- Glag
- Count
- 134
Glagolitic is a script value used for the medieval Slavic writing system, encompassing letters and symbols for languages like Old Church Slavonic.
- Abbr
- Gong
- Count
- 63
Gunjala Gondi is a script used for writing the Gondi language, primarily in central India.
- Abbr
- Gonm
- Count
- 75
Masaram Gondi is a Script value representing the writing system used for the Gondi language, encoded for digital text.
- Abbr
- Goth
- Count
- 27
Gothic is a script used for the East Germanic language, with characters encoded for historical and linguistic documentation.
- Abbr
- Gran
- Count
- 85
Grantha is a script value representing the historical Grantha writing system, used primarily for Sanskrit in South India.
- Abbr
- Grek
- Count
- 520
Greek is a script value covering the Greek alphabet, used for Ancient and Modern Greek, plus Coptic and some mathematical symbols.
- Abbr
- Gujr
- Count
- 91
Gujarati is a script used for writing the Gujarati language, primarily spoken in the Indian state of Gujarat.
- Abbr
- Gukh
- Count
- 58
Gurung Khema is a script used for writing the Tamu language, primarily in Nepal, with its own distinct Unicode character block.
- Abbr
- Guru
- Count
- 80
Gurmukhi is a script used primarily for writing the Punjabi language, with its Unicode value covering characters from U+0A00 to U+0A7F.
- Abbr
- Hanb
- Count
- 0
Han with Bopomofo (alias for Han + Bopomofo) is a script value that groups Han ideographs with Bopomofo phonetic symbols for unified text processing.
- Abbr
- Hang
- Count
- 11739
Hangul is the script value assigned to Korean characters, covering the modern Hangul syllables and related jamo letters.
- Abbr
- Hani
- Count
- 103352
Han is the script value for Chinese characters, used historically across China, Japan, Korea, and Vietnam, with simplified and traditional forms.
- Abbr
- Hano
- Count
- 21
Hanunoo is a script used for writing the Hanunó’o language, featuring an indigenous syllabic system with 48 characters.
- Abbr
- Hans
- Count
- 0
Han (Simplified variant) is a script value for Chinese characters using simplified forms, distinct from traditional variants, with a single sentence summary.
- Abbr
- Hant
- Count
- 0
Han (Traditional variant) is a script value covering Chinese characters in their traditional forms, used primarily in Taiwan, Hong Kong, and Macau.
- Abbr
- Hatr
- Count
- 26
Hatran is a script value representing the Aramaic-derived writing system used for inscriptions in the ancient city of Hatra.
- Abbr
- Hebr
- Count
- 136
Hebrew is the script used for writing Hebrew, Yiddish, and other Jewish languages, with right to left orientation and shared letterforms.
- Abbr
- Hira
- Count
- 382
Hiragana is a script used for native Japanese words, grammatical elements, and phonetic readings, distinct from katakana and kanji.
- Abbr
- Hluw
- Count
- 583
Anatolian Hieroglyphs is a script used in ancient Anatolia, encoded for digital text to represent its syllabic and logographic signs.
- Abbr
- Hmng
- Count
- 127
Pahawh Hmong is a script used to write the Hmong language, uniquely invented in 1959 by Shong Lue Yang.
- Abbr
- Hmnp
- Count
- 71
Nyiakeng Puachue Hmong is a script used for writing the Hmong language, with its characters encoded in Unicode.
- Abbr
- Hrkt
- Count
- 0
Katakana Or Hiragana is a script value for characters that can be written in either Japanese syllabary, with usage determined by context.
- Abbr
- Hung
- Count
- 108
Old Hungarian is a script used for the ancient Hungarian language, with characters historically carved on wood or stone.
- Abbr
- Inds
- Count
- 0
Indus (Harappan) is a script value representing the undeciphered writing system of the ancient Indus Valley civilization.
- Abbr
- Ital
- Count
- 39
Old Italic is a script used for ancient Italian languages, including Etruscan and Oscan, with letters derived from Greek.
- Abbr
- Jamo
- Count
- 0
Jamo (alias for Jamo subset of Hangul) is used for Korean phonetic components, marking consonants and vowels as distinct script characters.
- Abbr
- Java
- Count
- 90
Javanese is a script used for writing the Javanese language, primarily on the Indonesian island of Java.
- Abbr
- Jpan
- Count
- 0
Japanese (alias for Han + Hiragana + Katakana) is a script value used to group CJK ideographs with Japanese syllabaries for text processing.
- Abbr
- Jurc
- Count
- 965
Jurchen is a historical script used for the Jurchen language, encoded for scholarly and digital preservation of medieval Manchurian texts.
- Abbr
- Kali
- Count
- 47
Kayah Li is a script used for writing the Karen languages in Myanmar and Thailand, featuring 48 characters including vowels and tone marks.
- Abbr
- Kana
- Count
- 327
Katakana is a Japanese syllabary used primarily for writing foreign loanwords, onomatopoeia, and emphasis, distinct from hiragana and kanji.
- Abbr
- Kawi
- Count
- 87
Kawi is a historic script used across Southeast Asia, primarily for Old Javanese, Balinese, and Malay inscriptions and literary texts.
- Abbr
- Khar
- Count
- 68
Kharoshthi is a script used in ancient Gandhara for writing Prakrit and Sanskrit, with a right to left direction.
- Abbr
- Khmr
- Count
- 146
Khmer is a script used for writing the Khmer language, primarily in Cambodia, with characters ranging from consonants and vowels to digits and diacritics.
- Abbr
- Khoj
- Count
- 65
Khojki is a script used historically for writing Sindhi and other languages, primarily in the Indian subcontinent.
- Abbr
- Kitl
- Count
- 0
Khitan large script is a historical writing system used for the Khitan people, encoded for scholarly and digital text preservation.
- Abbr
- Kits
- Count
- 477
Khitan Small Script is a writing system used for the extinct Khitan language, encoded in Unicode to represent its unique syllabic and logographic characters.
- Abbr
- Knda
- Count
- 92
Kannada is a Unicode script value representing the writing system used for the Kannada language, primarily in Karnataka, India.
- Abbr
- Kore
- Count
- 0
Korean (alias for Hangul + Han) is a script value covering both the Korean alphabet Hangul and Chinese characters Han used in Korean writing.
- Abbr
- Kpel
- Count
- 0
Kpelle is a script used for writing the Kpelle language of West Africa, characterized by syllabic symbols.
- Abbr
- Krai
- Count
- 58
Kirat Rai is a script used for writing the Kirat languages of Nepal, with unique character shapes and historical significance.
- Abbr
- Kthi
- Count
- 68
Kaithi is a script used historically for writing languages like Bhojpuri, Magahi, and Maithili in northern India.
- Abbr
- Lana
- Count
- 127
Tai Tham is a script used for writing Northern Thai, Tai Lue, and Khün languages, primarily in religious and historical texts.
- Abbr
- Laoo
- Count
- 83
Lao is a script used for writing the Lao language, primarily in Laos, with distinctive rounded letterforms and no spaces between words.
- Abbr
- Latf
- Count
- 0
Latin (Fraktur variant) is a script value denoting blackletter style letterforms, historically used for German texts, now decorative.
- Abbr
- Latg
- Count
- 0
Latin (Gaelic variant) is a script identifier for Irish and Scottish Gaelic orthography, covering letters like á, é, í, ó, ú, and ḃ, ċ, ḋ, ḟ, ġ, ṁ, ṗ, ṡ, ṫ.
- Abbr
- Latn
- Count
- 1653
Latin is a writing system value used for texts in languages like English, Spanish, and French, covering letters and symbols.
- Abbr
- Leke
- Count
- 0
Leke is a script value used in text processing to identify characters written in the Leke script, primarily for digital rendering and language support.
- Abbr
- Lepc
- Count
- 74
Lepcha is a script used for writing the Lepcha language of Sikkim, India, with its own distinct letters and tonal marks.
- Abbr
- Limb
- Count
- 68
Limbu is a script used for writing the Limbu language, primarily in Nepal and India, with distinct glyph shapes.
- Abbr
- Lina
- Count
- 341
Linear A is a script value used to encode the ancient Minoan writing system, distinct from Linear B and other scripts.
- Abbr
- Linb
- Count
- 211
Linear B is a script value used to encode ancient Mycenaean Greek syllabic writing, covering signs for syllables, ideograms, and numbers.
- Abbr
- Lisu
- Count
- 49
Lisu is a script used for the Lisu language, originating from Myanmar and China, with alphabetic letters written left to right.
- Abbr
- Loma
- Count
- 0
Loma is a writing system used for the Loma language, primarily in Guinea and Liberia, with its own distinct script for tonal syllables.
- Abbr
- Lyci
- Count
- 29
Lycian is a script used for the ancient Anatolian language, encoded in Unicode for historical text representation.
- Abbr
- Lydi
- Count
- 27
Lydian is a script value representing the ancient Anatolian writing system used for the Lydian language, encoded in Unicode.
- Abbr
- Mahj
- Count
- 39
Mahajani is a script used historically for accounting and mercantile records in northern India, encoded separately in Unicode.
- Abbr
- Maka
- Count
- 25
Makasar is a script used for writing the Makassarese language, historically employed in Sulawesi, Indonesia, with modern revival efforts.
- Abbr
- Mand
- Count
- 29
Mandaic is a script used for the liturgical language of the Mandaean religion, originating from ancient Mesopotamia and written right to left.
- Abbr
- Mani
- Count
- 51
Manichaean is a script value used for writing the ancient Manichaean religion’s texts, including liturgical and doctrinal works.
- Abbr
- Marc
- Count
- 68
Marchen is a script used for the historical Tibetan-related Marchen language, with Unicode codepoints assigned for its letters and digits.
- Abbr
- Maya
- Count
- 0
Mayan hieroglyphs is a script value representing the historical writing system of the Maya civilization, used for inscriptions and codices.
- Abbr
- Medf
- Count
- 91
Medefaidrin is a script used for the Medefaidrin language, with its Unicode characters supporting this indigenous Nigerian writing system.
- Abbr
- Mend
- Count
- 213
Mende Kikakui is a script used for the Mende language of Sierra Leone, with its own unique syllabic writing system.
- Abbr
- Merc
- Count
- 90
Meroitic Cursive is a script value used for writing the ancient Meroitic language, encoded for digital text representation.
- Abbr
- Mero
- Count
- 32
Meroitic Hieroglyphs is a script value used for the ancient cursive and hieroglyphic writing system of the Kingdom of Kush, encoded for digital text.
- Abbr
- Mlym
- Count
- 118
Malayalam is a script used to write the Malayalam language, primarily in Kerala, India, with its own distinct letterforms and conjuncts.
- Abbr
- Modi
- Count
- 79
Modi is a script value used to denote text written in the historical Modi script, primarily for the Marathi language.
- Abbr
- Mong
- Count
- 168
Mongolian is a script value used for the traditional vertical writing system of the Mongolian language, including its historic forms like Manchu.
- Abbr
- Moon
- Count
- 0
Moon (Moon code, Moon script, Moon type) is a tactile writing system for the blind, using raised curved lines and shapes rather than embossed dots.
- Abbr
- Mroo
- Count
- 43
Mro is a script value for the Mru language, used in Myanmar and Bangladesh, with about 50,000 speakers.
- Abbr
- Mtei
- Count
- 79
Meetei Mayek is a script value used for writing the Manipuri language, primarily in northeastern India.
- Abbr
- Mult
- Count
- 38
Multani is a script used for writing the Saraiki language, historically in the Multan region of Pakistan.
- Abbr
- Mymr
- Count
- 243
Myanmar is a script used for writing Burmese and related languages, featuring circular letters and no spaces between words.
- Abbr
- Nagm
- Count
- 42
Nag Mundari is a script used for writing the Mundari language, primarily in Jharkhand, India, with a distinct character set.
- Abbr
- Nand
- Count
- 65
Nandinagari is a historical script used primarily for writing Sanskrit and Kannada manuscripts, especially in southern India.
- Abbr
- Narb
- Count
- 32
Old North Arabian is a script used for ancient inscriptions in the Arabian Peninsula, encoded for historical text representation.
- Abbr
- Nbat
- Count
- 40
Nabataean is a historical script used for writing the Nabataean language and Aramaic, featuring cursive, right-to-left letterforms.
- Abbr
- Newa
- Count
- 97
Newa is a script used for writing the Nepal Bhasa language, primarily in the Kathmandu Valley, with historical and modern applications.
- Abbr
- Nkdb
- Count
- 0
Naxi Dongba (na²¹ɕi³³ to³³ba²¹, Nakhi Tomba) is a script value identifying pictographic writing used for ritual texts and everyday records.
- Abbr
- Nkgb
- Count
- 0
Naxi Geba (na²¹ɕi³³ gʌ²¹ba²¹, 'Na-'Khi ²Ggŏ-¹baw, Nakhi Geba) is a script used for writing the Naxi language, with historical syllabic and pictographic forms.
- Abbr
- Nkoo
- Count
- 62
Nko is a script used for writing the Manding languages of West Africa, with characters encoded in the U+07C0 to U+07FF range.
- Abbr
- Nshu
- Count
- 397
Nushu is a script used exclusively by women in Hunan, China, encoding a syllabic writing system for the local Xiangnan Tuhua dialect.
- Abbr
- Ogam
- Count
- 29
Ogham is a script used for writing the early Irish language, typified by linear strokes along stone edges, and encoded in Unicode.
- Abbr
- Olck
- Count
- 48
Ol Chiki is a script used for writing the Santali language, with its own distinct Unicode character assignments.
- Abbr
- Onao
- Count
- 44
Ol Onal is a script used for writing the Ho language, primarily in eastern India, with an alphabetic system distinct from other regional scripts.
- Abbr
- Orkh
- Count
- 73
Old Turkic is a script used for the Orkhon and Yenisei inscriptions, written right to left, with distinct runic-like letterforms.
- Abbr
- Orya
- Count
- 93
Oriya is a script used for writing the Odia language, primarily in the Indian state of Odisha.
- Abbr
- Osge
- Count
- 72
Osage is a script value used for writing the Osage language, with characters encoded in the Unicode standard.
- Abbr
- Osma
- Count
- 40
Osmanya is a script used for writing the Somali language, encoded for digital text support.
- Abbr
- Ougr
- Count
- 26
Old Uyghur is a script used for the Turkic language, written right‑to‑left with distinct letter forms.
- Abbr
- Palm
- Count
- 32
Palmyrene is a script used for the Aramaic dialect of the ancient city of Palmyra, written right to left.
- Abbr
- Pauc
- Count
- 57
Pau Cin Hau is a script used for writing the Zomi language, with its characters encoded in Unicode.
- Abbr
- Pcun
- Count
- 164
Proto-Cuneiform is a script value used for the earliest Mesopotamian writing system, predating cuneiform and dating to around 3300 BCE.
- Abbr
- Pelm
- Count
- 0
Proto-Elamite is a script value used for the ancient writing system of Iran, distinct from Linear Elamite, and encoded for historical text representation.
- Abbr
- Perm
- Count
- 43
Old Permic is a script used for writing the Komi language, primarily in medieval northeastern Europe.
- Abbr
- Phag
- Count
- 56
Phags Pa is a script used historically for Mongolian and Chinese, now encoded for scholarly and digital text preservation.
- Abbr
- Phli
- Count
- 27
Inscriptional Pahlavi is a script used for writing Middle Persian, primarily in inscriptions from the Parthian and early Sasanian periods.
- Abbr
- Phlp
- Count
- 29
Psalter Pahlavi is a script used for Middle Persian texts, written right to left, with 29 letters.
- Abbr
- Phlv
- Count
- 0
Book Pahlavi is a script used for writing Middle Persian, primarily in Zoroastrian religious texts from the 3rd to 10th centuries.
- Abbr
- Phnx
- Count
- 29
Phoenician is a script used for ancient Semitic inscriptions, with letters read right to left and encoded separately from similar alphabets.
- Abbr
- Plrd
- Count
- 149
Miao is a script used for writing the Hmong language, primarily in China and Southeast Asia, with distinct tonal and consonant representations.
- Abbr
- Piqd
- Count
- 0
Klingon (KLI pIqaD) is a constructed script used for writing the Klingon language, featuring distinct glyphs and a left-to-right orientation.
- Abbr
- Prti
- Count
- 30
Inscriptional Parthian is a script used for writing the Parthian language, primarily on stone inscriptions from ancient Iran and Mesopotamia.
- Abbr
- Psin
- Count
- 0
Proto-Sinaitic is an ancient script value marking inscriptions from the Sinai Peninsula, dating to roughly the 19th to 16th century BCE.
- Abbr
- Qaaa
- Count
- 0
Reserved for private use (start) is the initial code point in a range that applications can assign custom characters to without standardized meaning.
- Abbr
- Qabx
- Count
- 0
Reserved for private use (end) is a sentinel value marking the upper boundary of the private use script range.
- Abbr
- Ranj
- Count
- 0
Ranjana is a script used historically for writing Nepali and Maithili, characterized by curved, ornate letterforms often seen in Buddhist manuscripts.
- Abbr
- Rjng
- Count
- 37
Rejang is a script used historically for writing the Rejang language of Sumatra, characterized by its distinctive abugida syllabic system.
- Abbr
- Rohg
- Count
- 50
Hanifi Rohingya is a script used for writing the Rohingya language, with letters encoded for digital text.
- Abbr
- Roro
- Count
- 0
Rongorongo is a script value used for the undeciphered glyphs of Easter Island, encoded in the Unicode standard.
- Abbr
- Runr
- Count
- 86
Runic is a script value identifying characters used for ancient Germanic inscriptions, including Elder and Younger Futhark, with regional variants.
- Abbr
- Samr
- Count
- 61
Samaritan is a script used for the Samaritan Pentateuch, with writing from right to left and roots in ancient Hebrew.
- Abbr
- Sara
- Count
- 0
Sarati is a script value identifying the constructed alphabet used for invented languages, primarily in Tolkien’s fictional works.
- Abbr
- Sarb
- Count
- 32
Old South Arabian is a historical script value used for encoding ancient Yemeni inscriptions, including Sabaean and Minaic texts.
- Abbr
- Saur
- Count
- 82
Saurashtra is a script used historically for the Saurashtra language in Gujarat, India, now revived in modern digital text.
- Abbr
- Seal
- Count
- 11328
- Abbr
- Sgnw
- Count
- 672
SignWriting is a script value indicating a visual notation system for sign languages, using iconic symbols for handshapes, movements, and facial expressions.
- Abbr
- Shaw
- Count
- 48
Shavian is a script value used to encode texts in the phonemic alphabet devised by George Bernard Shaw for English.
- Abbr
- Shrd
- Count
- 104
Sharada is a script used for writing Sanskrit and Kashmiri, primarily in historical inscriptions and manuscripts from northwestern India.
- Abbr
- Shui
- Count
- 0
Shuishu is a script value used to encode the historical Shuishu writing system, primarily for ritual texts of the Shui people.
- Abbr
- Sidd
- Count
- 92
Siddham is an ancient Indian script used to write Sanskrit, primarily in Buddhist texts and esoteric rituals across East Asia.
- Abbr
- Sidt
- Count
- 26
Sidetic is an ancient script from Asia Minor, used for inscriptions, encoded in Unicode for digital representation.
- Abbr
- Sind
- Count
- 69
Khudawadi is a script used for writing the Sindhi language, primarily in the Khudawadi region of India.
- Abbr
- Sinh
- Count
- 111
Sinhala is a script used for writing the Sinhalese language, primarily in Sri Lanka, with distinct character shapes and vowel marks.
- Abbr
- Sogd
- Count
- 42
Sogdian is a script used for the ancient Eastern Iranian language, written right to left with cursive and separated forms.
- Abbr
- Sogo
- Count
- 40
Old Sogdian is a script used for an ancient Eastern Iranian language, written right-to-left, with letters denoting consonants and some vowels.
- Abbr
- Sora
- Count
- 35
Sora Sompeng is a script used for writing the Sora language, primarily spoken in eastern India.
- Abbr
- Soyo
- Count
- 83
Soyombo is a historical script from Mongolia and Tibet, used mainly for Buddhist texts and the Mongolian flag’s national emblem.
- Abbr
- Sund
- Count
- 72
Sundanese is the script value used for writing the Sundanese language, primarily in West Java, Indonesia, with its own distinct characters.
- Abbr
- Sunu
- Count
- 44
Sunuwar is a script used for writing the Sunuwar language, primarily in Nepal, with its own distinct glyphs and characters.
- Abbr
- Sylo
- Count
- 45
Syloti Nagri is a script used for writing the Sylheti language, primarily in northeastern Bangladesh and nearby Indian regions.
- Abbr
- Syrc
- Count
- 88
Syriac is a writing system used for Aramaic dialects, with distinct letter forms for classical, Eastern, and Western variants.
- Abbr
- Syre
- Count
- 0
Syriac (Estrangelo variant) is a script value used for classical Syriac texts, written right to left, distinct from other Syriac forms.
- Abbr
- Syrj
- Count
- 0
Syriac (Western variant) is a script value identifying text written in the West Syriac tradition, used mainly by the Syriac Orthodox Church.
- Abbr
- Syrn
- Count
- 0
Syriac (Eastern variant) is used to write texts in the classical Syriac language, primarily in liturgical and historical contexts.
- Abbr
- Tagb
- Count
- 18
Tagbanwa is a script used for writing the Tagbanwa language, primarily in the Philippines, with its own distinct character shapes and syllabic structure.
- Abbr
- Takr
- Count
- 68
Takri is a script used historically for languages like Dogri, Kangri, and Chamba, now encoded for digital text.
- Abbr
- Tale
- Count
- 35
Tai Le is a script used for writing the Tai Nüa language, primarily in southwestern China and parts of Myanmar.
- Abbr
- Talu
- Count
- 83
New Tai Lue is a script used for writing the Tai Lü language, primarily in southern China and Southeast Asia, with distinct rounded letterforms.
- Abbr
- Taml
- Count
- 123
Tamil is a script used for writing the Tamil language, primarily in southern India and Sri Lanka, with distinct letterforms.
- Abbr
- Tang
- Count
- 7061
Tangut is a script used for the Tangut language, primarily in the Western Xia dynasty, with characters modeled on Chinese but unique in structure.
- Abbr
- Tavt
- Count
- 72
Tai Viet is a script used for writing Tai Dam and related languages, with letters arranged in a distinctive syllabic style.
- Abbr
- Tayo
- Count
- 55
Tai Yo is a script used for writing the Tai Yo language, primarily in Vietnam and Laos, with distinct glyph forms.
- Abbr
- Telu
- Count
- 101
Telugu is a script used primarily for writing the Telugu language, predominantly spoken in the Indian state of Andhra Pradesh and Telangana.
- Abbr
- Teng
- Count
- 0
Tengwar is a script used to write invented languages, like Quenya and Sindarin, with values assigned in Unicode for encoding its letters.
- Abbr
- Tfng
- Count
- 59
Tifinagh is a script used for writing Berber languages, with Unicode assigning characters for its letters and diacritics across two blocks.
- Abbr
- Tglg
- Count
- 23
Tagalog is a script used for writing the Philippine language, with its Unicode characters representing syllables and historical Baybayin forms.
- Abbr
- Thaa
- Count
- 50
Thaana is the script used for writing Dhivehi, the official language of the Maldives, with right‑to‑left orientation and derived from Arabic numerals.
- Abbr
- Thai
- Count
- 86
Thai is a script used for writing the Thai language, characterized by its distinct, rounded letterforms and no spaces between words.
- Abbr
- Tibt
- Count
- 207
Tibetan is the script value for characters used to write the Tibetan language, including its historical and religious texts.
- Abbr
- Tirh
- Count
- 82
Tirhuta is a script used historically for writing the Maithili language, primarily in the Mithila region of Nepal and India.
- Abbr
- Tnsa
- Count
- 89
Tangsa is a script used for writing several Tibeto-Burman languages spoken in northeastern India and adjacent Myanmar, with a distinct set of letters.
- Abbr
- Todr
- Count
- 52
Todhri is a script value used for the historical Albanian Todhri alphabet, distinct from Latin and Greek.
- Abbr
- Tols
- Count
- 54
Tolong Siki is a script used for the Bugis language, primarily written in South Sulawesi, Indonesia, with its own distinct character forms.
- Abbr
- Toto
- Count
- 31
Toto is a script value identifying a specific writing system, used to classify characters by their linguistic origin and historical usage.
- Abbr
- Tutg
- Count
- 80
Tulu Tigalari is a script used historically for writing Tulu and Sanskrit in the Karnataka region of India.
- Abbr
- Ugar
- Count
- 31
Ugaritic is a script used for the ancient Semitic language of Ugarit, written in cuneiform signs.
- Abbr
- Vaii
- Count
- 300
Vai is a script used for the Vai language of Liberia, featuring a unique syllabary with 300 characters.
- Abbr
- Visp
- Count
- 0
Visible Speech is a script value identifying characters used for phonetic notation, distinct from Latin and other writing systems.
- Abbr
- Vith
- Count
- 70
Vithkuqi is a script used historically for writing the Albanian language, featuring distinct letterforms and a right-to-left direction.
- Abbr
- Wara
- Count
- 84
Warang Citi is a script used for writing the Ho language, encoded with its own distinct character set.
- Abbr
- Wcho
- Count
- 59
Wancho is a script used for the Wancho language of northeast India, featuring an alphabetic system with 49 letters.
- Abbr
- Wole
- Count
- 0
Woleai is a writing system used historically for the Woleaian language on the Caroline Islands, now encoded for digital text.
- Abbr
- Xpeo
- Count
- 50
Old Persian is a script used for cuneiform inscriptions of ancient Iran, valued for encoding royal texts from the Achaemenid era.
- Abbr
- Xsux
- Count
- 1393
Cuneiform is a script value representing the ancient Sumerian, Akkadian, and other Mesopotamian writing systems, encoded for digital text.
- Abbr
- Yezi
- Count
- 47
Yezidi is a writing system used for the Kurdish Yezidi language, encoded with its own script in digital text.
- Abbr
- Yiii
- Count
- 1220
Yi is a script value used for the Yi syllabary and logograms, primarily for the Nuosu language in southwestern China.
- Abbr
- Zanb
- Count
- 72
Zanabazar Square is a script value used for the historical Mongolian script, encoded for texts in Buddhist and secular documents.
- Abbr
- Zinh
- Count
- 695
Inherited is used when a character’s script is determined by surrounding characters, like combining marks or punctuation.
- Abbr
- Zmth
- Count
- 0
Mathematical notation is used to identify which mathematical script, like Latin or Fraktur, a character belongs to for semantic rendering.
- Abbr
- Zsye
- Count
- 0
Symbols (Emoji variant) is a script value used for characters that represent emoji, pictographs, and related symbolic glyphs without a distinct written script.
- Abbr
- Zsym
- Count
- 0
Symbols is a script value for characters used in symbolic, non-alphabetic writing systems, including math, currency, and technical notation.
- Abbr
- Zxxx
- Count
- 0
Code for unwritten documents is a special script value used to mark text in unknown, invented, or undeciphered writing systems.
- Abbr
- Zyyy
- Count
- 9276
Common is a script value used for characters shared across writing systems, like punctuation, digits, and symbols, lacking a specific script identity.
- Abbr
- Zzzz
- Count
- 137468
Unknown is a fallback for characters not assigned to any specific script, covering unassigned, private use, and non-script codepoints.