Tibetan script explained

Tibetan
Type:	Abugida
Time:	–present
Fam1:	Egyptian
Fam2:	Proto-Sinaitic
Fam3:	Phoenician
Fam4:	Aramaic
Fam5:	Brahmi
Fam6:	Gupta^[1] ^[2]
Sisters:	Meitei,^[3] Sharada, Siddham, Kalinga, Bhaiksuki
Children:	Lepcha Khema Phagspa Marchen Tamyig
Sample:	Om Mani Padme Hum mantra.svg
Caption:	The mantra "Om mani padme hum"
Unicode:	U+0F00 - U+0FFF Final Accepted Script Proposal of the First Usable Edition (3.0)
Iso15924:	Tibt

The Tibetan script is a segmental writing system, or abugida, derived from Brahmic scripts and Gupta script, and used to write certain Tibetic languages, including Tibetan, Dzongkha, Sikkimese, Ladakhi, Jirel and Balti. It was originally developed by Tibetan minister Thonmi Sambhota for King Songtsen Gampo.^[4] ^[5]

The Tibetan script has also been used for some non-Tibetic languages in close cultural contact with Tibet, such as Thakali,^[6] Nepali^[7] and Old Turkic. The printed form is called uchen script while the hand-written cursive form used in everyday writing is called umê script. This writing system is used across the Himalayas and Tibet.

The script is closely linked to a broad ethnic Tibetan identity, spanning across areas in India, Nepal, Bhutan and Tibet.^[8] The Tibetan script is of Brahmic origin from the Gupta script and is ancestral to scripts such as Lepcha,^[9] Marchen and the multilingual ʼPhags-pa script, and is also closely related to Meitei.^[3]

History

According to Tibetan historiography, the Tibetan script was developed during the reign of King Songtsen Gampo by his minister Thonmi Sambhota, who was sent to India with 16 other students to study Buddhism along with Sanskrit and written languages. They developed the Tibetan script from the Gupta script^[10] while at the Pabonka Hermitage.

This occurred, towards the beginning of the king's reign. There were 21 Sutra texts held by the King which were afterward translated. In the first half of the 7th century, the Tibetan script was used for the codification of these sacred Buddhist texts,^[11] ^[12] for written civil laws, and for a Tibetan Constitution.

A contemporary academic suggests that the script was instead developed in the second half of the 11th century.^[13] New research and writings also suggest that there were one or more Tibetan scripts in use prior to the introduction of the script by Songtsen Gampo and Thonmi Sambhota. The incomplete Dunhuang manuscripts are their key evidence for their hypothesis,^[14] while the few discovered and recorded Old Tibetan Annals manuscripts date from 650 and therefore post-date the c. 620 date of development of the original Tibetan script.

Three orthographic standardisations were developed. The most important, an official orthography aimed to facilitate the translation of Buddhist scriptures emerged during the early 9th century. Standard orthography has not been altered since then, while the spoken language has changed by, for example, losing complex consonant clusters. As a result, in all modern Tibetan dialects and in particular in the Standard Tibetan of Lhasa, there is a great divergence between current spelling, which still reflects the 9th-century spoken Tibetan, and current pronunciation. This divergence is the basis of an argument in favour of spelling reform, to write Tibetan as it is pronounced; for example, writing Kagyu instead of Bka'-rgyud.^[15]

The nomadic Amdo Tibetan and the western dialects of the Ladakhi language, as well as the Balti language, come very close to the Old Tibetan spellings.^[13] Despite that, the grammar of these dialectical varieties has considerably changed. To write the modern varieties according to the orthography and grammar of Classical Tibetan would be similar to writing Italian according to Latin orthography, or to writing Hindi according to Sanskrit orthogrophy.^[13] However, modern Buddhist practitioners in the Indian subcontinent state that the classical orthography should not be altered even when used for lay purposes. This became an obstacle for many modern Tibetic languages wishing to modernize or to introduce a written tradition. Amdo Tibetan was one of a few examples where Buddhist practitioners initiated a spelling reform.^[13] A spelling reform of the Ladakhi language was controversial in part because it was first initiated by Christian missionaries.^[13]

Description

Basic alphabet

In the Tibetan script, the syllables are written from left to right. Syllables are separated by a tsek (་); since many Tibetan words are monosyllabic, this mark often functions almost as a space. Spaces are not used to divide words.^[16]

The Tibetan alphabet has thirty basic letters, sometimes known as "radicals", for consonants. As in other Indic scripts, each consonant letter assumes an inherent vowel; in the Tibetan script it is /a/. The letter is also the base for dependent vowel marks.

Although some Tibetan dialects are tonal, the language had no tone at the time of the script's invention, and there are no dedicated symbols for tone. However, since tones developed from segmental features, they can usually be correctly predicted by the archaic spelling of Tibetan words.

	Unaspirated high	Aspirated medium	Voiced low	Nasal low
	Letter	IPA	Letter	IPA	Letter	IPA	Letter	IPA
*Guttural*		pronounced as //ka//		pronounced as //kʰa//		pronounced as //ɡa//		pronounced as //ŋa//
*Palatal*		pronounced as //tʃa//		pronounced as //tʃʰa//		pronounced as //dʒa//		pronounced as //ɲa//
*Dental*		pronounced as //ta//		pronounced as //tʰa//		pronounced as //da//		pronounced as //na//
*Labial*		pronounced as //pa//		pronounced as //pʰa//		pronounced as //ba//		pronounced as //ma//
*Dental*		pronounced as //tsa//		pronounced as //tsʰa//		pronounced as //dza//		pronounced as //wa//
*low*		pronounced as //ʒa//		pronounced as //za//		pronounced as //ɦa//^[17]		pronounced as //ja//
*medium*		pronounced as //ra//		pronounced as //la//		pronounced as //ʃa//		pronounced as //sa//
*high*		pronounced as //ha//		pronounced as //a//

Consonant clusters

One aspect of the Tibetan script is that the consonants can be written either as radicals or they can be written in other forms, such as subscript and superscript forming consonant clusters.

To understand how this works, one can look at the radical /ka/ and see what happens when it becomes /kra/ or /rka/ (pronounced /ka/). In both cases, the symbol for /ka/ is used, but when the /ra/ is in the middle of the consonant and vowel, it is added as a subscript. On the other hand, when the /ra/ comes before the consonant and vowel, it is added as a superscript. /ra/ actually changes form when it is above most other consonants, thus rka. However, an exception to this is the cluster /ɲa/. Similarly, the consonants /ra/, and /ja/ change form when they are beneath other consonants, thus /ʈ ~ ʈʂa/; /ca/.

Besides being written as subscripts and superscripts, some consonants can also be placed in prescript, postscript, or post-postscript positions. For instance, the consonants /kʰa/, /tʰa/, /pʰa/, /ma/ and /a/ can be used in the prescript position to the left of other radicals, while the position after a radical (the postscript position), can be held by the ten consonants /kʰa/, /na/, /pʰa/, /tʰa/, /ma/, /a/, /ra/, /ŋa/, /sa/, and /la/. The third position, the post-postscript position is solely for the consonants /tʰa/ and /sa/.

Head letters

The head (in Tibetan, Wylie: mgo) letter, or superscript, position above a radical is reserved for the consonants /ra/, /la/, and /sa/.

When /ra/, /la/, and /sa/ are in superscript position with /ka/, /t͡ʃa/, /ta/, /pa/ and /t͡sa/, there are no changes to their sounds in Lhasa Tibetan, for example:
- /ka/, /ta/, /pa/, /t͡sa/
- /ka/, /t͡ʃa/, /ta/, /pa/,
- /ka/, /ta/, /pa/, /t͡sa/
When /ra/, /la/, and /sa/ are in superscript position with /kʰa/, /t͡ʃʰa/, /tʰa/, /pʰa/ and /t͡sʰa/, they lose their aspiration and become voiced in Lhasa Tibetan, for example:
- /ga/, /d͡ʒa/, /da/, /ba/, /dza/
- /ga/, /d͡ʒa/, /da/, /ba/,
- /ga/, /da/, /ba/
When /ra/, /la/, and /sa/ are in superscript position with the nasal consonants /ŋa/, /ɲa/, /na/ and /ma/, they receive a high tone in Lhasa Tibetan, for example:
- /ŋa/, /ɲa/, /na/, /ma/
- /ŋa/
- /ŋa/, /ɲa/, /na/, /ma/
When /la/ is in superscript position with /ha/, it becomes a voiceless alveolar lateral approximant in Lhasa Tibetan:
- /l̥a/,

Sub-joined letters

The subscript position under a radical can only be occupied by the consonants /ja/, /ra/, /la/, and /wa/. In this position they are described as (Wylie: btags, IPA: /taʔ/), in Tibetan meaning "hung on/affixed/appended", for example (IPA: /pʰa.ja.taʔ.t͡ʃʰa/), except for, which is simply read as it usually is and has no effect on the pronunciation of the consonant to which it is subjoined, for example (IPA: /ka.wa.suː.ka/).

Vowel marks

The vowels used in the alphabet are /a/, /i/, /u/, /e/, and /o/. While the vowel /a/ is included in each consonant, the other vowels are indicated by marks; thus /ka/, /ki/, /ku/, /ke/, /ko/. The vowels /i/, /e/, and /o/ are placed above consonants as diacritics, while the vowel /u/ is placed underneath consonants. Old Tibetan included a reversed form of the mark for /i/, the gigu 'verso', of uncertain meaning. There is no distinction between long and short vowels in written Tibetan, except in loanwords, especially transcribed from the Sanskrit.

Numerical digits

See main article: Tibetan numerals.

Tibetan numerals
Devanagari numerals	०	१	२	३	४	५	६	७	८	९
Arabic numerals	0	1	2	3	4	5	6	7	8	9
Tibetan fractions
Arabic fractions	-0.5	0.5	1.5	2.5	3.5	4.5	5.5	6.5	7.5	8.5

Punctuation marks

Symbol/ Graphemes	Name	Function
+
	yig mgo	marks beginning of a text, before a headline, front page of a pecha
	gter yig mgo	used in place of the yig mgo in terma texts
	yig mgo a phyed	used in place of the yig mgo in terma texts
	dpe rnying yig mgo	a variant of the yig mgo found in very old Tibetan texts
	bskur yig mgo	list enumerator (Dzongkha)
	tseg	syllable delimiter, also used as a spacer to justify text in pechas
	shad	full stop, comma, or semicolon (marks end of a sentence or clause, and originates from the danda of Indic scripts)
	nyis shad	marks end of a paragraph or topic (cp. pilcrow)
	bzhi shad	marks end of a chapter or entire section
	gsum shad	same as bzhi shad, but used when the preceding character is ཀ or ག
	rin chen spungs shad	replaces shad after single, orphaned syllables, indicating to the reader that the preceding syllable continues from text on the previous line
	tsheg shad	variant of rin chen spungs shad
	nyis tsheg shad	variant of rin chen spungs shad
	sbrul shad	marks the start of a new text, often in a collection of texts, separates chapters, and surrounds inserted text
	gter shad	replaces shad and variants thereof in terma texts
	rgya gram shad	sometimes used in place of the yig mgo in terma texts
	che mgo	literally, "big head"—used preceding a reference to the Dalai Lama or the name of another important lama or tulku that demands great respect
	bsdus rtags	repetition
	dzud rtags me long can	caret (indicates text insertion)
	ang khang g.yon 'khor	left roof bracket
	ang khang g.yas 'khor	right roof bracket
	gug rtags g.yon	left bracket
	gug rtags g.yas	right bracket

Extended use

The Tibetan alphabet, when used to write other languages such as Balti, Chinese and Sanskrit, often has additional and/or modified graphemes taken from the basic Tibetan alphabet to represent different sounds.

Extended alphabet

Letter	Used in	Romanization & IPA
+
		qa pronounced as //qa// (/q/)
		ɽa pronounced as //ɽa// (/ɽ/)
		xa pronounced as //χa// (/χ/)
		ɣa pronounced as //ʁa// (/ʁ/)
		fa pronounced as //fa// (/f/)
		va pronounced as //va// (/v/)
		gha pronounced as //ɡʱ//
		jha pronounced as //ɟʱ, d͡ʒʱ//
		ṭa pronounced as //ʈ//
		ṭha pronounced as //ʈʰ//
		ḍa pronounced as //ɖ//
		ḍha pronounced as //ɖʱ//
		ṇa pronounced as //ɳ//
		dha pronounced as //d̪ʱ//
		bha pronounced as //bʱ//
		ṣa pronounced as //ʂ//
		kṣa pronounced as //kʂ//

In Balti, consonants ka, ra are represented by reversing the letters (ka, ra) to give (qa, ɽa).
The Sanskrit retroflex consonants ṭa, ṭha, ḍa, ṇa, ṣa are represented in Tibetan by reversing the letters (ta, tha, da, na, sha) to give (ṭa, ṭha, ḍa, ṇa, ṣa).
It is a classical rule to transliterate Sanskrit ca, cha, ja, jha, to Tibetan (tsa, tsha, dza, dzha), respectively. Nowadays, (ca, cha, ja, jha) can also be used.

Extended vowel marks and modifiers

Vowel Mark	Used in	Romanization & IPA
+
		ā pronounced as //aː//
		ī pronounced as //iː//
		ū pronounced as //uː//
		ai pronounced as //ɐi̯//
		au pronounced as //ɐu̯//
		ṛ /r̩/
		ṝ pronounced as //r̩ː//
		ḷ pronounced as //l̩//
		ḹ pronounced as //l̩ː//
		aṃ pronounced as //◌̃//
		aṃ pronounced as //◌̃//
		aḥ pronounced as //h//

Symbol/ Graphemes	Name	Function
+
	srog med	suppresses the inherent vowel sound
	paluta	used for prolonging vowel sounds

Consonant clusters

In addition to the use of supplementary graphemes, the rules for constructing consonant clusters are amended, allowing any character to occupy the superscript or subscript position, negating the need for the prescript and postscript positions.

Romanization and transliteration

Romanization and transliteration of the Tibetan script is the representation of the Tibetan script in the Latin script. Multiple Romanization and transliteration systems have been created in recent years, but do not fully represent the true phonetic sound. While the Wylie transliteration system is widely used to Romanize Standard Tibetan, others include the Library of Congress system and the IPA-based transliteration (Jacques 2012).

Below is a table with Tibetan letters and different Romanization and transliteration system for each letter, listed below systems are: Wylie transliteration (W), Tibetan pinyin (TP), Dzongkha phonetic (DP), ALA-LC Romanization (A)^[18] and THL Simplified Phonetic Transcription (THL).

Letter	W	TP	DP	A	THL	W	TP	DP	A	THL	W	TP	DP	A	THL	W	TP	DP	A	THL
	ka	g	ka	ka	ka	kha	k	kha	kha	kha	ga*	k*	kha*	ga*	ga*	nga	ng	nga	nga	nga
	ca	j	ca	ca	cha	cha	q	cha	cha	cha	ja*	q*	cha*	ja*	ja*	nya	ny	nya	nya	nya
	ta	d	ta	ta	ta	tha	t	tha	tha	ta	da*	t*	tha*	da*	da*	na	n	na	na	na
	pa	b	pa	pa	pa	pha	p	pha	pha	pa	ba*	p*	pha*	ba*	ba*	ma	m	ma	ma	ma
	tsa	z	tsa	tsa	tsa	tsha	c	tsha	tsha	tsa	dza*	c*	tsha*	dza*	dza*	wa	w	wa	wa	wa
	zha*	x*	sha*	zha*	zha*	za*	s*	sa*	za*	za*	'a	-	a	'a	a	ya	y	ya	ya	ya
	ra	r	ra	ra	ra	la	l	la	la	la	sha	x	sha	sha	sha	sa	s	sa	sa	sa
	ha	h	ha	ha	ha	a	a	a	a	a
– Only in loanwords

Input method and keyboard layout

Tibetan

The first version of Microsoft Windows to support the Tibetan keyboard layout is MS Windows Vista. The layout has been available in Linux since September 2007. In Ubuntu 12.04, one can install Tibetan language support through Dash / Language Support / Install/Remove Languages, the input method can be turned on from Dash / Keyboard Layout, adding Tibetan keyboard layout. The layout applies the similar layout as in Microsoft Windows.

Mac OS-X introduced Tibetan Unicode support with OS-X version 10.5 and later, now with three different keyboard layouts available: Tibetan-Wylie, Tibetan QWERTY and Tibetan-Otani.

Dzongkha

See main article: Dzongkha keyboard layout. The Dzongkha keyboard layout scheme is designed as a simple means for inputting Dzongkha text on computers. This keyboard layout was standardized by the Dzongkha Development Commission (DDC) and the Department of Information Technology (DIT) of the Royal Government of Bhutan in 2000.

It was updated in 2009 to accommodate additional characters added to the Unicode & ISO 10646 standards since the initial version. Since the arrangement of keys essentially follows the usual order of the Dzongkha and Tibetan alphabet, the layout can be quickly learned by anyone familiar with this alphabet. Subjoined (combining) consonants are entered using the Shift key.

The Dzongkha (dz) keyboard layout is included in Microsoft Windows, Android, and most distributions of Linux as part of XFree86.

Unicode

See main article: Tibetan (Unicode block).

Tibetan was originally one of the scripts in the first version of the Unicode Standard in 1991, in the Unicode block U+1000 - U+104F. However, in 1993, in version 1.1, it was removed (the code points it took up would later be used for the Burmese script in version 3.0). The Tibetan script was re-added in July, 1996 with the release of version 2.0.

The Unicode block for Tibetan is U+0F00 - U+0FFF. It includes letters, digits and various punctuation marks and special symbols used in religious texts:

References

Sources

Asher, R. E. ed. The Encyclopedia of Language and Linguistics. Tarrytown, NY: Pergamon Press, 1994. 10 vol.
Beyer, Stephan V. (1993). The Classical Tibetan Language. Reprinted by Delhi: Sri Satguru.
Chamberlain, Bradford Lynn. 2008. Script Selection for Tibetan-related Languages in Multiscriptal Environments. International Journal of the Sociology of Language 192:117–132.
Csoma de Kőrös, Alexander. (1983). A Grammar of the Tibetan Language. Reprinted by Delhi: Sri Satguru.
Csoma de Kőrös, Alexander (1980–1982). Sanskrit-Tibetan-English Vocabulary. 2 vols. Reprinted by Delhi: Sri Satguru.
Daniels, Peter T. and William Bright. The World's Writing Systems. New York: Oxford University Press, 1996.
Das, Sarat Chandra: "The Sacred and Ornamental Characters of Tibet". Journal of the Asiatic Society of Bengal, vol. 57 (1888), pp. 41–48 and 9 plates.
Das, Sarat Chandra. (1996). An Introduction to the Grammar of the Tibetan Language. Reprinted by Delhi: Motilal Banarsidass.
Jacques, Guillaume 2012. A new transcription system for Old and Classical Tibetan, Linguistics of the Tibeto-Burman Area, 35.3:89-96.
Jäschke, Heinrich August. (1989). Tibetan Grammar. Corrected by Sunil Gupta. Reprinted by Delhi: Sri Satguru.

External links

Tibetan Calligraphy —Online guide for writing Tibetan script.
Elements of the Tibetan writing system.
Unicode area U0F00-U0FFF, Tibetan script (162KB)
Encoding Model of the Tibetan Script in the UCS
Digital Tibetan —Online resource for the digitalization of Tibetan.
Tibetan Scripts, Fonts & Related Issues—THDL articles on Unicode font issues; free cross-platform OpenType fonts—Unicode compatible.
Free Tibetan Fonts Project
Ancient Scripts: Tibetan

Notes and References

Book: Language in South Asia . Kachru . Braj B. . Kachru . Yamuna . Sridhar . S. N. . Writing systems of major and minor languages . Daniels . Peter T. . 285–308 . https://doi.org/10.1017/CBO9780511619069.017 . January 2008. 10.1017/CBO9780511619069.017 . 978-0-521-78653-9 .
Book: Masica . Colin . The Indo-Aryan languages . 1993 . 143.
Book: Chelliah . Shobhana Lakshmi . A Grammar of Meithei . "Meithei Mayek is part of the Tibetan group of scripts, which originated from the Gupta Brahmi script" . De Gruyter . 2011 . 355 . 9783110801118 . 2023-03-19 . 2023-04-13 . https://web.archive.org/web/20230413062923/https://books.google.com/books?id=noCHVvu0P8oC . live .
Tibet: A Political History, p. 12. 1967. Tsepon W. D. Shakabpa. Yale University Press, New Haven and London.
The White Annals, pp. 70-73. Gedun Choephel, translated by Samten Norboo. 1978. Tibetan Library and Archives, Dharamsala, H.P., India.
Manzardo . Andrew E. . Impression Management and Economic Growth: The Case of the Thakalis of Dhaulagiri Zone . . 2023-11-20 . 2023-11-20 . https://web.archive.org/web/20231120023404/https://himalaya.socanth.cam.ac.uk/collections/journals/kailash/pdf/kailash_09_01_02.pdf . live .
Book: Karmācārya, Mādhavalāla . Results of the Nepal German Project on High Mountain Archaeology: Ten documents from Mustang in the Nepali language (1667-1975 A.D.) . 2001 . VGH Wissenschaftsverlag . 978-3-88280-061-6 . en.
Chamberlain 2008
Daniels, Peter T. and William Bright. The World's Writing Systems. New York: Oxford University Press, 1996,
Claude Arpi, Glimpses on the Tibet History, Dharamsala: Tibet Museum, 2016.
William Woodville Rockhill,, United States National Museum, page 671
Berzin, Alexander. A Survey of Tibetan History - Reading Notes Taken by Alexander Berzin from Tsepon, W. D. Shakabpa, Tibet: A Political History. New Haven, Yale University Press, 1967: http://studybuddhism.com/web/en/archives/e-books/unpublished_manuscripts/survey_tibetan_history/chapter_1.html .
Book: Zeisler, Bettina . Why Ladakhi must not be written – Being part of the Great Tradition Another kind of global thinking. 2006. Lesser-Known Languages of South Asia. Anju Saxena. Lars Borin. 178.
Book: Phuntsok . Thubten . བོད་ཀྱི་ལོ་རྒྱུས་སྤྱི་དོན་པདྨ་ར་གཱའི་ལྡེ་མིག "A General History of Tibet".
Book: Gamble, R. . Reincarnation in Tibetan Buddhism: The Third Karmapa and the Invention of a Tradition . Oxford University Press . 2018 . 978-0-19-069078-6 . 2024-05-12 . 62.
Book: Chan . A. . Noble . A. . Sounds in Translation: Intersections of Music, Technology and Society . ANU E Press . DOAB Directory of Open Access Books . 2009 . 978-1-921536-55-7 . 2024-05-12 . 146.
Hill . Nathan W. . 2005b . Once more on the letter འ . Linguistics of the Tibeto-Burman Area . 28 . 2 . 111–141 . 2022-06-01 . 2022-06-16 . https://web.archive.org/web/20220616054618/https://eprints.soas.ac.uk/5632/1/once_more_on_the_letter.pdf . live.

Hill . Nathan W. . 2009 . Tibetan <ḥ-> as a plain initial and its place in Old Tibetan phonology . Linguistics of the Tibeto-Burman Area . 32 . 1 . 115–140 . 2022-06-01 . 2022-06-01 . https://web.archive.org/web/20220601181159/https://eprints.soas.ac.uk/7625/1/04-Hill-Dialect-reflexes-of-Tib-v.pdf . live.
Web site: ALA-LC Romanization of Tibetan script (PDF) . . 2017-12-29 . 2018-04-13 . https://web.archive.org/web/20180413080847/http://www.loc.gov/catdir/cpso/romanization/tibetan.pdf . live .

Tibetan script explained

History

Description

Basic alphabet

Consonant clusters

Head letters

Sub-joined letters

Vowel marks

Numerical digits

Punctuation marks

Extended use

Extended alphabet

Extended vowel marks and modifiers

Consonant clusters

Romanization and transliteration

Input method and keyboard layout

Tibetan

Dzongkha

Unicode

See also

References

Sources

External links

Notes and References