Tabassaran Alphabet at a Glance

  • Tabassaran and Tabasaran are two spellings of the same Northeast Caucasian language — both names refer to the same Lezgic language of southern Dagestan, Russia
  • The language has approximately 126,000 speakers in the Tabasaransky and Khivsky Districts of Dagestan
  • Tabassaran is world-famous among linguists for having up to 48 grammatical cases — one of the largest case inventories of any documented language [1]
  • The Cyrillic alphabet uses the palochka (Ӏ) for ejective consonants and distinctive lateral affricates Цц and Чч found only in the Lezgic branch [2]
  • Tabassaran belongs to the Lezgic branch of the Northeast Caucasian family alongside Lezgi, Agul, and Rutul [3]
  • Tabassaran holds official regional language status in the Republic of Dagestan, supported by school education and local media
  • 7 vowels and 35 consonants form the Tabassaran alphabet — the same script as Tabasaran, since both names denote the same language

Tabassaran Vowel Letters

Tabassaran has 7 vowels — А, Е, И, О, У, Ю, Я — the same vowel set as Tabasaran (both names, same language). The vowel inventory is modest compared to the language's extraordinarily complex case system.

The vowels Ю and Я appear mainly in Russian loanwords. The core native vowels А, И, О, У, Э form the foundation of Tabassaran phonology.

Tabasaran Vowels (Uppercase)

А
[AH]
Е
[YEH]
И
[EE]
О
[OH]
У
[OO]
Ю
[YOO]
Я
[YAH]

Tabasaran Vowels (Lowercase)

а
[ah]
е
[yeh]
и
[ee]
о
[oh]
у
[oo]
ю
[yoo]
я
[yah]

Tabassaran Consonant Letters

Tabassaran has 35 consonants including ejective stops (КӀ, ТӀ, ПӀ, ЦӀ, ЧӀ), uvular sounds (Гъ, Хъ, Кк), pharyngeal fricatives (Гь, Хь), and distinctive lateral affricates (Цц, Чч).

The lateral affricates Цц and Чч — sounds released from the sides of the tongue — are rare globally and particularly characteristic of the Lezgic branch of Northeast Caucasian languages.

Tabasaran Consonants (Uppercase)

Б
[BEH]
В
[VEH]
Г
[GEH]
Гъ
[GH]
Гь
[H]
Д
[DEH]
Ж
[ZHEH]
З
[ZEH]
К
[KAH]
КӀ
[K']
Кк
[QQ]
КкӀ
[Q']
Л
[EHL]
М
[EHM]
Н
[EHN]
П
[PEH]
ПӀ
[P']
Р
[EHR]
С
[EHS]
Т
[TEH]
ТӀ
[T']
Тт
[TH]
Ф
[FEH]
Х
[KH]
Хъ
[QH]
Хь
[HH]
Ц
[TS]
ЦӀ
[TS']
Цц
[TTS]
Ч
[CH]
ЧӀ
[CH']
Чч
[CCH]
Ш
[SH]
Й
[YEH]

Tabasaran Consonants (Lowercase)

б
[beh]
в
[veh]
г
[geh]
гъ
[gh]
гь
[h]
д
[deh]
ж
[zheh]
з
[zeh]
к
[kah]
кӀ
[k']
кк
[qq]
ккӀ
[q']
л
[ehl]
м
[ehm]
н
[ehn]
п
[peh]
пӀ
[p']
р
[ehr]
с
[ehs]
т
[teh]
тӀ
[t']
тт
[th]
ф
[feh]
х
[kh]
хъ
[qh]
хь
[hh]
ц
[ts]
цӀ
[ts']
цц
[tts]
ч
[ch]
чӀ
[ch']
чч
[cch]
ш
[sh]
й
[yeh]

Special Characters

The palochka (Ӏ) is a unique Cyrillic character used in Tabassaran and other Northeast Caucasian languages to mark ejective consonants following the base letter.

The hard sign (ъ) and soft sign (ь) function as consonant modifiers in Tabassaran digraphs for uvular (гъ, хъ) and pharyngeal (гь, хь) sounds, not in their standard Russian grammatical roles.

Ӏ
[palochka]
ъ
[hard sign]
ь
[soft sign]
«
[open guillemet]
»
[close guillemet]

Tabassaran Number Words

Tabassaran uses Arabic numerals (0–9) in modern writing. Native Tabassaran number words descend from Proto-Lezgic, with cognates across related languages.

The number system is base-10 (decimal). Native words like саб (one), губ (two), хуб (three) are distinctively Lezgic in form.

0
[ts'if]
1
[sab]
2
[gub]
3
[xub]
4
[ugub]
5
[xurub]
6
[ragub]
7
[yidib]
8
[muhub]
9
[yicib]
10
[bits'ib]

Complete Tabassaran Alphabet

The complete Tabassaran Cyrillic alphabet in standard order — identical to the Tabasaran alphabet, as both names refer to the same language. Includes 7 vowels, 35 consonants, ejective digraphs, and lateral affricates.

А а
[ah]
Б б
[beh]
В в
[veh]
Г г
[geh]
Гъ гъ
[gh]
Гь гь
[h]
Д д
[deh]
Е е
[yeh]
Ж ж
[zheh]
З з
[zeh]
И и
[ee]
Й й
[yeh]
К к
[kah]
КӀ кӀ
[k']
Кк кк
[qq]
Л л
[ehl]
М м
[ehm]
Н н
[ehn]
О о
[oh]
П п
[peh]
ПӀ пӀ
[p']
Р р
[ehr]
С с
[ehs]
Т т
[teh]
ТӀ тӀ
[t']
Тт тт
[th]
У у
[oo]
Ф ф
[feh]
Х х
[kh]
Хъ хъ
[qh]
Хь хь
[hh]
Ц ц
[ts]
ЦӀ цӀ
[ts']
Цц цц
[tts]
Ч ч
[ch]
ЧӀ чӀ
[ch']
Чч чч
[cch]
Ш ш
[sh]
Ю ю
[yoo]
Я я
[yah]

Frequently Asked Questions (FAQs)

References:

  • [1] Dryer, Matthew S. & Haspelmath, Martin (eds.) 2013. "World Atlas of Language Structures Online — Case Syncretism" — WALS chapter on case systems documents Tabasaran as having one of the largest case inventories in the world, sometimes described as having up to 48 grammatical cases. Retrieved from WALS Online — Case Syncretism
  • [2] Unicode Consortium. "Cyrillic Unicode Block (U+0400-U+04FF)". Retrieved from Unicode Cyrillic Block
  • [3] Dryer, Matthew S. & Haspelmath, Martin (eds.) 2013. "World Atlas of Language Structures Online" — WALS Online provides typological data on consonant inventories and other structural features of the world's languages, including Northeast Caucasian languages such as Agul, which are noted for large consonant inventories. Retrieved from WALS Online — Consonant Inventories
Sambhu Raj SinghSambhu Raj Singh · LinkedIn · GitHub · Npm

Updated:


The language with one of the world's most complex case systems...
Explore the Agul Cyrillic alphabet with ejective and pharyngeal consonants...
Explore the Tsez alphabet — an endangered Dagestanian language also called Dido...
Explore the Abaza Cyrillic script with ejective consonants and unique letters...
Explore the Abkhaz Cyrillic alphabet with unique extended letters not found in Russian...