Tabasaran Alphabet at a Glance

  • 7 vowels and 35 consonants form the Tabasaran Cyrillic alphabet, with ejective sounds, uvular stops, and lateral affricates unique to Northeast Caucasian languages
  • Spoken by approximately 126,000 people in the Tabasaransky and Khivsky Districts of southern Dagestan — one of the largest Northeast Caucasian language communities
  • Tabasaran is famous among linguists for its extraordinary case system, with some analyses identifying up to 48 grammatical cases — one of the largest in any known language [1]
  • The language uses the palochka (Ӏ) to mark ejective consonants and double-letter digraphs (Цц, Чч) for lateral affricates unique to Tabasaran phonology [2]
  • Tabasaran belongs to the Lezgic branch of the Northeast Caucasian family, closely related to Agul, Lezgi, and Rutul [3]
  • Tabasaran has official regional language status in the Republic of Dagestan, alongside Russian and 13 other languages of the region
  • "Tabasaran" and "Tabassaran" are two common spellings of the same language — both refer to the same Lezgic language and writing system

Tabasaran Vowel Letters

Tabasaran has 7 vowels — А, Е, И, О, У, Ю, Я — a modest vowel system typical of Northeast Caucasian languages. The letters Ю and Я appear mainly in Russian loanwords.

Despite its modest vowel inventory, Tabasaran compensates with an extraordinarily rich nominal morphology — the 48-case system encodes spatial, directional, and relational meanings through case suffixes on nouns.

Tabasaran Vowels (Uppercase)

А
[AH]
Е
[YEH]
И
[EE]
О
[OH]
У
[OO]
Ю
[YOO]
Я
[YAH]

Tabasaran Vowels (Lowercase)

а
[ah]
е
[yeh]
и
[ee]
о
[oh]
у
[oo]
ю
[yoo]
я
[yah]

Tabasaran Consonant Letters

Tabasaran has 35 consonants including ejective stops (КӀ, ТӀ, ПӀ, ЦӀ, ЧӀ), uvular sounds (Гъ, Хъ, Кк, КкӀ), pharyngeal fricatives (Гь, Хь), and distinctive lateral affricates (Цц, Чч) rare outside Northeast Caucasia.

The lateral affricates (Цц and Чч) are among the most distinctive features of Tabasaran — sounds produced with the tongue touching the palate while air escapes from the sides, absent from most other language families.

Tabasaran Consonants (Uppercase)

Б
[BEH]
В
[VEH]
Г
[GEH]
Гъ
[GH]
Гь
[H]
Д
[DEH]
Ж
[ZHEH]
З
[ZEH]
К
[KAH]
КӀ
[K']
Кк
[QQ]
КкӀ
[Q']
Л
[EHL]
М
[EHM]
Н
[EHN]
П
[PEH]
ПӀ
[P']
Р
[EHR]
С
[EHS]
Т
[TEH]
ТӀ
[T']
Тт
[TH]
Ф
[FEH]
Х
[KH]
Хъ
[QH]
Хь
[HH]
Ц
[TS]
ЦӀ
[TS']
Цц
[TTS]
Ч
[CH]
ЧӀ
[CH']
Чч
[CCH]
Ш
[SH]
Й
[YEH]

Tabasaran Consonants (Lowercase)

б
[beh]
в
[veh]
г
[geh]
гъ
[gh]
гь
[h]
д
[deh]
ж
[zheh]
з
[zeh]
к
[kah]
кӀ
[k']
кк
[qq]
ккӀ
[q']
л
[ehl]
м
[ehm]
н
[ehn]
п
[peh]
пӀ
[p']
р
[ehr]
с
[ehs]
т
[teh]
тӀ
[t']
тт
[th]
ф
[feh]
х
[kh]
хъ
[qh]
хь
[hh]
ц
[ts]
цӀ
[ts']
цц
[tts]
ч
[ch]
чӀ
[ch']
чч
[cch]
ш
[sh]
й
[yeh]

Special Characters

The palochka (Ӏ) marks ejective consonants in Tabasaran; the hard sign (ъ) and soft sign (ь) form digraphs for uvular (гъ, хъ) and pharyngeal (гь, хь) sounds.

These modifier signs work differently than in Russian — in Tabasaran they create distinct consonant categories rather than serving grammatical vowel-softening roles.

Ӏ
[palochka]
ъ
[hard sign]
ь
[soft sign]
«
[open guillemet]
»
[close guillemet]

Tabasaran Number Words

Tabasaran uses Arabic numerals (0–9) for modern writing. The native Tabasaran number words reflect the language's Lezgic heritage.

Native numerals share cognates with related Lezgic languages (Agul: сад/sab "one", губ/gub "two"), revealing the shared Proto-Lezgic counting system.

0
[ts'if]
1
[sab]
2
[gub]
3
[xub]
4
[ugub]
5
[xurub]
6
[ragub]
7
[yidib]
8
[muhub]
9
[yicib]
10
[bits'ib]

Complete Tabasaran Alphabet

The complete Tabasaran Cyrillic alphabet in standard order — 7 vowels and 35 consonants, including ejective digraphs (КӀ, ТӀ), uvular forms (Гъ, Хъ, Кк), pharyngeals (Гь, Хь), and lateral affricates (Цц, Чч).

А а
[ah]
Б б
[beh]
В в
[veh]
Г г
[geh]
Гъ гъ
[gh]
Гь гь
[h]
Д д
[deh]
Е е
[yeh]
Ж ж
[zheh]
З з
[zeh]
И и
[ee]
Й й
[yeh]
К к
[kah]
КӀ кӀ
[k']
Кк кк
[qq]
Л л
[ehl]
М м
[ehm]
Н н
[ehn]
О о
[oh]
П п
[peh]
ПӀ пӀ
[p']
Р р
[ehr]
С с
[ehs]
Т т
[teh]
ТӀ тӀ
[t']
Тт тт
[th]
У у
[oo]
Ф ф
[feh]
Х х
[kh]
Хъ хъ
[qh]
Хь хь
[hh]
Ц ц
[ts]
ЦӀ цӀ
[ts']
Цц цц
[tts]
Ч ч
[ch]
ЧӀ чӀ
[ch']
Чч чч
[cch]
Ш ш
[sh]
Ю ю
[yoo]
Я я
[yah]

Frequently Asked Questions (FAQs)

References:

  • [1] Dryer, Matthew S. & Haspelmath, Martin (eds.) 2013. "World Atlas of Language Structures Online — Case Syncretism" — WALS chapter on case systems documents Tabasaran as having one of the largest case inventories in the world, sometimes described as having up to 48 grammatical cases. Retrieved from WALS Online — Case Syncretism
  • [2] Unicode Consortium. "Cyrillic Unicode Block (U+0400-U+04FF)". Retrieved from Unicode Cyrillic Block
  • [3] Dryer, Matthew S. & Haspelmath, Martin (eds.) 2013. "World Atlas of Language Structures Online" — WALS Online provides typological data on consonant inventories and other structural features of the world's languages, including Northeast Caucasian languages such as Agul, which are noted for large consonant inventories. Retrieved from WALS Online — Consonant Inventories
Sambhu Raj SinghSambhu Raj Singh · LinkedIn · GitHub · Npm

Updated:


Explore the Agul Cyrillic alphabet with ejective and pharyngeal consonants...
Explore the Tabassaran Cyrillic alphabet — alternate name for Tabasaran...
Explore the Tsez alphabet — an endangered Dagestanian language also called Dido...
Explore the Abaza Cyrillic script with ejective consonants and unique letters...
Explore the Abkhaz Cyrillic alphabet with unique extended letters not found in Russian...