Skip to content
This is the editor’s draft of the Mongolian UTN.

Other letters, digits and punctuation

Other letters

The letters and marks below lie outside the cursive letter inventories of the writing systems described on the preceding pages, but the font covers them nonetheless. In the figures, the letter or mark under discussion is shown in black; a neighbouring letter quoted in an example is shown in gray; and the nirugu that stands for the cursive join is shown in blue.

MONGOLIAN LETTER ALI GALI BALUDA (U+1885) and MONGOLIAN LETTER ALI GALI THREE BALUDA (U+1886). U+1885 is used in the Hudum Ali Gali and Manchu Ali Gali writing to mark prodelision — the elision of a word-initial vowel — and attested in the Tongwen Yuntong. It derives from U+0F85 TIBETAN MARK PALUTA and corresponds to U+093D DEVANAGARI SIGN AVAGRAHA. U+1886 is equivalent to three U+1885 marks in a row, just as three Tibetan U+0F85 TIBETAN MARK PALUTA or three Devanagari U+093D DEVANAGARI SIGN AVAGRAHA would be written together.

BALUDA THREE BALUDA
U+1889, U+1885, U+1826 U+1889, U+1886, U+1826

MONGOLIAN LETTER ALI GALI ANUSVARA ONE (U+1880). This letter marks the anusvara. Its plain form is based on U+0F83 TIBETAN SIGN SNA LDAN and its FVS1 form on U+0F7E TIBETAN SIGN RJES SU NGA RO; c.f. U+0901 DEVANAGARI SIGN CANDRABINDU and U+0902 DEVANAGARI SIGN ANUSVARA.

isol isol with FVS1
U+1880 U+1880, U+180B

MONGOLIAN LETTER ALI GALI VISARGA ONE (U+1881). This letter marks the visarga. Its plain form is based on U+0F7F TIBETAN SIGN RNAM BCAD; c.f. U+0903 DEVANAGARI SIGN VISARGA.

isol isol with FVS1
U+1881 U+1881, U+180B

MONGOLIAN LETTER ALI GALI DAMARU (U+1882), UBADAMA (U+1883), and INVERTED UBADAMA (U+1884). These letters are used to transliterate the Vedic Sanskrit sounds jihvāmūlīya and upadhmānīya. DAMARU is based on U+0F88 TIBETAN SIGN LCE TSA CAN and corresponds to U+1CF5 VEDIC SIGN JIHVAMULIYA; UBADAMA and INVERTED UBADAMA are based on U+0F8C TIBETAN SIGN INVERTED MCHU CAN and U+0F89 TIBETAN SIGN MCHU CAN respectively and correspond to U+1CF6 VEDIC SIGN UPADHMANIYA.

DAMARU (isol) UBADAMA (isol) INVERTED UBADAMA (isol)
U+1882 U+1883 U+1884

MONGOLIAN LETTER ALI GALI DAGALGA (U+18A9). This letter is used in the Hudum Ali Gali to mark the dagalga — a syllable-final consonant cluster. It derives from U+0F84 TIBETAN MARK HALANTA and corresponds to U+094D DEVANAGARI SIGN VIRAMA. It is used in combination with the preceding letter to indicate that the following vowel is not pronounced.

DAGALGA
U+1892, U+18A9, U+1820

Digits

Ten Mongolian digits are encoded in the Mongolian block, U+1810–U+1819, named MONGOLIAN DIGIT ZERO through MONGOLIAN DIGIT NINE. In the Unicode Character Database they are decimal digits (general category Nd) of the Mongolian script, and they are the digits used by Hudum and Todo. They derive from the Tibetan digits U+0F20–U+0F29, whose glyphs the vertical script has turned for its own orientation. Manchu and Sibe use no script-specific digits: European digits will be used in Sibe books. See Section 13.5, “Mongolian” and the code chart for the Mongolian block.

Mongolian digits from zero to nine are not used in modern China: Mongolian (Hudum) text printed in China today writes numbers with European digits, and the Mongolian digits appear in the older literature. The Hudum and Todo literatures use these digits; they also appear a bit different between Hudum and Todo, but they are still in one system (the same ten code points).

The Mongolian digits are employed in modern Mongolia. They appear on Mongolian banknotes.

The Mongolian digits behave identically to European digits: they are decimal digits with the same numeric values, and they are sorted and processed by those values; only their shapes differ. In vertical text, they are set in either of the two ways that are common for the European digits. In the rotated layout, the digits are rotated, clockwise by 90° as the European digits are in vertical text, so that each digit runs in the direction of the column. In the upright layout, the digits are not rotated, and the digits of a whole number are set from left to right across the column rather than from top to bottom; such a horizontal run set upright within a vertical line is called “horizontal in vertical.” In horizontal text, the digits are simply upright.

Punctuation

NameCode pointWritten formHudum with Ali GaliTodo with Ali GaliSibeManchu with Ali Gali
MONGOLIAN BIRGAU+1800
MONGOLIAN BIRGA WITH ORNAMENTU+11660
MONGOLIAN BIRGA WITH DOUBLE ORNAMENTU+11664
MONGOLIAN DOUBLE BIRGA WITH ORNAMENTU+11662
MONGOLIAN TRIPLE BIRGA WITH ORNAMENTU+11663
MONGOLIAN ROTATED BIRGAU+11661
MONGOLIAN ROTATED BIRGA WITH ORNAMENTU+11665
MONGOLIAN ROTATED BIRGA WITH DOUBLE ORNAMENTU+11666
MONGOLIAN INVERTED BIRGAU+11667
MONGOLIAN INVERTED BIRGA WITH DOUBLE ORNAMENTU+11668
MONGOLIAN SWIRL BIRGAU+11669
MONGOLIAN SWIRL BIRGA WITH ORNAMENTU+1166A
MONGOLIAN SWIRL BIRGA WITH DOUBLE ORNAMENTU+1166B
MONGOLIAN TURNED SWIRL BIRGA WITH DOUBLE ORNAMENTU+1166C
MONGOLIAN ELLIPSISU+1801
MONGOLIAN COMMAU+1802
MONGOLIAN FULL STOPU+1803
MONGOLIAN COLONU+1804
MONGOLIAN FOUR DOTSU+1805
MONGOLIAN TODO SOFT HYPHENU+1806
MONGOLIAN MANCHU COMMAU+1808
MONGOLIAN MANCHU FULL STOPU+1809
MIDDLE DOTU+00B7
QUESTION EXCLAMATION MARKU+2048
EXCLAMATION QUESTION MARKU+2049
FULLWIDTH SEMICOLONU+FF1B
FULLWIDTH EXCLAMATION MARKU+FF01
FULLWIDTH QUESTION MARKU+FF1F
EM DASHU+2014
HORIZONTAL ELLIPSISU+2026
FULLWIDTH COMMAU+FF0C
IDEOGRAPHIC COMMAU+3001
IDEOGRAPHIC FULL STOPU+3002
FULLWIDTH COLONU+FF1A
LEFT SINGLE QUOTATION MARKU+2018
RIGHT SINGLE QUOTATION MARKU+2019
LEFT DOUBLE QUOTATION MARKU+201C
RIGHT DOUBLE QUOTATION MARKU+201D
FULLWIDTH LEFT PARENTHESISU+FF08
FULLWIDTH RIGHT PARENTHESISU+FF09
LEFT DOUBLE ANGLE BRACKETU+300A
RIGHT DOUBLE ANGLE BRACKETU+300B
LEFT ANGLE BRACKETU+3008
RIGHT ANGLE BRACKETU+3009
FULLWIDTH LEFT SQUARE BRACKETU+FF3B
FULLWIDTH RIGHT SQUARE BRACKETU+FF3D
LEFT CORNER BRACKETU+300C
RIGHT CORNER BRACKETU+300D
LEFT WHITE CORNER BRACKETU+300E
RIGHT WHITE CORNER BRACKETU+300F

Note that many punctuation marks are visually set with a space on both sides, but none of that space is built into the glyph: no punctuation glyph carries any space in its own advance, on either side. The space before a mark is not an actual space character — if a plain space were typed before the mark, a line could break there and move the mark to the beginning of the following line, and a line must not begin with that punctuation mark — so it is set with a non-breaking space U+00A0, which permits no line break. The space after a mark, by contrast, is a real breaking space character (U+0020) that still has to be typed. In short: to open the space before a mark, insert U+00A0 before it; to open the space after a mark, insert U+0020 after it.