Other letters, digits and punctuation
Other letters
The letters and marks below lie outside the cursive letter inventories of the writing systems described on the preceding pages, but the font covers them nonetheless. In the figures, the letter or mark under discussion is shown in black; a neighbouring letter quoted in an example is shown in gray; and the nirugu that stands for the cursive join is shown in blue.
MONGOLIAN LETTER ALI GALI BALUDA (U+1885) and MONGOLIAN LETTER ALI GALI THREE BALUDA (U+1886). U+1885 is used in the Hudum Ali Gali and Manchu Ali Gali writing to mark prodelision — the elision of a word-initial vowel — and attested in the Tongwen Yuntong. It derives from U+0F85 TIBETAN MARK PALUTA and corresponds to U+093D DEVANAGARI SIGN AVAGRAHA. U+1886 is equivalent to three U+1885 marks in a row, just as three Tibetan U+0F85 TIBETAN MARK PALUTA or three Devanagari U+093D DEVANAGARI SIGN AVAGRAHA would be written together.
| BALUDA | THREE BALUDA |
|---|---|
| | |
| U+1889, U+1885, U+1826 | U+1889, U+1886, U+1826 |
MONGOLIAN LETTER ALI GALI ANUSVARA ONE (U+1880). This letter marks the anusvara. Its plain form is based on U+0F83 TIBETAN SIGN SNA LDAN and its FVS1 form on U+0F7E TIBETAN SIGN RJES SU NGA RO; c.f. U+0901 DEVANAGARI SIGN CANDRABINDU and U+0902 DEVANAGARI SIGN ANUSVARA.
| isol | isol with FVS1 |
|---|---|
| | |
| U+1880 | U+1880, U+180B |
MONGOLIAN LETTER ALI GALI VISARGA ONE (U+1881). This letter marks the visarga. Its plain form is based on U+0F7F TIBETAN SIGN RNAM BCAD; c.f. U+0903 DEVANAGARI SIGN VISARGA.
| isol | isol with FVS1 |
|---|---|
| | |
| U+1881 | U+1881, U+180B |
MONGOLIAN LETTER ALI GALI DAMARU (U+1882), UBADAMA (U+1883), and INVERTED UBADAMA (U+1884). These letters are used to transliterate the Vedic Sanskrit sounds jihvāmūlīya and upadhmānīya. DAMARU is based on U+0F88 TIBETAN SIGN LCE TSA CAN and corresponds to U+1CF5 VEDIC SIGN JIHVAMULIYA; UBADAMA and INVERTED UBADAMA are based on U+0F8C TIBETAN SIGN INVERTED MCHU CAN and U+0F89 TIBETAN SIGN MCHU CAN respectively and correspond to U+1CF6 VEDIC SIGN UPADHMANIYA.
| DAMARU (isol) | UBADAMA (isol) | INVERTED UBADAMA (isol) |
|---|---|---|
| | | |
| U+1882 | U+1883 | U+1884 |
MONGOLIAN LETTER ALI GALI DAGALGA (U+18A9). This letter is used in the Hudum Ali Gali to mark the dagalga — a syllable-final consonant cluster. It derives from U+0F84 TIBETAN MARK HALANTA and corresponds to U+094D DEVANAGARI SIGN VIRAMA. It is used in combination with the preceding letter to indicate that the following vowel is not pronounced.
| DAGALGA |
|---|
| |
| U+1892, U+18A9, U+1820 |
Digits
Ten Mongolian digits are encoded in the Mongolian block, U+1810–U+1819, named MONGOLIAN DIGIT ZERO through MONGOLIAN DIGIT NINE. In the Unicode Character Database they are decimal digits (general category Nd) of the Mongolian script, and they are the digits used by Hudum and Todo. They derive from the Tibetan digits U+0F20–U+0F29, whose glyphs the vertical script has turned for its own orientation. Manchu and Sibe use no script-specific digits: European digits will be used in Sibe books. See Section 13.5, “Mongolian” and the code chart for the Mongolian block.
Mongolian digits from zero to nine are not used in modern China: Mongolian (Hudum) text printed in China today writes numbers with European digits, and the Mongolian digits appear in the older literature. The Hudum and Todo literatures use these digits; they also appear a bit different between Hudum and Todo, but they are still in one system (the same ten code points).
The Mongolian digits are employed in modern Mongolia. They appear on Mongolian banknotes.
The Mongolian digits behave identically to European digits: they are decimal digits with the same numeric values, and they are sorted and processed by those values; only their shapes differ. In vertical text, they are set in either of the two ways that are common for the European digits. In the rotated layout, the digits are rotated, clockwise by 90° as the European digits are in vertical text, so that each digit runs in the direction of the column. In the upright layout, the digits are not rotated, and the digits of a whole number are set from left to right across the column rather than from top to bottom; such a horizontal run set upright within a vertical line is called “horizontal in vertical.” In horizontal text, the digits are simply upright.
Punctuation
| Name | Code point | Written form | Hudum with Ali Gali | Todo with Ali Gali | Sibe | Manchu with Ali Gali |
|---|---|---|---|---|---|---|
| MONGOLIAN BIRGA | U+1800 | | ✔ | |||
| MONGOLIAN BIRGA WITH ORNAMENT | U+11660 | | ✔ | ✔ | ||
| MONGOLIAN BIRGA WITH DOUBLE ORNAMENT | U+11664 | | ✔ | |||
| MONGOLIAN DOUBLE BIRGA WITH ORNAMENT | U+11662 | | ✔ | |||
| MONGOLIAN TRIPLE BIRGA WITH ORNAMENT | U+11663 | | ✔ | |||
| MONGOLIAN ROTATED BIRGA | U+11661 | | ✔ | |||
| MONGOLIAN ROTATED BIRGA WITH ORNAMENT | U+11665 | | ✔ | |||
| MONGOLIAN ROTATED BIRGA WITH DOUBLE ORNAMENT | U+11666 | | ✔ | ✔ | ||
| MONGOLIAN INVERTED BIRGA | U+11667 | | ✔ | |||
| MONGOLIAN INVERTED BIRGA WITH DOUBLE ORNAMENT | U+11668 | | ✔ | |||
| MONGOLIAN SWIRL BIRGA | U+11669 | | ✔ | |||
| MONGOLIAN SWIRL BIRGA WITH ORNAMENT | U+1166A | | ✔ | |||
| MONGOLIAN SWIRL BIRGA WITH DOUBLE ORNAMENT | U+1166B | | ✔ | |||
| MONGOLIAN TURNED SWIRL BIRGA WITH DOUBLE ORNAMENT | U+1166C | | ✔ | |||
| MONGOLIAN ELLIPSIS | U+1801 | | ✔ | ✔ | ✔ | |
| MONGOLIAN COMMA | U+1802 | | ✔ | ✔ | ||
| MONGOLIAN FULL STOP | U+1803 | | ✔ | ✔ | ||
| MONGOLIAN COLON | U+1804 | | ✔ | ✔ | ||
| MONGOLIAN FOUR DOTS | U+1805 | | ✔ | ✔ | ||
| MONGOLIAN TODO SOFT HYPHEN | U+1806 | | ✔ | |||
| MONGOLIAN MANCHU COMMA | U+1808 | | ✔ | |||
| MONGOLIAN MANCHU FULL STOP | U+1809 | | ✔ | |||
| MIDDLE DOT | U+00B7 | | ✔ | ✔ | ✔ | |
| QUESTION EXCLAMATION MARK | U+2048 | | ✔ | ✔ | ||
| EXCLAMATION QUESTION MARK | U+2049 | | ✔ | ✔ | ||
| FULLWIDTH SEMICOLON | U+FF1B | | ✔ | ✔ | ✔ | |
| FULLWIDTH EXCLAMATION MARK | U+FF01 | | ✔ | ✔ | ✔ | |
| FULLWIDTH QUESTION MARK | U+FF1F | | ✔ | ✔ | ✔ | |
| EM DASH | U+2014 | | ✔ | ✔ | ✔ | |
| HORIZONTAL ELLIPSIS | U+2026 | | ✔ | |||
| FULLWIDTH COMMA | U+FF0C | | ✔ | ✔ | ||
| IDEOGRAPHIC COMMA | U+3001 | | ✔ | ✔ | ||
| IDEOGRAPHIC FULL STOP | U+3002 | | ✔ | ✔ | ||
| FULLWIDTH COLON | U+FF1A | | ✔ | |||
| LEFT SINGLE QUOTATION MARK | U+2018 | | ✔ | |||
| RIGHT SINGLE QUOTATION MARK | U+2019 | | ✔ | |||
| LEFT DOUBLE QUOTATION MARK | U+201C | | ✔ | |||
| RIGHT DOUBLE QUOTATION MARK | U+201D | | ✔ | |||
| FULLWIDTH LEFT PARENTHESIS | U+FF08 | | ✔ | ✔ | ✔ | |
| FULLWIDTH RIGHT PARENTHESIS | U+FF09 | | ✔ | ✔ | ✔ | |
| LEFT DOUBLE ANGLE BRACKET | U+300A | | ✔ | ✔ | ✔ | |
| RIGHT DOUBLE ANGLE BRACKET | U+300B | | ✔ | ✔ | ✔ | |
| LEFT ANGLE BRACKET | U+3008 | | ✔ | ✔ | ✔ | |
| RIGHT ANGLE BRACKET | U+3009 | | ✔ | ✔ | ✔ | |
| FULLWIDTH LEFT SQUARE BRACKET | U+FF3B | | ✔ | ✔ | ✔ | |
| FULLWIDTH RIGHT SQUARE BRACKET | U+FF3D | | ✔ | ✔ | ✔ | |
| LEFT CORNER BRACKET | U+300C | | ✔ | |||
| RIGHT CORNER BRACKET | U+300D | | ✔ | |||
| LEFT WHITE CORNER BRACKET | U+300E | | ✔ | |||
| RIGHT WHITE CORNER BRACKET | U+300F | | ✔ |
Note that many punctuation marks are visually set with a space on both sides, but none of that space is built into the glyph: no punctuation glyph carries any space in its own advance, on either side. The space before a mark is not an actual space character — if a plain space were typed before the mark, a line could break there and move the mark to the beginning of the following line, and a line must not begin with that punctuation mark — so it is set with a non-breaking space U+00A0, which permits no line break. The space after a mark, by contrast, is a real breaking space character (U+0020) that still has to be typed. In short: to open the space before a mark, insert U+00A0 before it; to open the space after a mark, insert U+0020 after it.