Unicode 14 is due for release in September. As a preview, I have released ex_unicode version 1.12.0-rc.0. The key features of Unicode 14 are:
-
Add 838 characters, for a total of 144,697 characters. These additions include 5 new scripts, for a total of 159 scripts, as well as 37 new emoji characters.
-
Add support for lesser-used languages and unique written requirements worldwide, including numerous symbols additions. Funds from the Adopt-a-Character program provided support for some of these additions. The new scripts and characters include:
- Toto, used to write the Toto language in India near Bhutan
- Cypro-Minoan, an undeciphered historical script primarily used on the island of Cyprus
- Vithkuqi, an historic script used to write Albanian, and undergoing a modern revival
- Old Uyghur, an historic script used in Central Asia and elsewhere to write Turkic, Chinese, Mongolian, Tibetan, and Arabic languages
- Tangsa, a modern script used to write the Tangsa language, which is spoken in India and Myanmar
- Many Latin additions for extended IPA
- Arabic script additions used to write languages across Africa and in Iran, Pakistan, Malaysia, Indonesia, Java, and Bosnia, and to write honorifics, and additions for Quranic use
- Other character additions support languages of the Philippines, North America, India, and Mongolia
-
Popular symbol additions:
- 37 emoji characters. For complete statistics regarding all emoji as of Unicode 14.0, see Emoji Counts. For more information about emoji additions in version 14.0, including new emoji ZWJ sequences and emoji modifier sequences, see Emoji Recently Added, v14.0.






















