Unicode Versions Table

–published versions
VersionDateScriptsCharactersHeadline additions
1.0.0October 1991247,129The first book: Arabic, Cyrillic, Greek, Hebrew, Hangul, Hiragana/Katakana and the other founding scripts
1.0.1June 19922528,327The first 20,902 CJK unified ideographs arrive in one drop
1.1June 19932434,1684,306 new Hangul syllables; Tibetan removed (returned in 2.0); 33 characters reclassified as controls
2.0July 19962538,885The surrogate mechanism opens the 16-plane future; Hangul rebuilt in place (11,172 syllables); Tibetan readmitted
2.1May 19982538,887One headline glyph: the euro sign €
3.0September 19993849,194Cherokee, Geʸez, Khmer, Mongolian, Ogham, runes, Syriac, Thaana, braille patterns and the Yi syllables
3.1March 20014194,140Deseret, Gothic and Old Italic - the first dead-alphabet revivals - plus 42,711 CJK ideographs
3.2March 20024595,156The Philippine four: Buhid, Hanunoo, Tagalog (Baybayin) and Tagbanwa
4.0April 20035296,382Linear B and Cypriot syllabary (writing decrypted twice), Shavian, Osmanya, Ugaritic, Tai Le, Limbu
4.1March 20055997,655Glagolitic, Kharoshthi, Old Persian cuneiform, Tifinagh, Buginese, New Tai Lue, Sylheti Nagari
5.0July 20066499,024Sumero-Akkadian cuneiform lands whole; Balinese, NʻKo, Phags-pa, Phoenician
5.1April 200875100,648Lycian, Lydian, Carian, Vai, Saurashtra, Sundanese, Ol Chiki, Rejang, Cham, Kayah Li, Lepcha - and capital ị
5.2October 200990107,296The Sasanian wave: Avestan, Imperial Aramaic, Inscriptional Pahlavi, Inscriptional Parthian, Old Turkic, Old South Arabian, Samaritan, Egyptian hieroglyphs, Tai Tham, Tai Viet, Javanese, Kaithi, Lisu, Meetei Mayek, Bamum, Vedic extensions
6.0October 201093109,384Brahmi, Batak, Mandaic - and the first emoticons/emoji plus playing cards, traffic signs and alchemy
6.1January 2012100110,116Chakma, Miao, Sharada, Takri, Sora Sompeng and both Meroitic scripts
6.2September 2012100110,117One character: the Turkish lira sign ₺
6.3September 2013100110,122Five bidirectional formatting characters - plumbing, not glyphs
7.0June 2014123112,956The largest script drop ever: Caucasian Albanian, Manichaean, Nabataean, Palmyrene, Psalter Pahlavi, Elbasan, Linear A, Grantha, Modi, Siddham, Tirhuta, Duployan, Mende Kikakui, Warang Citi, Pau Cin Hau and friends
8.0June 2015129120,672Anatolian hieroglyphs, Ahom, Hatran, Multani, Old Hungarian, SignWriting, Cherokee lowercase and emoji skin tones
9.0June 2016135128,172Adlam, Osage, Newa, Marchen, Bhaiksuki - and the Tangut corpus of nearly 6,000 characters
10.0June 2017139136,690Zanabazar Square, Soyombo, Masaram Gondi, Nushu, Hentaigana - and the bitcoin sign ₿
11.0June 2018146137,374Old Sogdian and Sogdian, Hanifi Rohingya, Makasar, Medefaidrin, Dogra, Maya numerals and 145 emoji
12.0March 2019150137,928Elymaic, Nandinagari, Wancho, Nyiakeng Puachue Hmong, small kana extensions and 61 emoji
12.1May 2019150137,929One character: ㋿ Reiwa, the new Japanese era, rushed out on day one
13.0March 2020154143,859Chorasmian, Dhives Akuru, the Khitan small script, Yezidi and 55 emoji
14.0September 2021159144,697Delayed six months by COVID: Vithkuqi, Old Uyghur, Toto, Cypro-Minoan, Tangsa and 37 emoji
15.0September 2022161149,186Kawi and Mundari Nag, 4,192 CJK ideographs and 20 emoji
15.1September 2023161149,813A CJK repertoire extension plus the new CJK Extension H
16.0September 2024168154,998Garay, Sunuwar, Gurung Khema, Kirat Rai, Ol Onal, Todhri, Tulu-Tigalari and 3,995 new Egyptian hieroglyphs
17.0September 2025172159,801Beria Erfe, Tai Yo, Sidetic, Tolong Siki, the Saudi riyal sign ⃁ and 4,316 CJK ideographs
18.0September 2026175172,808Proto-cuneiform, Jurchen and Seal script - the biggest single jump (+13,007) since 3.1
This is the release history of the Unicode character universe, from the 24-script, 7,129-character first edition of October 1991 to version 18.0’s 175 scripts and 172,808 characters (September 2026) - a 24-fold expansion in thirty-five years. The shape of the curve tells three stories: the CJK land-grab (1.0.1’s 20,902 ideographs, 3.1’s 42,711 more), the dead-alphabet revival project that started with Deseret, Gothic and Old Italic in 3.1 and now underwrites most of the historic-script charts on this site, and the emoji era that began in 6.0 (2010) and now forces an annual September release. Wikipedia records the rhythm’s hiccups too: version 14.0 slipped six months to September 2021 because of COVID-19, and 12.1 shipped a single character - ㋿ Reiwa - the day the new Japanese era was announced.
Reading notes: scripts counts are the Wikipedia tabulation (some versions add characters without adding scripts - 6.2 and 6.3 are pure singles), and the character columns are cumulative repertoire size, not new characters; 18.0’s +13,007 jump is the largest single-version gain since the 42,711-drop of 3.1, driven by proto-cuneiform, Jurchen and Seal script. Version 5.2 (October 2009) is this site’s favourite: one release admitted Avestan, Imperial Aramaic, Inscriptional Pahlavi, Inscriptional Parthian, Old Turkic, Old South Arabian and Samaritan - seven charts’ worth of history in one October. Source: verified against Wikipedia’s enumerated-versions table and the Unicode Consortium’s official enumerated versions index. Deep dives on this site: the Samaritan alphabet (a 5.2 release-mate), the Mandaic (6.0), the Manichaean (7.0) and the Old Uyghur (14.0) charts, each tagged with the release that admitted it.

Every app, font and operating system on earth inherits its characters from the release calendar below: thirty-one published versions of the Unicode Standard, from the 7,129-character first edition of October 1991 to version 18.0 (September 2026) with 172,808 characters and 175 scripts - a twenty-four-fold expansion that has run at a steady annual September rhythm since 2010.

The table is more than a changelog; it is a map of digital archaeology. The dead-alphabet revival project - the reason most historic scripts are typeable at all - began in 3.1 (2001) with Deseret, Gothic and Old Italic, and the 5.2 release of October 2009 admitted an entire Sasanian civilization in one drop: Avestan, Imperial Aramaic, Inscriptional Pahlavi, Inscriptional Parthian, Old Turkic, Old South Arabian and Samaritan. The emoji era started quietly in 6.0 (2010) and now anchors the whole schedule.

Reading the table: the Scripts and Characters columns are cumulative counts (the Wikipedia tabulation), so each row shows the state of the character universe after that release - and the Headline column names what each version is actually remembered for, from the euro sign’s solo debut in 2.1 to proto-cuneiform’s 13,007-character landing in 18.0.

How to use

  1. Find a character’s birth year: if you know which script or symbol you care about (say, Cherokee or the euro sign or emoji), scan the Headline column - each script appears exactly once, at the version that first admitted it, so the row gives you the minimum Unicode version any font or platform must support to render it.
  2. Use the version numbers as compatibility keys: software often advertises “Unicode 9.0+” or “Unicode 15.1 support” - read that as the table row (9.0 = 128,172 characters, June 2016, Adlam and Tangut) and you know exactly which characters are safe to ship.
  3. Watch the two counterintuitive columns: Characters is cumulative repertoire, not new additions (so 6.3’s five bidirectional characters is not a typo), and Scripts counts Wikipedia’s tabulation of writing systems - versions can add thousands of characters without adding a single script, and vice versa.

Frequently asked questions

What is the latest Unicode version?

Version 18.0, published September 2026: 175 scripts and 172,808 characters - the largest single-version gain (+13,007) since 3.1, led by proto-cuneiform, Jurchen and Seal script. New versions ship every September (with rare slips - 14.0 was delayed from March to September 2021 by COVID-19).

When was emoji added to Unicode?

Emoji proper arrived in version 6.0 (October 2010), which Wikipedia’s table headlines with “emoticons” alongside playing cards, traffic signs and alchemical symbols. Emoji then became the release calendar’s engine: nearly every version since has carried a headline emoji batch, and the skin-tone modifiers of 8.0 (2015) turned them into a combinatorial system.

Which version added [my script]?

Each script appears exactly once. Some landmarks: Cherokee in 3.0 (1999), Linear B in 4.0 (2003), Tifinagh in 4.1 (2005), cuneiform in 5.0 (2006), Avestan and Samaritan in 5.2 (2009), Brahmi in 6.0 (2010), Manichaean and Psalter Pahlavi in 7.0 (2014), Old Hungarian in 8.0 (2015), Adlam in 9.0 (2016), Sogdian in 11.0 (2018), Old Uyghur in 14.0 (2021) and proto-cuneiform in 18.0 (2026).

Why did version numbers jump from 13.0 to 14.0 a year apart?

They did not - but the dates moved. 13.0 shipped March 2020 on the old spring schedule; in April 2020 the Consortium announced 14.0 would slip six months to September 2021 because of the COVID-19 pandemic (Wikipedia cites this as the schedule’s only major break), and from then on September became the permanent release month.

How many characters does Unicode have now?

After version 18.0 (September 2026): 172,808 characters across 175 scripts, out of a codespace of 1,114,112 possible code points. The gaps are deliberate - roughly 880,000 code points remain unassigned, reserve space that has already shaped the standard’s surrogate and plane architecture since version 2.0 (1996).

Related tools