Please Note: This article is written for users of the following Microsoft Word versions: 2007, 2010, 2013, and 2016. If you are using an earlier version (Word 2003 or earlier), this tip may not work for you. For a version of this tip written specifically for earlier versions of Word, click here: Understanding Unicode Characters.

Understanding Unicode Characters

Written by Allen Wyatt (last updated April 26, 2021)
This tip applies to Word 2007, 2010, 2013, and 2016


2

You may have heard of the term Unicode before, and wondered what it meant. Normal single-byte encoding schemes (such as ASCII and ANSI) allow only up to 256 unique individual characters to be encoded and displayed on the computer. In the global computer community, where each member is required to work in their own language, this is a problem. There are far more than 256 characters in common use throughout the world.

This is where Unicode comes into play.

Depending on the version of Unicode being used, the standard requires anywhere from two to five bytes for encoding each character. As of this writing, the current Unicode standard is 9.0.0, which uses five bytes and 128,172 characters defined. This standard, devised and promoted by the Unicode Consortium (http://www.unicode.org), allows for the display of virtually all the unique language characters in the world. A team of computer professionals, linguists, and scholars continue to work on the actual development of Unicode.

The use of multiple bytes to define each character means that Unicode can be used to encode most of the characters used in the world's major languages. There is an extension mechanism built into the standard, as well, which means it is possible to encode close to a million more characters, if necessary. This ability should be sufficient for all known language requirements, plus the encoding of all the historic scripts of the world. (This includes languages and symbols that are no longer in use.)

As presently defined, Unicode 9.0.0 (the latest version, released in June 2016) includes codes for characters used in the major written languages of the world, including Arabic, Armenian, Balinese, Bengali, Bopomofo, Buhid, Canadian Syllabics, Cherokee, Chinese, Cyrillic, Deseret, Devanagari, Ethiopic, Georgian, Gothic, Greek, Gujarati, Gurmukhi, Han, Hangul, Hanun—o, Hebrew, Hiragana, Kannada, Katakana, Khmer, Lao, Latin, Malayalam, Mongolian, Myanmar, Ogham, Old Italic (Etruscan), Oriya, Phoenician, Runic, Sinhala, Syriac, Tagalog, Tagbanwa, Tamil, Telugu, Thaana, Thai, Tibetan, and Yi. Work is progressing to add more characters from lesser-known languages.

In addition, Unicode also includes many different symbols, including numbers, general diacritics, general punctuation, general symbols, dingbats, emojis, arrows, blocks, box drawing forms, geometric shapes, mathematical symbols, musical symbols (western and byzantine), technical symbols, braille patterns, and Kangxi radicals.

Unicode is supported in all modern versions of Windows and Word. Exactly what standard of Unicode that is supported depends on the version of Windows and Word in question.

WordTips is your source for cost-effective Microsoft Word training. (Microsoft Word is the most popular word processing software in the world.) This tip (11277) applies to Microsoft Word 2007, 2010, 2013, and 2016. You can find a version of this tip for the older menu interface of Word here: Understanding Unicode Characters.

Author Bio

Allen Wyatt

With more than 50 non-fiction books and numerous magazine articles to his credit, Allen Wyatt is an internationally recognized author. He is president of Sharon Parq Associates, a computer and publishing services company. ...

MORE FROM ALLEN

Understanding and Creating Lists

There are two types of common lists you can use in a document: bulleted lists and numbered lists. This tip explains the ...

Discover More

Calculating an Age On a Given Date

Start putting dates in a worksheet (especially birthdates), and sooner or later you will need to calculate an age based ...

Discover More

Working with Table Columns and Rows

Need to add or delete columns and rows from a table? It's easy to do using the tools provided in Word.

Discover More

Do More in Less Time! Are you ready to harness the full power of Word 2013 to create professional documents? In this comprehensive guide you'll learn the skills and techniques for efficiently building the documents you need for your professional and your personal life. Check out Word 2013 In Depth today!

More WordTips (ribbon)

Multiple Taskbar Icons for Documents

If you like to see your open documents in multiple icons on the Taskbar, you may wonder how to make that happen. This is ...

Discover More

Changing the Document Page Color

Word's default black text and a white page background may not appeal to everyone. Here's how you can easily change the ...

Discover More

Changing from Pirated to Permitted Software

When you install Microsoft Office, you are required to enter a product key that unlocks the software for your use. This ...

Discover More
Subscribe

FREE SERVICE: Get tips like this every week in WordTips, a free productivity newsletter. Enter your address and click "Subscribe."

View most recent newsletter.

Comments

If you would like to add an image to your comment (not an avatar, but an image to help in making the point of your comment), include the characters [{fig}] (all 7 characters, in the sequence shown) in your comment text. You’ll be prompted to upload your image when you submit the comment. Maximum image size is 6Mpixels. Images larger than 600px wide or 1000px tall will be reduced. Up to three images may be included in a comment. All images are subject to review. Commenting privileges may be curtailed if inappropriate images are posted.

What is 6 + 5?

2017-04-18 13:31:43

Rod Grealish

James, Open web page https://unicode-table.com/en/. On the right, click "Open in separate page". This will display a list of Unicode blocks. Select the one in which you are interested. You can also enter a character in the search box near the top of the page to find a Unicode value.


2017-04-17 13:50:11

James

Allen, Thank you for the theory, but practically, a few questions:

1) where or how do we find the Unicode of a character?
2) or vice versa, how do we decode the Unicode to find out what character it represents
3) How can this be done with a macro?
Regards.


This Site

Got a version of Word that uses the ribbon interface (Word 2007 or later)? This site is for you! If you use an earlier version of Word, visit our WordTips site focusing on the menu interface.

Videos
Subscribe

FREE SERVICE: Get tips like this every week in WordTips, a free productivity newsletter. Enter your address and click "Subscribe."

(Your e-mail address is not shared with anyone, ever.)

View the most recent newsletter.