Phonetic symbols in Unicode


Unicode supports several phonetic scripts and notation systems through its existing scripts and the addition of extra blocks with phonetic characters. These phonetic characters are derived from an existing script, usually Latin, Greek or Cyrillic. Apart from the International Phonetic Alphabet, extensions to the IPA and IPA symbols, these blocks also contain characters from the Uralic Phonetic Alphabet and the Americanist Phonetic Alphabet.

Phonetic scripts

The International Phonetic Alphabet (IPA) makes use of letters from other writing systems as most phonetic scripts do. IPA notably uses Latin, Greek and Cyrillic characters. Combining diacritics also add meaning to the phonetic text. Finally, these phonetic alphabets make use of modifier letters, that are specially constructed for phonetic meaning. A "modifier letter" is strictly intended not as an independent grapheme but as a modification of the preceding character resulting in a distinct grapheme, notably in the context of the International Phonetic Alphabet. For example, ʰ should not occur on its own but modifies the preceding or following symbol. Thus, is a single IPA symbol, distinct from. In practice, however, several of these "modifier letters" are also used as full graphemes, e.g. ʿ as transliterating Semitic ayin or Hawaiian okina, or ˚ transliterating Abkhaz ә.

From to Unicode

Consonants

The following tables indicates the Unicode code point sequences for phonemes as used in the International Phonetic Alphabet. A bold code point indicates that the Unicode chart provides an application note such as "voiced retroflex lateral" for. An entry in bold italics indicates the character name itself refers to a phoneme such as

Vowels

The following figures depict the phonetic vowels and their Unicode / UCS code points, arranged to represent the phonetic vowel trapezium. Vowels appearing in pairs in the figure to the right indicate rounded and unrounded variations respectively. Again, characters with Unicode names referring to phonemes are indicated by bold text. Those with explicit application notes are indicated by bold italic text. Those from borrowed unchanged from another script are indicated by italics. Before and after a bullet are the unrounded • rounded vowels.
FrontCentralBack
Close
U+0069

U+0079

U+0268

U+0289

U+026F

U+0075
Near-close
U+026A

U+028F

U+026A U+0308

U+028A U+0308



U+028A
Close-mid
U+0065

U+00F8

U+0258

U+0275

U+0264

U+006F
Mid
U+0065 U+031E

U+00F8 U+031E


U+0259

U+0264 U+031E

U+006F U+031E
Open-mid
U+025B

U+0153

U+025C

U+025E

U+028C

U+0254
Near-open
U+00E6



U+0250
Open
U+0061

U+0276

U+0061 U+0308


U+0251

U+0252

Diacritics

Diacritics may be encoded as either modifier or combining characters.

Unicode blocks

Unicode blocks with many phonetic symbols

Six Unicode blocks contain many phonetic symbols:

Spacing Modifier Letters (U+02B0–02FF)

The characters in the "Spacing Modifier Letters" block are intended as forming a unity with the preceding letter. E.g. the character isn't intended simply as a superscript h, but as the mark of aspiration placed after the letter being aspirated, as in "aspirated voiceless bilabial plosive". The block contains:
  • Latin superscript modifier letters: : ʰ aspiration; ʱ breathy voice, murmured; ʲ palatalization; ʳ, ʴ, ʵ, ʶ r-coloring or r-offglides; ʷ labialization; ʸ palatalization, Americanist usage for U+02B2
  • Miscellaneous phonetic modifiers: : ʹ ʺ ʻ ʼ ʽ ʾ ʿ ˀ ˁ ˂ ˃ ˄ ˅ ˆ ˇ ˈ ˉ ˊ ˋ ˌ ˍ ˎ ˏ ː ˑ ˒ ˓ ˔ ˕ ˖ ˗
  • Spacing clones of diacritics: : ˘ breve; ˙ dot above; ˚ ring above; ˛ ogonek; ˜ small tilde; ˝ double acute accent
  • Additions based on 1989 IPA: : ˞ ˟ ˠ ˡ ˢ ˣ ˤ
  • Tone letters: : ˥ ˦ ˧ ˨ ˩
  • Extended Bopomofo tone marks: ;
  • IPA modifiers:, unaspirated
  • Other modifier letters: for Nenets
  • Uralic Phonetic Alphabet modifiers: : ˯ ˰ ˱ ˲ ˳ ˴ ˵ ˶ ˷ ˸ ˹ ˺ ˻ ˼ ˽ ˾ ˿

Phonetic Extensions (U+1D00–1D7F)

This block, together with Phonetic Extensions Supplement below, contains:
  • Small capitals "ɢ ɪ ɴ ɶ ʀ ʏ ʙ ʜ ʟ"
  • Turned small letters "ɐ ɥ ɯ ɹ ɺ ɻ ʇ ʌ ʍ ʎ ʞ ʮ ʯ"
  • Extra small capitals "ʁ ʛ ᴀ ᴁ ᴃ ᴄ ᴅ ᴆ ᴇ ᴊ ᴋ ᴌ ᴍ ᴎ ᴏ ᴐ ᴘ ᴙ ᴚ ᴛ ᴜ ᴠ ᴡ ᴢ ᴣ ᴦ ᴧ ᴨ ᴩ ᴪ"
  • Letters with palatal hooks "ƫ ᶀ ᶁ ᶂ ᶃ ᶄ ᶅ ᶆ ᶇ ᶈ ᶉ ᶊ ᶋ ᶌ ᶍ ᶎ ᶪ ᶵ"
  • Letters with retroflex hooks "ᶏ ᶐ ᶒ ᶓ ᶔ ᶕ ᶖ ᶗ ᶘ ᶙ ᶚ ᶩ ᶯ ᶼ"

Input by selection from a screen

Many systems provide a way to select Unicode characters visually. ISO/IEC 14755 refers to this as a screen-selection entry method.
Microsoft Windows has provided a Unicode version of the Character Map program since version NT 4.0 – appearing in the consumer edition since XP. This is limited to characters in the Basic Multilingual Plane. Characters are searchable by Unicode character name, and the table can be limited to a particular code block. More advanced third-party tools of the same type are also available.
macOS provides a "character palette" with much the same functionality, along with searching by related characters, glyph tables in a font, etc. It can be enabled in the input menu in the menu bar under System Preferences → International → Input Menu or can be viewed under Edit → Emoji & Symbols in many programs.
Equivalent tools – such as gucharmap or kcharselect – exist on most Linux desktop environments.