• Makes pronunciation explicit. Especially useful for names, abbreviations, and languages where spelling does not map cleanly to sound.
  • Byte-based models take another route: consume UTF-8 bytes and learn the mapping without a separate tokenizer or pronunciation dictionary.

Connections