- Makes pronunciation explicit. Especially useful for names, abbreviations, and languages where spelling does not map cleanly to sound.
- Byte-based models take another route: consume UTF-8 bytes and learn the mapping without a separate tokenizer or pronunciation dictionary.
Connections
- part_of Text-to-Speech