Neural Text Normalization with Subword Units

Courtney Mansfield; Ming Sun; Yuzong Liu; Ankur Gandhe; Björn Hoffmeister

Publication

Neural Text Normalization with Subword Units

By Courtney Mansfield, Ming Sun, Yuzong Liu, Ankur Gandhe, Björn Hoffmeister

2019

Download Copy BibTeX

Share

Download

Copy BibTeX

Share

Text normalization (TN) is an important step in conversational systems. It converts written text to its spoken form to facilitate speech recognition, natural language understanding and text-to-speech synthesis. Finite state transducers (FSTs) are commonly used to build grammars that handle text normalization (Sproat, 1996; Roark et al., 2012). However, translating linguistic knowledge into grammars requires extensive effort. In this paper, we frame TN as a machine translation task and tackle it with sequence-to-sequence (seq2seq) models. Previous research focuses on normalizing a word (or phrase) with the help of limited word-level context, while our approach directly normalizes full sentences. We find subword models with additional linguistic features yield the best performance (with a word error rate of 0:17%).

Neural Text Normalization with Subword Units

Latest news

Work with us