RNN Transducer Models for Spoken Language Understanding

Thomas, Samuel; Kuo, Hong-Kwang Jeff; Saon, George; Tüske, Zoltán; Kingsbury, Brian; Kurata, Gakuto; Kons, Zvi; Hoory, Ron

doi:10.1109/icassp39728.2021.9414029

Cited by 9 publications

(1 citation statement)

References 29 publications

(36 reference statements)

Supporting

Mentioning

Contrasting

Order By: Relevance

“…The graph clearly shows that additional data helps to improve the accuracy of the classifier. Our second dataset was composed of recordings from a call center with manual transcription and labeling [12][13][14]. There are in total about 30,000 samples split into test (5,600), train (21,900), and development (2,500) sets.…”

Section: Dataset Amentioning

confidence: 99%

A new data augmentation method for intent classification enhancement and its application on spoken conversation datasets

Kons¹,

Satt²,

Kuo³

et al. 2022

Preprint

Self Cite

View full text Add to dashboard Cite

Intent classifiers are vital to the successful operation of virtual agent systems. This is especially so in voice activated systems where the data can be noisy with many ambiguous directions for user intents. Before operation begins, these classifiers are generally lacking in real-world training data. Active learning is a common approach used to help label large amounts of collected user input. However, this approach requires many hours of manual labeling work. We present the Nearest Neighbors Scores Improvement (NNSI) algorithm for automatic data selection and labeling. The NNSI reduces the need for manual labeling by automatically selecting highly-ambiguous samples and labeling them with high accuracy. This is done by integrating the classifier's output from a semantically similar group of text samples. The labeled samples can then be added to the training set to improve the accuracy of the classifier. We demonstrated the use of NNSI on two large-scale, real-life voice conversation systems. Evaluation of our results showed that our method was able to select and label useful samples with high accuracy. Adding these new samples to the training data significantly improved the classifiers and reduced error rates by up to 10%.

show abstract

Section: Dataset Amentioning

confidence: 99%