Wenning Wei scite author profile

Wenning Wei

5Publications

99Citation Statements Received

49Citation Statements Given

How they've been cited

143

How they cite others

123

Affiliations

Microsoft (United States), Microsoft (Finland), Microsoft Research Asia (China)

Publications

Order By: Most citations

Developing RNN-T Models Surpassing High-Performance Hybrid Models with Customization Capability

Zhao

Meng

et al. 2020

View full text Add to dashboard Cite

Because of its streaming nature, recurrent neural network transducer (RNN-T) is a very promising end-to-end (E2E) model that may replace the popular hybrid model for automatic speech recognition. In this paper, we describe our recent development of RNN-T models with reduced GPU memory consumption during training, better initialization strategy, and advanced encoder modeling with future lookahead. When trained with Microsoft's 65 thousand hours of anonymized training data, the developed RNN-T model surpasses a very well trained hybrid model with both better recognition accuracy and lower latency. We further study how to customize RNN-T models to a new domain, which is important for deploying E2E models to practical scenarios. By comparing several methods leveraging text-only data in the new domain, we found that updating RNN-T's prediction and joint networks using text-to-speech generated from domain-specific text is the most effective.

show abstract

Using Personalized Speech Synthesis and Neural Language Generator for Rapid Speaker Adaptation

Huang

Wei

Gale

et al. 2020

View full text Add to dashboard Cite

Adaptation of RNN Transducer with Text-To-Speech Technology for Keyword Spotting

Sharma

Wei

et al. 2020

View full text Add to dashboard Cite

Developing RNN-T Models Surpassing High-Performance Hybrid Models with Customization Capability

Li¹,

Zhao²,

Meng³

et al. 2020

Preprint

View full text Add to dashboard Cite

Rapid RNN-T Adaptation Using Personalized Speech Synthesis and Neural Language Generator

Huang

Wei

et al. 2020

View full text Add to dashboard Cite

scite is a Brooklyn-based organization that helps researchers better discover and understand research articles through Smart Citations–citations that display the context of the citation and describe whether the article provides supporting or contrasting evidence. scite is used by students and researchers from around the world and is funded in part by the National Science Foundation and the National Institute on Drug Abuse of the National Institutes of Health.

Contact Info

hi@scite.ai

10624 S. Eastern Ave., Ste. A-614

Henderson, NV 89052, USA

Blog Terms and Conditions API Terms Privacy Policy Contact Cookie Preferences Do Not Sell or Share My Personal Information

Made with 💙 for researchers

Part of the Research Solutions Family.

Wenning Wei

Developing RNN-T Models Surpassing High-Performance Hybrid Models with Customization Capability

Using Personalized Speech Synthesis and Neural Language Generator for Rapid Speaker Adaptation

Adaptation of RNN Transducer with Text-To-Speech Technology for Keyword Spotting

Developing RNN-T Models Surpassing High-Performance Hybrid Models with Customization Capability

Rapid RNN-T Adaptation Using Personalized Speech Synthesis and Neural Language Generator

Contact Info

Product

Resources

About