Dinh, Tu Anh scite author profile

Dinh, Tu Anh

3Publications

0Citation Statements Received

26Citation Statements Given

How they've been cited

How they cite others

Affiliations

Maastricht University

Publications

Order By: Most citations

Zero-shot Speech Translation

Anh¹

2021

Preprint

View full text Add to dashboard Cite

Speech Translation (ST) is the task of translating speech in one language into text in another language. Traditional cascaded approaches for ST, using Automatic Speech Recognition (ASR) and Machine Translation (MT) systems, are prone to error propagation. End-to-end approaches use only one system to avoid propagating error, yet are difficult to employ due to data scarcity. We explore zero-shot translation, which enables translating a pair of languages that is unseen during training, thus avoid the use of end-to-end ST data. Zero-shot translation has been shown to work for multilingual machine translation, yet has not been studied for speech translation. We attempt to build zero-shot ST models that are trained only on ASR and MT tasks but can do ST task during inference. The challenge is that the representation of text and audio is significantly different, thus the models learn ASR and MT tasks in different ways, making it non-trivial to perform zero-shot. These models tend to output the wrong language when performing zero-shot ST. We tackle the issues by including additional training data and an auxiliary loss function that minimizes the text-audio difference. Our experiment results and analysis show that the methods are promising for zero-shot ST. Moreover, our methods are particularly useful in the fewshot settings where a limited amount of ST data is available, with improvements of up to +11.8 BLEU points compared to direct end-to-end ST models and +3.9 BLEU points compared to ST models fine-tuned from pre-trained ASR model.

show abstract

Tackling Data Scarcity in Speech Translation Using Zero-Shot Multilingual Machine Translation Techniques

Anh

Liu

Niehues

2022

View full text Add to dashboard Cite

KIT’s Multilingual Speech Translation System for IWSLT 2023

Liu¹,

Nguyen²,

Koneru³

et al. 2023

View full text Add to dashboard Cite

scite is a Brooklyn-based organization that helps researchers better discover and understand research articles through Smart Citations–citations that display the context of the citation and describe whether the article provides supporting or contrasting evidence. scite is used by students and researchers from around the world and is funded in part by the National Science Foundation and the National Institute on Drug Abuse of the National Institutes of Health.

Contact Info

hi@scite.ai

10624 S. Eastern Ave., Ste. A-614

Henderson, NV 89052, USA

Blog Terms and Conditions API Terms Privacy Policy Contact Cookie Preferences Do Not Sell or Share My Personal Information

Made with 💙 for researchers

Part of the Research Solutions Family.

Dinh, Tu Anh

Zero-shot Speech Translation

Tackling Data Scarcity in Speech Translation Using Zero-Shot Multilingual Machine Translation Techniques

KIT’s Multilingual Speech Translation System for IWSLT 2023

Contact Info

Product

Resources

About