2021
DOI: 10.48550/arxiv.2107.01922
|View full text |Cite
Preprint
|
Sign up to set email alerts
|

Investigation of Practical Aspects of Single Channel Speech Separation for ASR

Jian Wu,
Zhuo Chen,
Sanyuan Chen
et al.

Abstract: Speech separation has been successfully applied as a frontend processing module of conversation transcription systems thanks to its ability to handle overlapped speech and its flexibility to combine with downstream tasks such as automatic speech recognition (ASR). However, a speech separation model often introduces target speech distortion, resulting in a sub-optimum word error rate (WER). In this paper, we describe our efforts to improve the performance of a single channel speech separation system. Specifical… Show more

Help me understand this report

Search citation statements

Order By: Relevance

Paper Sections

Select...
1

Citation Types

0
1
0

Year Published

2022
2022
2022
2022

Publication Types

Select...
1

Relationship

1
0

Authors

Journals

citations
Cited by 1 publication
(1 citation statement)
references
References 36 publications
0
1
0
Order By: Relevance
“…This is typically realized with window-based processing, where a speech separation neural network model is used to process each windowed signal. CSS can be performed with either single microphone [10,11] or a microphone array [5,12] with the latter being more effective in many cases.…”
Section: Introductionmentioning
confidence: 99%
“…This is typically realized with window-based processing, where a speech separation neural network model is used to process each windowed signal. CSS can be performed with either single microphone [10,11] or a microphone array [5,12] with the latter being more effective in many cases.…”
Section: Introductionmentioning
confidence: 99%