TY - GEN
T1 - Two-Stage Multi-Target Joint Learning for Monaural Speech Separation
AU - Nie, Shuai
AU - Liang, Shan
AU - Xue, Wei
AU - Zhang, Xueliang
AU - Liu, Wenju
AU - Dong, Like
AU - Yang, Hong
N1 - Publisher Copyright:
Copyright © 2015 ISCA.
PY - 2015
Y1 - 2015
N2 - Recently, supervised speech separation has been extensively studied and shown considerable promise. Due to the temporal continuity of speech, speech auditory features and separation targets present prominent spectro-temporal structures and strong correlations over the time-frequency (T-F) domain, which can be exploited for speech separation. However, many supervised speech separation methods independently model each T-F unit with only one target and much ignore these useful information. In this paper, we propose a two-stage multi-target joint learning method to jointly model the related speech separation targets at the frame level. Systematic experiments show that the proposed approach consistently achieves better separation and generalization performances in the low signal-to-noise ratio(SNR) conditions.
AB - Recently, supervised speech separation has been extensively studied and shown considerable promise. Due to the temporal continuity of speech, speech auditory features and separation targets present prominent spectro-temporal structures and strong correlations over the time-frequency (T-F) domain, which can be exploited for speech separation. However, many supervised speech separation methods independently model each T-F unit with only one target and much ignore these useful information. In this paper, we propose a two-stage multi-target joint learning method to jointly model the related speech separation targets at the frame level. Systematic experiments show that the proposed approach consistently achieves better separation and generalization performances in the low signal-to-noise ratio(SNR) conditions.
KW - Computational auditory scene analysis (CASA)
KW - Multi-target learning
KW - Speech separation
UR - http://www.scopus.com/inward/record.url?scp=84959170129&partnerID=8YFLogxK
M3 - Conference Proceeding
AN - SCOPUS:84959170129
VL - 2015-January
T3 - Proceedings of the Annual Conference of the International Speech Communication Association, INTERSPEECH
SP - 1503
EP - 1507
BT - 16th Annual Conference of the International Speech Communication Association, INTERSPEECH 2015
T2 - 16th Annual Conference of the International Speech Communication Association, INTERSPEECH 2015
Y2 - 6 September 2015 through 10 September 2015
ER -