Online Monaural Speech Enhancement Based on Periodicity Analysis and A Priori SNR Estimation

被引：13

作者：

Chen, Zhangli ^{[1
,2
]}

Hohmann, Volker ^{[1
,2
]}

机构：

[1] Carl von Ossietzky Univ Oldenburg, Med Phys, D-26129 Oldenburg, Germany

[2] Carl von Ossietzky Univ Oldenburg, Cluster Excellence Hearing4all, D-26129 Oldenburg, Germany

来源：

IEEE-ACM TRANSACTIONS ON AUDIO SPEECH AND LANGUAGE PROCESSING | 2015年 / 23卷 / 11期

关键词：

A priori signal-to-noise ratio (SNR) estimation; monaural speech enhancement; online implementation; periodicity analysis; NOISE; SEGREGATION; TRACKING;

D O I：

10.1109/TASLP.2015.2456423

中图分类号：

O42 [声学];

学科分类号：

070206 ; 082403 ;

摘要：

This paper describes an online algorithm for enhancing monaural noisy speech. First, a novel phase-corrected low-delay gammatone filterbank is derived for signal subband decomposition and resynthesis; the subband signals are then analyzed frame by frame. Second, a novel feature named periodicity degree (PD) is proposed to be used for detecting and estimating the fundamental period (P-0) in each frame and for estimating the signal-to-noise ratio (SNR) in each frame-subband signal unit. The PD is calculated in each unit as the multiplication of the normalized autocorrelation and the comb filter ratio, and shown to be robust in various low-SNR conditions. Third, the noise energy level in each signal unit is estimated recursively based on the estimated SNR for units with high PD and based on the noisy signal energy level for units with low PD. Then the a priori SNR is estimated using a decision-directed approach with the estimated noise level. Finally, a revised Wiener gain is calculated, smoothed, and applied to each unit; the processed units are summed across subbands and frames to form the enhanced signal. The detection accuracy of the algorithm was evaluated on two corpora and showed comparable performance on one corpus and better performance on the other corpus when compared to a recently published pitch detection algorithm. The speech enhancement effect of the algorithm was evaluated on one corpus with two objective criteria and showed better performance in one highly non-stationary noise and comparable performance in two other noises when compared to a state-of-the-art statistical-model based algorithm.

引用

页码：1904 / 1916

页数：13

共 50 条

[1] Monaural speech enhancement based on periodicity analysis
Chen, Z.
Hohmann, V
BIOMEDICAL ENGINEERING-BIOMEDIZINISCHE TECHNIK, 2014, 59 : S736 - S736
[2] IMPROVED A PRIORI SNR ESTIMATION IN SPEECH ENHANCEMENT
Nahma, Lara
Yong, Pei Chee
Dam, Hai Huyen
Nordholm, Sven
2017 23RD ASIA-PACIFIC CONFERENCE ON COMMUNICATIONS (APCC): BRIDGING THE METROPOLITAN AND THE REMOTE, 2017, : 253 - 257
[3] A priori SNR estimation and noise estimation for speech enhancement
Yao, Rui
Zeng, ZeQing
Zhu, Ping
EURASIP JOURNAL ON ADVANCES IN SIGNAL PROCESSING, 2016,
[4] GMM-based a priori SNR estimation in speech enhancement
Lei, Jianjun
Wang, Jian
Liu, Gang
Guo, Jun
WCICA 2006: SIXTH WORLD CONGRESS ON INTELLIGENT CONTROL AND AUTOMATION, VOLS 1-12, CONFERENCE PROCEEDINGS, 2006, : 4293 - +
[5] A priori SNR estimation and noise estimation for speech enhancement
Rui Yao
ZeQing Zeng
Ping Zhu
EURASIP Journal on Advances in Signal Processing, 2016
[6] A PRIORI SNR COMPUTATION FOR SPEECH ENHANCEMENT BASED ON CEPSTRAL ENVELOPE ESTIMATION
Elshamy, Samy
Madhu, Nilesh
Tirry, Wouter
Fingscheidt, Tim
2018 16TH INTERNATIONAL WORKSHOP ON ACOUSTIC SIGNAL ENHANCEMENT (IWAENC), 2018, : 351 - 355
[7] Improved Wavelet Based A-priori SNR Estimation for Speech Enhancement
Lun, Daniel Pak-Kong
Hsung, Tai-Chiu
2010 IEEE INTERNATIONAL SYMPOSIUM ON CIRCUITS AND SYSTEMS, 2010, : 2382 - 2385
[8] Speech enhancement employing modified a priori SNR estimation
Ou, Shifeng
Zhao, Xiaohui
Gao, Ying
SNPD 2007: EIGHTH ACIS INTERNATIONAL CONFERENCE ON SOFTWARE ENGINEERING, ARTIFICIAL INTELLIGENCE, NETWORKING, AND PARALLEL/DISTRIBUTED COMPUTING, VOL 3, PROCEEDINGS, 2007, : 827 - +
[9] Speech Enhancement Algorithm of Binary Mask Estimation Based on a Priori SNR Constraints
Wang, Jie
Yang, Chengcheng
Yan, Linhuang
Huang, Manlu
Sang, Jinqiu
2018 ASIA-PACIFIC SIGNAL AND INFORMATION PROCESSING ASSOCIATION ANNUAL SUMMIT AND CONFERENCE (APSIPA ASC), 2018, : 937 - 943
[10] A Priori SNR Estimation Based on a Recurrent Neural Network for Robust Speech Enhancement
Xia, Yangyang
Stern, Richard M.
19TH ANNUAL CONFERENCE OF THE INTERNATIONAL SPEECH COMMUNICATION ASSOCIATION (INTERSPEECH 2018), VOLS 1-6: SPEECH RESEARCH FOR EMERGING MARKETS IN MULTILINGUAL SOCIETIES, 2018, : 3274 - 3278

← 1 2 3 4 5 →