Jin Chu Wu,Alvin F Martin,Craig S Greenberg et al.
Jin Chu Wu et al.
The data dependency due to multiple use of the same subjects has impact on the standard error (SE) of the detection cost function (DCF) in speaker recognition evaluation. The DCF is defined as a weighted sum of the probabilities of type I a...
Xiao-Lei Zhang,DeLiang Wang
Xiao-Lei Zhang
Monaural speech separation is a fundamental problem in robust speech processing. Recently, deep neural network (DNN)-based speech separation methods, which predict either clean speech or an ideal time-frequency mask, have demonstrated remar...
James M Kates,Kathryn H Arehart
James M Kates
This paper presents an index designed to predict music quality for individuals listening through hearing aids. The index is "intrusive", that is, it compares the degraded signal being evaluated to a reference signal. The index is based on a...
Donald S Williamson,Yuxuan Wang,DeLiang Wang
Donald S Williamson
Speech separation systems usually operate on the short-time Fourier transform (STFT) of noisy speech, and enhance only the magnitude spectrum while leaving the phase spectrum unchanged. This is done because there was a belief that the phase...
Relationships between vocal function measures derived from an acoustic microphone and a subglottal neck-surface accelerometer [0.03%]
基于声学麦克风和亚声带颈部表面加速度计的发声功能指标之间的关系
Daryush D Mehta,Jarrad H Van Stan,Robert E Hillman
Daryush D Mehta
Monitoring subglottal neck-surface acceleration has received renewed attention due to the ability of low-profile accelerometers to confidentially and noninvasively track properties related to normal and disordered voice characteristics and ...
Improving Robustness of Deep Neural Network Acoustic Models via Speech Separation and Joint Adaptive Training [0.03%]
基于说话人分离与联合自适应改进深度神经网络声学模型的鲁棒性
Arun Narayanan,DeLiang Wang
Arun Narayanan
Although deep neural network (DNN) acoustic models are known to be inherently noise robust, especially with matched training and testing data, the use of speech separation as a frontend and for deriving alternative feature representations h...
Yishan Jiao,Visar Berisha,Ming Tu et al.
Yishan Jiao et al.
Speaking rate estimation directly from the speech waveform is a long-standing problem in speech signal processing. In this paper, we pose the speaking rate estimation problem as that of estimating a temporal density function whose integral ...
Yuxuan Wang,Arun Narayanan,DeLiang Wang
Yuxuan Wang
Formulation of speech separation as a supervised learning problem has shown considerable promise. In its simplest form, a supervised learning algorithm, typically a deep neural network, is trained to learn a mapping from noisy features to a...