Front-end for far-field speech recognition based on frequency domain linear prediction
暂无分享,去创建一个
Hynek Hermansky | Samuel Thomas | Sriram Ganapathy | H. Hermansky | Samuel Thomas | Sriram Ganapathy
[1] Aaron E. Rosenberg,et al. Cepstral channel normalization techniques for HMM-based speaker verification , 1994, ICSLP.
[2] John Mourjopoulos,et al. Modelling and enhancement of reverberant speech using an envelope convolution method , 1983, ICASSP.
[3] S. Marple. Computing the discrete-time 'analytic' signal via FFT , 1997 .
[4] J. Makhoul,et al. Linear prediction: A tutorial review , 1975, Proceedings of the IEEE.
[5] Nelson Morgan,et al. Double the trouble: handling noise and reverberation in far-field automatic speech recognition , 2002, INTERSPEECH.
[6] Hynek Hermansky,et al. Temporal processing of speech in a time-feature space , 1997 .
[7] James David Johnston,et al. Enhancing the Performance of Perceptual Audio Coders by Using Temporal Noise Shaping (TNS) , 1996 .
[8] H Hermansky,et al. Perceptual linear predictive (PLP) analysis of speech. , 1990, The Journal of the Acoustical Society of America.
[9] J. Flanagan,et al. Computer‐steered microphone arrays for sound transduction in large rooms , 1985 .
[10] Fumitada Itakura,et al. An approach to dereverberation using multi-microphone sub-band envelope estimation , 1991, [Proceedings] ICASSP 91: 1991 International Conference on Acoustics, Speech, and Signal Processing.
[11] David Pearce,et al. The aurora experimental framework for the performance evaluation of speech recognition systems under noisy conditions , 2000, INTERSPEECH.
[12] Hynek Hermansky,et al. On the effects of short-term spectrum smoothing in channel normalization , 1997, IEEE Trans. Speech Audio Process..
[13] Daniel P. W. Ellis,et al. LP-TRAP: linear predictive temporal patterns , 2004, INTERSPEECH.