|
[1]王小川,”語音訊號處理”,全華科技圖書,2004. [2] Y. Gong, “Speech Recognition in Noisy Environments: A Survey”, Speech Communication 16, 1995. [3] M.J.F. Gales, “Model-based Techniques for Noise Robust Speech Recognition ”, University of Cambridge, Sep. 1995. [4] Boll, S. F, “Suppression of Acoustic Noise in Speech Using Spectral Subtraction”,IEEE Trans. on ASSP, Vol. 27, No. 2, pp.113-120.1979. [5] P. Lockwood and J. Boudy, “Experiments with a Nonlinear Spectral Subtractor (NSS) , Hidden Markov Models and the Projection, for Robust Speech Recognition in Cars”, Eurospeech 1991. [6] ITU-T Recommendation G.729 – Annex B: A silence compression sceme for G. 729 optimized for terminals conforming to Recommendation V.70. [7] B.A. Mellor and A.P. Varga, “Noise Masking in the MFCC Domain for the Recognition of Speech in Background Noise”, ICASSP 1992. [8] Y. Ephraim and H.L. Van Trees, “A Signal Subspace Approach for Speech Enhancement”, IEEE Trans. on Speech and Audio Processing, 1995. [9] S. Furui, "Cepstral Analysis Technique for Automatic Speaker Verification", IEEE Trans. Acoust. Speech Signal Process. 1981. [10] O. Viikki and K. Laurila, “Noise Robust HMM-based Speech Recognition Using Segmental Cepstral Feature Vector Normalization”, in ESCA NATO Workshop Robust Speech Recognition Unknown Communication Channels, Pont-a-Mousson, France, 1997, pp. 107–110. [11]H. Hermansky and N. Morgan, “RASTA Processing of Speech”. IEEE Trans. on Speech and Audio Processing. 2, pp. 578-589, 1994 . [12]Kuo-Hwei Yuo and Hsiao-Chuan Wang, “Robust Features for Noisy Speech Recognition Based on Temporal Trajectory Filtering of Short-Time Autocorrelation Sequences”, Speech Communication 28, 1999. [13]J.W. Hung, J.L. Shen, L.S. Lee, “New Approaches for Domain Transformation and Parameter Combination for Improved Accuracy in Parallel Model Combination ( PMC) Techniques”, IEEE Trans. on Speech and Audio Processing, Nov. 2001. [14]J.L. Gauiain and C.H.Lee, “Maximum a Posteriori Estimation for Multivariate Gaussian Mixture Observations of Markov Chains”, IEEE Trans. on Speech and Audio Processing, 1994. [15]C.J. Leggetter and P.C. Woodland, “Maximum Likelihood Linear Regression for Speaker Adaptation of Continuous Density Hidden Markov Models”, Computer Speech and Language, 1995. [16]呂麗如, “Improved Techniques for Continuous Mandarin Speech Recognition Under Telephone Environment”,國立台灣大學碩士論文,June 1999. [17]ITU-T Recommendation G.729 (Annex B): A Silence Compression Scheme for G.729, Optimized for Terminals Conforming to Recommendation V.70, ITU,1996. [18]Hemant Misra_, Shajith Ikbal_, Herv´e Bourlard_, Hynek Hermansky, “Spectral Entropy Based feature for Robust ASR”,ICASSP 2004. [19]郭正雄, “Robust Speech Recognition: Improved Spectral Subtraction”,國立暨南國際大學碩士論文,June 2004. [20]Harold Gene Longbotham,Alan Conrad Bovik, “Theory of Order Statistic Filter and Their Relationship to Linear FIR Filters”,IEEE TRANSACTIONS ON ACOUSTICS. SPEECH. AND SIGNAL PROCESSING. VOL. 37. NO. 2. FEBRUARY 1989. [21]Jos´e C. Segura, Javier Ram´ırez, Carmen Ben´ıtez, Angel de la Torre, Antonio Rubio, “Feature Extraction Combining Spectral Noise Reduction and Cepstral Histogram Equalization for Robust ASR”,ICSLP 2002. [22]Jos´e C. Segura, Javier Ram´ırez, Carmen Ben´ıtez, Angel de la Torre, Antonio Rubio, “A New Voice Activity Detector Using Subband Order-Statistics Filters for Robust Speech Recognition”,ICASSP 2004. [23]Jos´e C. Segura, Javier Ram´ırez, Carmen Ben´ıtez, Angel de la Torre, Antonio Rubio, “Improved Feature Extraction Based on Spectral Noise Reduction and Nonlinear Feature Normalization”,EUROSPEECH 2003. [24]Jos´e C. Segura, Javier Ram´ırez, Carmen Ben´ıtez, Angel de la Torre, Antonio Rubio, “Voice Activity Detection With Noise Reduction and Long-Term Spectral Divergence Estimation”,ICASSP 2004. [25]Jos´e C. Segura, Javier Ram´ırez, Carmen Ben´ıtez, Angel de la Torre, Antonio Rubio, “A New Adaptive Long-Term Spectral Estimation Voice Activity Detector”, EUROSPEECH 2003-GENEVA. [26]Jos´e C. Segura, Javier Ram´ırez, Carmen Ben´ıtez, Angel de la Torre, Antonio Rubio, “Improved Voice Activity Detection Combining Noise Reduction and Subband Divergence Measures”, INTERSPEECH 2004 – ICSLP. [27]Beena Ahmed and W. Harvey Holmes,“A Voice Activity Detector Using The Chi-Square Test”,ICASLP 2004. [28] R. Stern, A. Acero, F.-H. Liu, and Y. Ohshima, "Signal processing for robust speech recognition," Automatic Speech and Speaker Recognition. Advanced Topics. Kluwer Academic Pub., pp. 357--384, 1997. [29] Ananthakrishnan, K. S. "A comparison of modified k-means(MKM) and NN based real time adaptive clustering algorithms for articulatory space codebook formation", In ICSLP-1996, 1253-1256.
|