Automatic detection of Voice Disorders: recent literature advancements
Abstract
1. Introduction
2. Databases
2.1. Massachusetts Eye and Ear Infirmary (MEEI) database
2.3. Arabic voice pathology database (AVPD)
2.4. VOice ICar fEDerico II (VOICED)
2.5. The Cantonese perceptual evaluation of voice (CanPEV) database
2.6. The HUPA Database
2.7. The Advanced Voice Function Assessment Database (AVFAD)
2.8. The Voice disordered and Healthy Adults Speech Database
2.9. The LANNA children speech corpus
2.10. Parkinson’s UI Machine Learning database
3. Some representative recent research papers on automatic voice disorders detection
3.1. Mobile apps for automatic detection of voice disorders
4. Automatic classification of voice disorders
5. Speech impairments by central nervous system disorders.
5.1. Parkinons’ desease (PD)
6. Specific language impairment (SLI)
7. Conclusions
References
- Abbas, Musatafa, Abbood Albadr, and Sabrina Tiun. 2017. Extreme Learning Machine: A Review. International Journal of Applied Engineering Research 12: 4610–4623. [Google Scholar]
- Airaksinen, Manu, Tuomo Raitio, Brad Story, and Paavo Alku. 2014. Quasi Closed Phase Glottal Inverse Filtering Analysis with Weighted Linear Prediction. IEEE Transactions on Audio, Speech and Language Processing 22 (3): 596–607. [Google Scholar] [CrossRef] [Scilit]
- AL-Dhief, Fahad Taha, Nurul Mu’azzah Abdul Latiff, Nik Noordini Nik Abd. Malik, Naseer Sabri, Marina Mat Baki, Musatafa Abbas Abbood Albadr, Aymen Fadhil Abbas, Yaqdhan Mahmood Hussein, and Mazin Abed Mohammed. 2020. Voice Pathology Detection Using Machine Learning Technique. In 2020 IEEE 5th International Symposium on Telecommunication Technologies (ISTT), 99–104. IEEE. [CrossRef] [Scilit]
- Alhussein, Musaed, and Ghulam Muhammad. 2018. Voice Pathology Detection Using Deep Learning on Mobile Healthcare Framework. IEEE Access 6: 41034–41. [Google Scholar] [CrossRef] [Scilit]
- Arias-Londoño, Julián David, Juan I. GodinoLlorente, Maria Markaki, and Yannis Stylianou. 2011. On Combining Information from Modulation Spectra and Mel-Frequency Cepstral Coefficients for Automatic Detection of Pathological Voices. Logopedics Phoniatrics Vocology 36 (2): 60–69. [Google Scholar] [CrossRef] [Scilit] [PubMed]
- Asmae, Ouhmida, Raihani Abdelhadi, Cherradi Bouchaib, Sandabad Sara, and Khalili Tajeddine. 2020. Parkinson’s Disease Identification Using KNN and ANN Algorithms Based on Voice Disorder. In 2020 1st International Conference on Innovative Research in Applied Science, Engineering and Technology (IRASET), 1–6. IEEE. [CrossRef] [Scilit]
- Barry, WJ, and M Putzer. 2020. Voice Database. Accessed: December 10, 2020. http://www.stimmdatenbank.coli.uni-saarland.de/.
- Behlau, M. 2003. Consensus Auditory-Perceptual Evaluation of Voice (CAPE-V). ASHA 9: 187–89. [Google Scholar]
- Breiman, L. 2001. Random Forest. Machine Learning 45 (1): 5–32. [Google Scholar] [PubMed]
- Cesari, Ugo, Giuseppe De Pietro, Elio Marciano, Ciro Niri, Giovanna Sannino, and Laura Verde. 2018. A New Database of Healthy and Pathological Voices. Computers and Electrical Engineering 68: 310–21. [Google Scholar] [CrossRef] [Scilit]
- Dixit, RP. 1988. On Defining Aspiration. In Proc. 12th Int. Conf. Linguistics, 606–10. Tokyo, Japan.
- Elemetrics, K. 1994. Kay Elemetrics Corp. Disordered Voice Data-Base. Model 4337. [Google Scholar]
- Eyben, Florian, Klaus R. Scherer, Bjorn W. Schuller, Johan Sundberg, Elisabeth Andre, Carlos Busso, Laurence Y. Devillers, and et al. 2016. The Geneva Minimalistic Acoustic Parameter Set (GeMAPS) for Voice Research and Affective Computing. IEEE Transactions on Affective Computing 7 (2): 190–202. [Google Scholar] [CrossRef] [Scilit]
- Georgopoulos, Voula C. 2020. Advanced TimeFrequency Analysis and Machine Learning for Pathological Voice Detection. In 2020 12th International Symposium on Communication Systems, Networks and Digital Signal Processing (CSNDSP), 1–5. IEEE. [CrossRef] [Scilit]
- Godino-Llorente, Juan Ignacio, Víctor Osma-Ruiz, Nicolás Sáenz-Lechón, Ignacio Cobeta-Marco, Ramón González-Herranz, and Carlos RamírezCalvo. 2008. Acoustic Analysis of Voice Using WPCVox: A Comparative Study with Multi Dimensional Voice Program. European Archives of Oto-Rhino-Laryngology 265 (4): 465–76. [Google Scholar] [CrossRef] [Scilit]
- Grill, Pavel, and Jana Tučková. 2016. Speech Databases of Typical Children and Children with SLI. Edited by Frederic Dick. PLOS ONE 11 (3): e0150365. [Google Scholar] [CrossRef] [Scilit]
- Hagan, Martin T., and Mohammad B. Menhaj. 1994. Training Feedforward Networks with the Marquardt Algorithm. IEEE Transactions on Neural Networks 5 (6): 989–93. [Google Scholar] [CrossRef] [Scilit] [PubMed]
- He, Kaiming, Xiangyu Zhang, Shaoqing Ren, and Jian Sun. 2016. Deep Residual Learning for Image Recognition. In 2016 IEEE Conference on Computer Vi-Sion and Pattern Recognition (CVPR), 770–78. pp. 770–78.
- Huang, Guang-Bin. 2015. What Are Extreme Learning Machines? Filling the Gap Between Frank Rosenblatt’s Dream and John von Neumann’s Puzzle. Cognitive Computation 7: 263–78. [Google Scholar] [CrossRef] [Scilit]
- Ilapakurti, Anitha, Sharat Kedari, Jaya Shankar Vuppalapati, Santosh Kedari, and Chandrasekar Vuppalapati. 2019. Artificial Intelligent (AI) Clinical Edge for Voice Disorder Detection. In 2019 IEEE Fifth International Conference on Big Data Computing Service and Applications (BigDataService), 340–45. IEEE. [CrossRef] [Scilit]
- Jesus, Luis M.T., Inês Belo, Jessica Machado, and Andreia Hall. 2017. The Advanced Voice Function Assessment Databases (AVFAD): Tools for Voice Clinicians and Speech Research. Advances in Speech-Language Pathology 9: 237255. [Google Scholar] [CrossRef] [Scilit]
- Kadiri, Sudarsana Reddy, and Paavo Alku. 2020. Analysis and Detection of Pathological Voice Using Glottal Source Features. IEEE Journal of Selected Topics in Signal Processing 14 (2): 367–79. [Google Scholar] [CrossRef] [Scilit]
- Kantz, Holger, and Thomas Schreiber. 2003. Nonlinear Time Series Analysis, Edited by Cambridge University Press. , 2nd ed. Cambridge, UK. [Google Scholar]
- Kodrasi, Ina, and Herve Bourlard. 2020. SpectroTemporal Sparsity Characterization for Dysarthric Speech Detection. IEEE/ACM Transactions on Audio, Speech, and Language Processing 28: 1210–22. [Google Scholar] [CrossRef] [Scilit]
- Kotarba, Katarzyna, and Michal Kotarba. 2020. Efficient Detection of Specific Language Impairment in Children Using ResNet Classifier. In 2020 Signal Processing: Algorithms, Architectures, Arrangements, and Applications (SPA), 169–73. IEEE. [CrossRef] [Scilit]
- Lauraitis, Andrius, Rytis Maskeliunas, Robertas Damasevicius, and Tomas Krilavicius. 2020. Detection of Speech Impairments Using Cepstrum, Auditory Spectrogram and Wavelet Time Scattering Domain Features. IEEE Access 8: 96162–72. [Google Scholar] [CrossRef] [Scilit]
- Law, T, K Lee, JH Lam, AC van Hasselt, and MCF Tong. 2010. The Construction of the Cantonese Percep-Tual Evaluation of Voice (CanPEV): The Content Validation Process. In Proc. 4th World Voice Congr. World Voice Consortium, 159. Seoul, Korea.
- Little, Max A., Patrick E. McSharry, Eric J. Hunter, Jennifer Spielman, and Lorraine O. Ramig. 2009. Suitability of Dysphonia Measurements for Telemonitoring of Parkinson’s Disease. IEEE Transactions on Biomedical Engineering 56 (4): 1015–22. [Google Scholar] [CrossRef] [Scilit]
- Liu, Yuanyuan, Tan Lee, Thomas Law, and Kathy Yuet-Sheung Lee. 2019. Acoustical Assessment of Voice Disorder With Continuous Speech Using ASR Posterior Features. IEEE/ACM Transactions on Audio, Speech, and Language Processing 27 (6): 1047–59. [Google Scholar] [CrossRef] [Scilit]
- Mesallam, Tamer A., Mohamed Farahat, Khalid H. Malki, Mansour Alsulaiman, Zulfiqar Ali, Ahmed Alnasheri, and Ghulam Muhammad. 2017. Development of the Arabic Voice Pathology Database and Its Evaluation by Using Speech Features and Machine Learning Algorithms. Journal of Healthcare Engineering 2017: 1–13. [Google Scholar] [CrossRef] [Scilit]
- Miliaresi, Ioanna, Kyriakos Poutos, and Aggelos Pikrakis. 2021. Combining Acoustic Features and Medical Data in Deep Learning Networks for Voice Pathology Classification. In 2020 28th European Signal Processing Conference (EUSIPCO), 1190–94. IEEE. [CrossRef] [Scilit]
- Moro-Velázquez, Laureano, Jorge Andrés GómezGarcía, Juan Ignacio Godino-Llorente, and Gustavo Andrade-Miranda. 2015. Modulation Spectra Morphological Parameters: A New Method to Assess Voice Pathologies According to the GRBAS Scale. BioMed Research International. [Google Scholar] [CrossRef] [Scilit]
- Muhammad, G, MF Alhamid, M Al-sulaiman, and B Gupta. 2018. Edge Computing with Cloud for Voice Disor-Der Assessment and Treatment. IEEE Commun. Mag 56 (4): 6065. [Google Scholar] [CrossRef] [Scilit]
- Oliveira, Brigada F. C., Deborah M. V. Magalhaes, Daniel S. Ferreira, and Fatima N. S. Medeiros. 2020. Combined Sustained Vowels Improve the Performance of the Haar Wavelet for Pathological Voice Characterization. In 2020 International Conference on Systems, Signals and Image Processing (IWSSIP), 381–86. IEEE. [CrossRef] [Scilit]
- Reddy, Mittapalle Kiran, Paavo Alku, and Krothapalli Sreenivasa Rao. 2020. Detection of Specific Language Impairment in Children Using Glottal Source Features. IEEE Access 8: 15273–79. [Google Scholar] [CrossRef] [Scilit]
- Sharma, Yogesh, and Bikesh Kumar Singh. 2020. Prediction of Specific Language Impairment in Children Using Speech Linear Predictive Coding Coefficients. In 2020 First International Conference on Power, Control and Computing Technologies (ICPC2T), 305–10. IEEE. [CrossRef] [Scilit]
- Shia, S. Emerald, and T. Jayasree. 2017. Detection of Pathological Voices Using Discrete Wavelet Transform and Artificial Neural Networks. In Proceedings of the 2017 IEEE International Conference on Intelligent Techniques in Control, Optimization and Signal Processing, INCOS 2017, 1–6. [CrossRef] [Scilit]
- Sri, K, Rama Murty, and B Yegnanarayana. 2008. Epoch Extraction From Speech Signals. IEEE Trans. Audio, Speech, Lang Process 16 (8). [Google Scholar] [CrossRef] [Scilit]
- Tulics, Miklos Gabriel, Gyorgy Szaszak, Krisztina Meszaros, and Klara Vicsi. 2019. Artificial Neural Network and SVM Based Voice Disorder Classification. In 2019 10th IEEE International Conference on Cognitive Infocommunications (CogInfoCom), 307–12. IEEE. [CrossRef] [Scilit]
- Tulics, Miklos Gabriel, Gyorgy Szaszak, Krisztina Meszaros, and Klara Vicsi. 2020. Using ASR Posterior Probability and Acoustic Features for Voice Disorder Classification. In 2020 11th IEEE International Conference on Cognitive Infocommunications (CogInfoCom), 000155–60. IEEE. [CrossRef] [Scilit]
- Verde, Laura, Giuseppe De Pietro, Mubarak Alrashoud, Ahmed Ghoneim, Khaled N. Al-Mutib, and Giovanna Sannino. 2019. Leveraging Artificial Intelligence to Improve Voice Disorder Identification Through the Use of a Reliable Mobile App. IEEE Access 7: 124048–54. [Google Scholar] [CrossRef] [Scilit]
- Verdolini, K, C Rosen, and R Branski. 2006. Classification Manual for Voice Disorders. Edited by I. Mahwah. Lawrence Erlbaum. [Google Scholar]
- Wu, Huiyi, John Soraghan, Anja Lowit, and Gaetano Di Caterina. 2018. A Deep Learning Method for Pathological Voice Detection Using Convolutional Deep Belief Networks, September, 446–50. [Google Scholar]
Disclaimer/Publisher’s Note: The statements, opinions and data contained in all publications are solely those of the individual author(s) and contributor(s) and not of MDPI and/or the editor(s). MDPI and/or the editor(s) disclaim responsibility for any injury to people or property resulting from any ideas, methods, instructions or products referred to in the content. |
Copyright © 2020. This article is licensed under a Creative Commons Attribution 4.0 International License.
Share and Cite
Sigona, F. Automatic detection of Voice Disorders: recent literature advancements. J. Interdiscip. Res. Appl. Med. 2020, 4, 21-30. https://doi.org/10.1285/i25327518v4i2p21
Sigona F. Automatic detection of Voice Disorders: recent literature advancements. Journal of Interdisciplinary Research Applied to Medicine. 2020; 4(2):21-30. https://doi.org/10.1285/i25327518v4i2p21
Chicago/Turabian StyleSigona, Francesco. 2020. "Automatic detection of Voice Disorders: recent literature advancements" Journal of Interdisciplinary Research Applied to Medicine 4, no. 2: 21-30. https://doi.org/10.1285/i25327518v4i2p21
APA StyleSigona, F. (2020). Automatic detection of Voice Disorders: recent literature advancements. Journal of Interdisciplinary Research Applied to Medicine, 4(2), 21-30. https://doi.org/10.1285/i25327518v4i2p21
