Search engine for discovering works of Art, research articles, and books related to Art and Culture
ShareThis
Javascript must be enabled to continue!

A Survey of Automatic Speech Recognition for Dysarthric Speech

View through CrossRef
Dysarthric speech has several pathological characteristics, such as discontinuous pronunciation, uncontrolled volume, slow speech, explosive pronunciation, improper pauses, excessive nasal sounds, and air-flow noise during pronunciation, which differ from healthy speech. Automatic speech recognition (ASR) can be very helpful for speakers with dysarthria. Our research aims to provide a scoping review of ASR for dysarthric speech, covering papers in this field from 1990 to 2022. Our survey found that the development of research studies about the acoustic features and acoustic models of dysarthric speech is nearly synchronous. During the 2010s, deep learning technologies were widely applied to improve the performance of ASR systems. In the era of deep learning, many advanced methods (such as convolutional neural networks, deep neural networks, and recurrent neural networks) are being applied to design acoustic models and lexical and language models for dysarthric-speech-recognition tasks. Deep learning methods are also used to extract acoustic features from dysarthric speech. Additionally, this scoping review found that speaker-dependent problems seriously limit the generalization applicability of the acoustic model. The scarce available speech data cannot satisfy the amount required to train models using big data.
Title: A Survey of Automatic Speech Recognition for Dysarthric Speech
Description:
Dysarthric speech has several pathological characteristics, such as discontinuous pronunciation, uncontrolled volume, slow speech, explosive pronunciation, improper pauses, excessive nasal sounds, and air-flow noise during pronunciation, which differ from healthy speech.
Automatic speech recognition (ASR) can be very helpful for speakers with dysarthria.
Our research aims to provide a scoping review of ASR for dysarthric speech, covering papers in this field from 1990 to 2022.
Our survey found that the development of research studies about the acoustic features and acoustic models of dysarthric speech is nearly synchronous.
During the 2010s, deep learning technologies were widely applied to improve the performance of ASR systems.
In the era of deep learning, many advanced methods (such as convolutional neural networks, deep neural networks, and recurrent neural networks) are being applied to design acoustic models and lexical and language models for dysarthric-speech-recognition tasks.
Deep learning methods are also used to extract acoustic features from dysarthric speech.
Additionally, this scoping review found that speaker-dependent problems seriously limit the generalization applicability of the acoustic model.
The scarce available speech data cannot satisfy the amount required to train models using big data.

Related Results

Empowering Dysarthric Speech: Leveraging Advanced LLMs for Accurate Speech Correction and Multimodal Emotion Analysis
Empowering Dysarthric Speech: Leveraging Advanced LLMs for Accurate Speech Correction and Multimodal Emotion Analysis
Dysarthria or Dysarthric speech as called is kind of a motor speech disorder which is caused by neurological damage that affects the muscles for speech production, which results in...
Enhancing dysarthric speech recognition through SepFormer and hierarchical attention network models with multistage transfer learning
Enhancing dysarthric speech recognition through SepFormer and hierarchical attention network models with multistage transfer learning
AbstractDysarthria, a motor speech disorder that impacts articulation and speech clarity, presents significant challenges for Automatic Speech Recognition (ASR) systems. This study...
Recent Advances in Dysarthric Speech Recognition: Approaches and Datasets
Recent Advances in Dysarthric Speech Recognition: Approaches and Datasets
Dysarthria is a neuromotor speech disorder that results from physical disability and limits speech intelligibility. Dysarthric speakers can make use of speech recognition systems t...
A Comprehensive Survey of Automatic Dysarthric Speech Recognition
A Comprehensive Survey of Automatic Dysarthric Speech Recognition
Automatic dysarthric speech recognition (DSR) is very crucial for many human computer interaction systems that enables the human to interact with machine in natural way. The object...
Pola Komunikasi Interpersonal Terapis Wicara pada Anak Telambat Bicara (Speech Delay)
Pola Komunikasi Interpersonal Terapis Wicara pada Anak Telambat Bicara (Speech Delay)
Abstract. The tittle of this study is Interpersonal Communication Patterns of Speech Therapists in Children with Speech Delay. Children with speech delay disorders are also social ...
AFM signal model for dysarthric speech classification using speech biomarkers
AFM signal model for dysarthric speech classification using speech biomarkers
Neurological disorders include various conditions affecting the brain, spinal cord, and nervous system which results in reduced performance in different organs and muscles througho...
Automatic speech recognition in voice-speech rehabilitation effectiveness evaluation in patients after laryngectomy
Automatic speech recognition in voice-speech rehabilitation effectiveness evaluation in patients after laryngectomy
Introduction. Lost voice function compensation determines the personal and social life of laryngectomees. Automatic speech recognition and synthesis methods are...
Ensemble Gaussian mixture model-based special voice command cognitive computing intelligent system
Ensemble Gaussian mixture model-based special voice command cognitive computing intelligent system
Dysarthria is a speech disorder caused by stroke, Parkinson’s disease, neurological injury, or tumors that damage the nervous system and weaken the speech quality. Developing a uni...

Back to Top