Search engine for discovering works of Art, research articles, and books related to Art and Culture
ShareThis
Javascript must be enabled to continue!

Statistical learning models of early phonetic acquisition struggle with child-centered audio data

View through CrossRef
Infants learn their native language(s) at an amazing speed. Before they even talk, their perception adapts to the language(s) they hear. However, the mechanisms responsible for this perceptual attunement still remain unclear. A long tradition in linguistics points to the importance of specialized language mechanisms that would allow us to quickly and effortlessly learn from the language(s) we are exposed to. However, the currently dominant explanation for perceptual attunement posits that infants apply a domain-general learning mechanism consisting in learning statistical regularities from the speech stream they hear, and which may be found in learning across domains and across species. Critically, the feasibility of employing purely domain-general statistical learning mechanisms has only been demonstrated with computational models on unrealistic and simplified input. This paper presents the first attempt to study perceptual attunement with 2,000 hours of ecological child-centered recordings in American English and Metropolitan French. We show that, when applied on ecologically-valid data, generic learning mechanisms develop a language-relevant perceptual space but fail to show evidence for perceptual attunement. It is only when supplemented with domain-specific audio filtering and augmentation mechanisms that computational models show a significant attunement to the language they have been exposed to. Hence, we conclude that, when learning from ecological audio, domain-specific mechanisms may be necessary to guide early language learning in the wild even if the learning itself is done through generic mechanisms. We anticipate our work to be a starting point for ecologically-valid computational models of perceptual attunement in other domains and species.
Title: Statistical learning models of early phonetic acquisition struggle with child-centered audio data
Description:
Infants learn their native language(s) at an amazing speed.
Before they even talk, their perception adapts to the language(s) they hear.
However, the mechanisms responsible for this perceptual attunement still remain unclear.
A long tradition in linguistics points to the importance of specialized language mechanisms that would allow us to quickly and effortlessly learn from the language(s) we are exposed to.
However, the currently dominant explanation for perceptual attunement posits that infants apply a domain-general learning mechanism consisting in learning statistical regularities from the speech stream they hear, and which may be found in learning across domains and across species.
Critically, the feasibility of employing purely domain-general statistical learning mechanisms has only been demonstrated with computational models on unrealistic and simplified input.
This paper presents the first attempt to study perceptual attunement with 2,000 hours of ecological child-centered recordings in American English and Metropolitan French.
We show that, when applied on ecologically-valid data, generic learning mechanisms develop a language-relevant perceptual space but fail to show evidence for perceptual attunement.
It is only when supplemented with domain-specific audio filtering and augmentation mechanisms that computational models show a significant attunement to the language they have been exposed to.
Hence, we conclude that, when learning from ecological audio, domain-specific mechanisms may be necessary to guide early language learning in the wild even if the learning itself is done through generic mechanisms.
We anticipate our work to be a starting point for ecologically-valid computational models of perceptual attunement in other domains and species.

Related Results

No Sudden Audio Switch – Preventing discontinuous POI audio playing in LBS
No Sudden Audio Switch – Preventing discontinuous POI audio playing in LBS
Abstract. Many LBS applications provide automatic audio playing functions for introducing POI’s. Appropriate automatic audio playing can improve users’ expressions during traveling...
Pshal P’shaw
Pshal P’shaw
Pshal P’shaw investigates the sonic instability of speech—where phonetic dissonance, vocal fragmentation, and gestural sound challenge structured linguistic norms. Developed during...
Feature selection for multimodal: acoustic event detection
Feature selection for multimodal: acoustic event detection
The detection of the Acoustic Events (AEs) naturally produced in a meeting room may help to describe the human and social activity. The automatic description of interactions betwee...
Selection of Injectable Drug Product Composition using Machine Learning Models (Preprint)
Selection of Injectable Drug Product Composition using Machine Learning Models (Preprint)
BACKGROUND As of July 2020, a Web of Science search of “machine learning (ML)” nested within the search of “pharmacokinetics or pharmacodynamics” yielded over 100...
The Impact of Father Involvement in the Early Childhood Problematic Behavior
The Impact of Father Involvement in the Early Childhood Problematic Behavior
Father's involvement is something that influences the child's problematic behavior. The purpose of this study is to investigate whether father involvement can influence children's ...
CREATING LEARNING MEDIA IN TEACHING ENGLISH AT SMP MUHAMMADIYAH 2 PAGELARAN ACADEMIC YEAR 2020/2021
CREATING LEARNING MEDIA IN TEACHING ENGLISH AT SMP MUHAMMADIYAH 2 PAGELARAN ACADEMIC YEAR 2020/2021
The pandemic Covid-19 currently demands teachers to be able to use technology in teaching and learning process. But in reality there are still many teachers who have not been able ...
Do parent posttraumatic stress symptoms (PTSS) predict later child PTSS?
Do parent posttraumatic stress symptoms (PTSS) predict later child PTSS?
Background: Although research suggests that parent and child PTSS are associated, the magnitude of this association at different time points and in the context of various covariate...
EFL Students' Perspective In Learning Phonetic Symbols
EFL Students' Perspective In Learning Phonetic Symbols
Phonetic symbols are symbols used to explain how a sound is formed. Phonetic symbols can help students explain the different sounds of various English words. The purpose of this st...

Back to Top