Perception of child-produced Polish sibilants: a comparison of native English speakers and Polish Heritage speakers
Bibliographic record
Abstract
The Polish language has a complex sibilant structure when compared to languages like English. Of particular interest here are the alveolo-palatal and retroflex sibilants. There have been some previous studies on Polish sibilants examining production and perception of children (under 5 years). However, there is a greater need for understanding adult perception of children’s productions and the perception of different populations listening to children’s productions. Contributing to perception studies would, therefore, allow for a more in-depth analysis of this field of research. This paper builds on the findings of a production-perception study of Polish sibilants in typical children (Zygis et al., 2023) and expands the results by examining English and Heritage Polish population perceptions of Polish children’s productions. The Zygis et al. study examined Polish children and their production and perception of the contrasting sibilants. The study looked at the perception of the children for their own production and adults’ productions. Their study acquired recordings of 80 Polish children aged 35–106 months producing words with /s, ʂ, ɕ/. One of their tasks involved the child participants hearing their own productions of word-medial sibilants: /kasa/, /kaʂa/ and, /kaɕa/ at random. They then had to choose between three images (corresponding to Polish words, e.g.: kasa for cash register) to indicate the stimuli they heard. Their study found that there were a number of acoustic parameters that children used to identify sibilants. They observed that especially the younger children, “appear [to] pay more attention to formants independent of the sibilant and [that] the cue weighting [for these young children] changes during the acquisition process” (Zygis et al., 2023). For the present study, we wanted to explore the perception of these word-medial sibilants for different phonetic environments and for non-native listener populations. The three phonetic conditions included: the whole word as in the original study, the isolated sibilant, and the (isolated) sibilant together with the preceding vowel. The audio files (taken from Zygis et al, 2023) were edited and played to both native English and Polish Heritage listeners at McMaster University in Hamilton, to determine the perception of the three-way Polish sibilant distinction. This distinction is non-existent in English for English listeners or influenced by both Heritage and English phonetics/phonology for Heritage speakers. The sibilant distinction in English lies between /s/ and /ʃ/, therefore the task for the English native participants was to choose between buttons that indicated “kasa | as | s” (for the /s/ sibilant) or “kasha | ash | sh” (for /ʃ/) to indicate which sibilant they perceived. The Heritage speakers of Polish were English participants with varying levels of Polish fluency residing in the Southern Ontario area. They used the same design (three-way sibilant distinction) as the original study. A total of 41 English and 13 Heritage listeners participated in the study. It was hypothesized that the English native listeners would categorize all Polish alveolars as (English) alveolars, but it was not clear how retroflex and alveolo-palatal contrasts from the children’s complex productions would be resolved by the English listeners. It was further assumed that the perception of stimuli with vowel transitions (e.g., /kasa/ and /as/ in contrast to isolated /s/) would significantly differ comparing English listeners and Polish Heritage listeners. In our results, English participants increasingly categorized all manipulations of /s/ as /s/, and /ɕ/ as the /ʃ/ sibilant, especially for the older children’s productions. Their perceptions for the retroflex /ʂ/ was split, half as /s/ perceptions, across conditions. Phonetic information in the form of formants (on top of the spectral noise of the isolated sibilant) did not significantly improve distinction for the English participants. The Polish Heritage speakers showed difficulty in correctly identifying /ʂ/ variations especially in the older children. Phonetic environment and age had varying effects depending on the sibilant. As Polish Heritage participants are familiar with three-way sibilant contrasts, it was interesting to see how these Heritage speakers’ classification differed from that of English participants, especially for stimuli from children who are in the very initial stages of speech development (i.e., decreased articulatory and acoustic accuracy).
Fetched live from OpenAlex and de-inverted. Abstracts are not stored in this database: the inverted indexes are 8.6 GB of the frame’s 9.3 GB of text, and the host has 13 GB free.
How this classification was reachedexpand
Full frame distilled prediction
Teacher imitationNot calibrated prevalence, not ground truth. Human validation pending. Learned from the 10,348 direct Codex labels and 10,348 direct Gemma labels. Candidate is the union of thresholded teacher heads; consensus is their intersection. These outputs are machine_predicted_unvalidated and are not human labels or direct frontier model labels.
Codex and Gemma teacher scores by category
| Category | Codex | Gemma |
|---|---|---|
| Metaresearch | 0.000 | 0.000 |
| Meta-epidemiology (narrow) | 0.000 | 0.000 |
| Meta-epidemiology (broad) | 0.001 | 0.000 |
| Bibliometrics | 0.000 | 0.001 |
| Science and technology studies | 0.000 | 0.000 |
| Scholarly communication | 0.000 | 0.000 |
| Open science | 0.000 | 0.000 |
| Research integrity | 0.000 | 0.001 |
| Insufficient payload (model declined to judge) | 0.013 | 0.000 |
Machine scores (provisional)
The two teacher heads of the student model, read on this work. A score orders the frame for review; it never asserts a category, and the validation status ships verbatim with every row.
Baseline scores from an immature model (maturity gate not passed, 7 training rounds). Scores rank; they never assert a category.
score_only:v0-immature-baseline · verbatim from the scoring run: score_only means the number may rank works, and no category label ships from itClassification
machine, unvalidatedMachine predicted; a candidate call from one teacher head, not a consensus.
How this classification was reached, model by model and score by score, is at the end of the page under "How this classification was reached".