探索英国口音：用函数数据分析建模Trap-Bath分裂

Exploring British Accents: Modelling the Trap–Bath Split with Functional Data Analysis

Journal of the Royal Statistical Society. Series C: Applied Statistics · 2022

被引 5

ABS 3

Aranya Koshy 通讯
Shahin Tavakoli

中文导读

研究了英国南北口音中‘class’等词中元音发音的差异，用函数数据分析和广义加性模型对语音曲线建模，并可视化地理口音变化。

Abstract

Abstract The sound of our speech is influenced by the places we come from. Great Britain contains a wide variety of distinctive accents which are of interest to linguistics. In particular, the ‘a’ vowel in words like ‘class’ is pronounced differently in the North and the South. Speech recordings of this vowel can be represented as formant curves or as mel-frequency cepstral coefficient curves. Functional data analysis and generalised additive models offer techniques to model the variation in these curves. Our first aim was to model the difference between typical Northern and Southern vowels /æ/ and /ɑ/, by training two classifiers on the North-South Class Vowels dataset collected for this paper. Our second aim is to visualise geographical variation of accents in Great Britain. For this we use speech recordings from a second dataset, the British National Corpus (BNC) audio edition. The trained models are used to predict the accent of speakers in the BNC, and then we model the geographical patterns in these predictions using a soap film smoother. This work demonstrates a flexible and interpretable approach to modelling phonetic accent variation in speech recordings.

语言学语音识别方言学函数数据分析

阅读原文 ↗