Home > Research > Publications & Outputs > Assessing the accuracy of existing forced align...

Associated organisational unit

Electronic data

  • Forced alignment paper - Vanguard

    Accepted author manuscript, 3.38 MB, PDF document

    Available under license: CC BY-NC: Creative Commons Attribution-NonCommercial 4.0 International License

Links

Text available via DOI:

View graph of relations

Assessing the accuracy of existing forced alignment software on varieties of British English

Research output: Contribution to Journal/MagazineJournal articlepeer-review

Published
Article number20180061
<mark>Journal publication date</mark>29/01/2020
<mark>Journal</mark>Linguistics Vanguard
Issue numbers1
Volume6
Publication StatusPublished
<mark>Original language</mark>English

Abstract

This paper presents an analysis of the performance and usability of automatic speech processing tools on six different varieties of English spoken in the British Isles. The tools used in the present study were developed for use with Mainstream American English, but we demonstrate that their forced alignment func- tionality nonetheless performs extremely well on a range of British varieties, encompassing both careful and casual speech. Where phone boundary placement is concerned, substantial errors in alignment occur infre- quently, and although small displacements between aligner-placed and human-placed phone boundaries are found regularly, these will rarely have a significant effect on measurements of interest for the researcher. We demonstrate that gross phone boundary placement errors, when they do arise, are particularly likely to be introduced in fast speech or with varieties that are radically different from Mainstream American English (e.g. Scots). We also observe occasional problems with phonetic transcription. Overall, we advise that, although forced alignment software is highly reliable and improving continuously, human confirmation is needed to correct errors which can displace entire stretches of speech. Nevertheless, sociolinguists can be assured that the output of these tools is generally highly accurate for a wide range of varieties.