Abstract
As speech datasets used in sociolinguistic research increase in size, laborious and time-intensive manual orthographic transcription is a challenge, limiting the amount of (transcribed) data which can be analysed. In this paper, I discuss the use of (commercial) automatic speech recognition (ASR) as a tool in sociolinguistic research in the context of a case study: the Lothian Diary Project. I describe the kinds of errors produced by two commercial ASR systems for British English within the broader context of algorithmic bias in ASR, and suggest some best practices when working with ASR in sociolinguistic work.
| Original language | English |
|---|---|
| Title of host publication | University of Pennsylvania Working Papers in Linguistics |
| Subtitle of host publication | Selected Papers from NWAV 49 |
| Publisher | Penn Graduate Linguistics Society |
| Number of pages | 12 |
| Volume | 28 |
| Edition | 2 |
| Publication status | Published - 19 Sept 2022 |
| Event | 49th meeting of New Ways of Analyzing Variation - The University of Texas, Austin, United States Duration: 19 Oct 2021 → 24 Oct 2021 Conference number: 49 https://www.nwav49.org/ |
Conference
| Conference | 49th meeting of New Ways of Analyzing Variation |
|---|---|
| Abbreviated title | NWAV 49 |
| Country/Territory | United States |
| City | Austin |
| Period | 19/10/21 → 24/10/21 |
| Internet address |
Fingerprint
Dive into the research topics of '(Commercial) Automatic Speech Recognition as a Tool in Sociolinguistic Research'. Together they form a unique fingerprint.Cite this
- APA
- Author
- BIBTEX
- Harvard
- Standard
- RIS
- Vancouver