Frontiers in Psychology (Jan 2024)

Simon Fraser University Speech Error Database (SFUSED) Cantonese: Methods, design, and usage

  • John Alderete

DOI
https://doi.org/10.3389/fpsyg.2024.1270433
Journal volume & issue
Vol. 15

Abstract

Read online

The Simon Fraser University Speech Error Database (SFUSED) is a multi-purpose database of speech errors based in audio recordings. The motivation for SFUSED Cantonese, a component of this database, is to create a linguistically rich data set for exploring language production processes in Cantonese, an under-studied language. We describe in detail the methods used to collect, analyze, and explore the database, including details of team workflows, time budgets, data quality, and explicit linguistic and processing assumptions. In addition to showing how to use the database, this account supports future research with a template for investigating additional under-studied languages, and it gives fresh perspective on the benefits and drawbacks of collecting speech error data from spontaneous speech. All of the data and supporting materials are available as open access data sets.

Keywords