Speech technology

Speech Synthesis and Recognition

HG4052 introduces the linguistic, acoustic, and computational foundations of speech recognition and synthesis for linguistics students, from the speech chain and DTW/HMMs to CTC, Whisper, and neural TTS, with hands-on Praat and Colab practicals and a focus on Singapore's multilingual speech environment.

Working with large-scale speech corpus for phonetic research: Pipeline and tools

Introduction to pipeline and tools for phonetic research using large-scale speech corpora.

Training your first ASR model: Introduction to ASR in linguistic research

Introduction to Automatic Speech Recognition (ASR) in linguistic research.

Person-specific Automatic Speaker Recognition

The ESRC-funded project "Person-specific automatic speaker recognition : understanding the behaviour of individuals for applications of ASR" is a three year project running from 2022 to 2025 led by Dr Vincent Hughes (PI), Professor Paul Foulkes (CI) and Dr Philip Harrison in the Department of Language and Linguistic Science at the University of York. The project involves collaboration with the Netherlands Forensic Institute and Oxford Wave Research.