Evidential value of voice quality acoustics in forensic voice comparison
Project Summary
Forensic voice comparison typically involves comparing a known voice of a suspect with an unknown voice of an offender. This process is crucial in cases involving speech recordings, such as hoax calls, ransom demands, threatening messages, or conversations with accomplices. In legal cases where speech recordings are involved (e.g. hoax calls, ransom demands, threatening voice message, conversation with an accomplice), the comparison of voices may assist the trier-of-fact (e.g. judge/jury) or investigating authorities (e.g. police) in determining whether the two voices originate from the same speaker or different speakers.
A key objective of forensic voice comparison research is to identify speech features that are useful for distinguishing voices. Voice quality (VQ) features (e.g. breathy/creaky/hoarse/nasalised voice) have been considered among the most useful for speaker comparison. However, VQ is typically assessed by experts through subjective auditory analysis and categorized qualitatively. This project was among the first to evaluate the evidential strength of VQ acoustics, marking an important step towards transforming the categorical auditory assessment of VQ features into quantitative analysis which fosters transparency and replicability. Analysis using the Bayesian likelihood ratio framework shows that, contrary to the widely held belief in the utility of VQ analysis, acoustic VQ parameters offer limited speaker-discriminatory value, especially when both speech style mismatch and non-contemporaneous recordings were involved.