Skip to main content
Start main content

Seminar | Exploring individual-level variability in automatic speaker recognition performance

Seminars / Lectures / Workshops

Seminar_23Sep_FB_X
  • Date

    23 Sep 2026

  • Organiser

    Department of English and Communication

  • Time

    17:00 - 18:00

  • Venue

    Online via Zoom  

Speaker

Prof. Vincent Hughes

Summary

In this talk, I will cover outcomes from the Person-specific Automatic Speaker Recognition (PASR) project led by the University of York, together with Oxford Wave Research and the Netherlands Forensic Institute. The talk will be in two parts. The first will explore automatic speaker recognition performance with a small, controlled set of data collected at York called SISTEM (Speech of Individual Speakers Across Time by Experts in Multiple Sessions). The corpus allows us to systematically assess when and how systems perform well and where they struggle. I will also discuss testing conducted with a very large, forensically realistic corpus provided by the UK Government. Our testing goes beyond the typical high-level metrics of overall system performance, to assess the behaviour of individual speakers and individual files within our dataset. The aim is to describe the range of speaker and file variability and then attempt to explain and even predict speakers and files that may produce poor performance. I will also report on testing conducted to assess the stability of individual speakers across generations of automatic speaker recognition systems.

Keynote Speaker

Prof. Vincent Hughes

Prof. Vincent Hughes

Professor, University of York, United Kingdom

Vincent Hughes is Professor of Forensic Speech Science in the Department of Language and Linguistic Science at the University of York. A central theme of his research involves assessing how individual voices differ from each other. This involves testing and validating phonetic and AI-based methods for forensic voice comparison and the statistical evaluation of expert forensic evidence. He has recently completed a large-scale ESRC-funded projects looking at the use of automatic speaker recognition systems as forensic evidence and has also led interdisciplinary projects focused on the individuality of the human voice. More recently, he has worked on applying forensic approaches to the analysis of deepfake audio and in the context of transcription in law enforcement. Professor Hughes is also co-lead of Forensic Speech Services at York, which bridges the gap between research and practice through forensic casework, CPD and training courses, consultancy services, and collaborative and contract research.

Your browser is not the latest version. If you continue to browse our website, Some pages may not function properly.

You are recommended to upgrade to a newer version or switch to a different browser. A list of the web browsers that we support can be found here