The Kriston AI System for the VoxCeleb Speaker Recognition Challenge 2022

09/23/2022

∙

This technical report describes our system for track 1, 2 and 4 of the VoxCeleb Speaker Recognition Challenge 2022 (VoxSRC-22). By combining several ResNet variants, our submission for track 1 attained a minDCF of 0:090 with EER 1:401 submission for track 2 achieved a minDCF of 0:072 with EER 1:119 our system consisted of voice activity detection (VAD), speaker embedding extraction, agglomerative hierarchical clustering (AHC) followed by a re-clustering step based on a Bayesian hidden Markov model and overlapped speech detection and handling. Our submission for track 4 achieved a diarisation error rate (DER) of 4.86 places for the corresponding tracks.

READ FULL TEXT

The Kriston AI System for the VoxCeleb Speaker Recognition Challenge 2022

Sign in with Google

Consider DeepAI Pro