Improved Relation Networks for End-to-End Speaker Verification and Identification

03/31/2022
by   Ashutosh Chaubey, et al.
0

Speaker identification systems in a real-world scenario are tasked to identify a speaker amongst a set of enrolled speakers given just a few samples for each enrolled speaker. This paper demonstrates the effectiveness of meta-learning and relation networks for this use case. We propose improved relation networks for speaker verification and few-shot (unseen) speaker identification. The use of relation networks facilitates joint training of the frontend speaker encoder and the backend model. Inspired by the use of prototypical networks in speaker verification and to increase the discriminability of the speaker embeddings, we train the model to classify samples in the current episode amongst all speakers present in the training set. Furthermore, we propose a new training regime for faster model convergence by extracting more information from a given meta-learning episode with negligible extra computation. We evaluate the proposed techniques on VoxCeleb, SITW and VCTK datasets on the tasks of speaker verification and unseen speaker identification. The proposed approach outperforms the existing approaches consistently on both tasks.

READ FULL TEXT

page 1

page 2

page 3

page 4

research
07/31/2020

Designing Neural Speaker Embeddings with Meta Learning

Neural speaker embeddings trained using classification objectives have d...
research
04/06/2020

Meta-Learning for Short Utterance Speaker Recognition with Imbalance Length Pairs

In realistic settings, a speaker recognition system needs to identify a ...
research
04/14/2020

Kinship Identification through Joint Learning Using Kinship Verification Ensemble

While kinship verification is a well-exploited task which only identifie...
research
03/29/2021

Improved Meta-learning training for Speaker Verification

Meta-learning (ML) has recently become a research hotspot in speaker ver...
research
09/20/2021

Improving Text-Independent Speaker Verification with Auxiliary Speakers Using Graph

The paper presents a novel approach to refining similarity scores betwee...
research
06/01/2023

Speaker-specific Thresholding for Robust Imposter Identification in Unseen Speaker Recognition

Speaker identification systems are deployed in diverse environments, oft...
research
10/24/2019

Meta-learning for robust child-adult classification from speech

Computational modeling of naturalistic conversations in clinical applica...

Please sign up or login with your details

Forgot password? Click here to reset