Wake Word Detection Based on Res2Net

09/30/2022
by   Qiuchen Yu, et al.
0

This letter proposes a new wake word detection system based on Res2Net. As a variant of ResNet, Res2Net was first applied to objection detection. Res2Net realizes multiple feature scales by increasing possible receptive fields. This multiple scaling mechanism significantly improves the detection ability of wake words with different durations. Compared with the ResNet-based model, Res2Net also significantly reduces the model size and is more suitable for detecting wake words. The proposed system can determine the positions of wake words from the audio stream without any additional assistance. The proposed method is verified on the Mobvoi dataset containing two wake words. At a false alarm rate of 0.5 per hour, the system reduced the false rejection of the two wake words by more than 12

READ FULL TEXT
research
10/28/2020

Replay and Synthetic Speech Detection with Res2net Architecture

Existing approaches for replay and synthetic speech detection still lack...
research
08/18/2022

Walking on Words

Take any word over some alphabet. If it is non-empty, go to any position...
research
10/02/2022

The lexicographically least square-free word with a given prefix

The lexicographically least square-free infinite word on the alphabet of...
research
10/24/2022

I see what you hear: a vision-inspired method to localize words

This paper explores the possibility of using visual object detection tec...
research
02/04/2020

3D ResNet with Ranking Loss Function for Abnormal Activity Detection in Videos

Abnormal activity detection is one of the most challenging tasks in the ...
research
07/22/2020

PhishZip: A New Compression-based Algorithm for Detecting Phishing Websites

Phishing has grown significantly in the past few years and is predicted ...

Please sign up or login with your details

Forgot password? Click here to reset