ProtoSound: A Personalized and Scalable Sound Recognition System for Deaf and Hard-of-Hearing Users

02/22/2022
by   Dhruv Jain, et al.
0

Recent advances have enabled automatic sound recognition systems for deaf and hard of hearing (DHH) users on mobile devices. However, these tools use pre-trained, generic sound recognition models, which do not meet the diverse needs of DHH users. We introduce ProtoSound, an interactive system for customizing sound recognition models by recording a few examples, thereby enabling personalized and fine-grained categories. ProtoSound is motivated by prior work examining sound awareness needs of DHH people and by a survey we conducted with 472 DHH participants. To evaluate ProtoSound, we characterized performance on two real-world sound datasets, showing significant improvement over state-of-the-art (e.g., +9.7 deployed ProtoSound's end-user training and real-time recognition through a mobile application and recruited 19 hearing participants who listened to the real-world sounds and rated the accuracy across 56 locations (e.g., homes, restaurants, parks). Results show that ProtoSound personalized the model on-device in real-time and accurately learned sounds across diverse acoustic contexts. We close by discussing open challenges in personalizable sound recognition, including the need for better recording interfaces and algorithmic improvements.

READ FULL TEXT

page 2

page 13

page 14

page 16

research
10/19/2018

Mobile Sound Recognition for the Deaf and Hard of Hearing

Human perception of surrounding events is strongly dependent on audio cu...
research
06/17/2019

Towards Real-Time Action Recognition on Mobile Devices Using Deep Models

Action recognition is a vital task in computer vision, and many methods ...
research
06/17/2020

ExSampling: a system for the real-time ensemble performance of field-recorded environmental sounds

We propose ExSampling: an integrated system of recording application and...
research
04/17/2022

Advances in Thunder Sound Synthesis

A recent comparative study evaluated all known thunder synthesis techniq...
research
07/02/2022

UserLibri: A Dataset for ASR Personalization Using Only Text

Personalization of speech models on mobile devices (on-device personaliz...
research
02/25/2020

IoT Based Real Time Noise Mapping System for Urban Sound Pollution Study

This paper describes the development of a system that enables real time ...
research
08/04/2022

Impact Makes a Sound and Sound Makes an Impact: Sound Guides Representations and Explorations

Sound is one of the most informative and abundant modalities in the real...

Please sign up or login with your details

Forgot password? Click here to reset