Personalized One-Shot Lipreading for an ALS Patient

11/02/2021
by   Bipasha Sen, et al.
0

Lipreading or visually recognizing speech from the mouth movements of a speaker is a challenging and mentally taxing task. Unfortunately, multiple medical conditions force people to depend on this skill in their day-to-day lives for essential communication. Patients suffering from Amyotrophic Lateral Sclerosis (ALS) often lose muscle control, consequently their ability to generate speech and communicate via lip movements. Existing large datasets do not focus on medical patients or curate personalized vocabulary relevant to an individual. Collecting a large-scale dataset of a patient, needed to train mod-ern data-hungry deep learning models is, however, extremely challenging. In this work, we propose a personalized network to lipread an ALS patient using only one-shot examples. We depend on synthetically generated lip movements to augment the one-shot scenario. A Variational Encoder based domain adaptation technique is used to bridge the real-synthetic domain gap. Our approach significantly improves and achieves high top-5accuracy with 83.2 compared to 62.6 evaluating our approach on the ALS patient, we also extend it to people with hearing impairment relying extensively on lip movements to communicate.

READ FULL TEXT

page 2

page 3

research
02/25/2022

Bridging the Gap Between Patient-specific and Patient-independent Seizure Prediction via Knowledge Distillation

Objective. Deep neural networks (DNN) have shown unprecedented success i...
research
07/13/2020

Designing Personalized Interaction of a Socially Assistive Robot for Stroke Rehabilitation Therapy

The research of a socially assistive robot has a potential to augment an...
research
11/23/2020

Automated Quality Assessment of Hand Washing Using Deep Learning

Washing hands is one of the most important ways to prevent infectious di...
research
05/12/2023

Design, Development, and Evaluation of an Interactive Personalized Social Robot to Monitor and Coach Post-Stroke Rehabilitation Exercises

Socially assistive robots are increasingly being explored to improve the...
research
05/17/2020

Learning Individual Speaking Styles for Accurate Lip to Speech Synthesis

Humans involuntarily tend to infer parts of the conversation from lip mo...
research
03/10/2021

Automatic Speaker Independent Dysarthric Speech Intelligibility Assessment System

Dysarthria is a condition which hampers the ability of an individual to ...
research
08/21/2020

Domain Adaptation of Learned Features for Visual Localization

We tackle the problem of visual localization under changing conditions, ...

Please sign up or login with your details

Forgot password? Click here to reset