MGTR: End-to-End Mutual Gaze Detection with Transformer

09/22/2022
by   Hang Guo, et al.
0

People's looking at each other or mutual gaze is ubiquitous in our daily interactions, and detecting mutual gaze is of great significance for understanding human social scenes. Current mutual gaze detection methods focus on two-stage methods, whose inference speed is limited by the two-stage pipeline and the performance in the second stage is affected by the first one. In this paper, we propose a novel one-stage mutual gaze detection framework called Mutual Gaze TRansformer or MGTR to perform mutual gaze detection in an end-to-end manner. By designing mutual gaze instance triples, MGTR can detect each human head bounding box and simultaneously infer mutual gaze relationship based on global image information, which streamlines the whole process with simplicity. Experimental results on two mutual gaze datasets show that our method is able to accelerate mutual gaze detection process without losing performance. Ablation study shows that different components of MGTR can capture different levels of semantic information in images. Code is available at https://github.com/Gmbition/MGTR

READ FULL TEXT

page 9

page 14

research
10/15/2020

Boosting Image-based Mutual Gaze Detection using Pseudo 3D Gaze

Mutual gaze detection, i.e., predicting whether or not two people are lo...
research
06/06/2023

Human-Object Interaction Prediction in Videos through Gaze Following

Understanding the human-object interactions (HOIs) from a video is essen...
research
02/16/2023

Social Visual Behavior Analytics for Autism Therapy of Children Based on Automated Mutual Gaze Detection

Social visual behavior, as a type of non-verbal communication, plays a c...
research
04/12/2021

Glance and Gaze: Inferring Action-aware Points for One-Stage Human-Object Interaction Detection

Modern human-object interaction (HOI) detection approaches can be divide...
research
08/23/2022

Multimodal Across Domains Gaze Target Detection

This paper addresses the gaze target detection problem in single images ...
research
08/08/2022

In the Eye of Transformer: Global-Local Correlation for Egocentric Gaze Estimation

In this paper, we present the first transformer-based model to address t...
research
06/12/2019

LAEO-Net: revisiting people Looking At Each Other in videos

Capturing the `mutual gaze' of people is essential for understanding and...

Please sign up or login with your details

Forgot password? Click here to reset