Self-Selective Context for Interaction Recognition

10/17/2020
by   Mert Kilickaya, et al.
8

Human-object interaction recognition aims for identifying the relationship between a human subject and an object. Researchers incorporate global scene context into the early layers of deep Convolutional Neural Networks as a solution. They report a significant increase in the performance since generally interactions are correlated with the scene (riding bicycle on the city street). However, this approach leads to the following problems. It increases the network size in the early layers, therefore not efficient. It leads to noisy filter responses when the scene is irrelevant, therefore not accurate. It only leverages scene context whereas human-object interactions offer a multitude of contexts, therefore incomplete. To circumvent these issues, in this work, we propose Self-Selective Context (SSC). SSC operates on the joint appearance of human-objects and context to bring the most discriminative context(s) into play for recognition. We devise novel contextual features that model the locality of human-object interactions and show that SSC can seamlessly integrate with the State-of-the-art interaction recognition models. Our experiments show that SSC leads to an important increase in interaction recognition performance, while using much fewer parameters.

READ FULL TEXT

page 1

page 4

page 7

research
10/17/2019

Deep Contextual Attention for Human-Object Interaction Detection

Human-object interaction detection is an important and relatively new cl...
research
12/17/2021

Distillation of Human-Object Interaction Contexts for Action Recognition

Modeling spatial-temporal relations is imperative for recognizing human ...
research
07/05/2022

Distance Matters in Human-Object Interaction Detection

Human-Object Interaction (HOI) detection has received considerable atten...
research
05/16/2023

Learning Higher-order Object Interactions for Keypoint-based Video Understanding

Action recognition is an important problem that requires identifying act...
research
04/07/2023

Devil's on the Edges: Selective Quad Attention for Scene Graph Generation

Scene graph generation aims to construct a semantic graph structure from...
research
06/20/2018

iMapper: Interaction-guided Joint Scene and Human Motion Mapping from Monocular Videos

A long-standing challenge in scene analysis is the recovery of scene arr...
research
03/25/2016

Early Detection of Combustion Instabilities using Deep Convolutional Selective Autoencoders on Hi-speed Flame Video

This paper proposes an end-to-end convolutional selective autoencoder ap...

Please sign up or login with your details

Forgot password? Click here to reset