Recurrent Space-time Graphs for Video Understanding

04/11/2019
by   Andrei Nicolicioiu, et al.
0

Visual learning in the space-time domain remains a very challenging problem in artificial intelligence. Current computational models for understanding video data are heavily rooted in the classical single-image based paradigm. It is not yet well understood how to integrate visual information from space and time into a single, general model. We propose a neural graph model, recurrent in space and time, suitable for capturing both the appearance and the complex interactions of different entities and objects within the changing world scene. Nodes and links in our graph have dedicated neural networks for processing information. Edges process messages between connected nodes at different locations and scales or between past and present time. Nodes compute over features extracted from local parts in space and time and over messages received from their neighbours and previous memory states. Messages are passed iteratively in order to transmit information globally and establish long range interactions. Our model is general and could learn to recognize a variety of high level spatio-temporal concepts and be applied to different learning tasks. We demonstrate, through extensive experiments, a competitive performance over strong baselines on the tasks of recognizing complex patterns of movement in video.

READ FULL TEXT

page 2

page 7

research
06/05/2018

Videos as Space-Time Region Graphs

How do humans recognize the action "opening a book" ? We argue that ther...
research
03/15/2022

Learning Spatio-Temporal Downsampling for Effective Video Upscaling

Downsampling is one of the most basic image processing operations. Impro...
research
09/07/2021

Improving Phenotype Prediction using Long-Range Spatio-Temporal Dynamics of Functional Connectivity

The study of functional brain connectivity (FC) is important for underst...
research
10/17/2016

Spatio-Temporal Attention Models for Grounded Video Captioning

Automatic video captioning is challenging due to the complex interaction...
research
03/18/2021

Ano-Graph: Learning Normal Scene Contextual Graphs to Detect Video Anomalies

Video anomaly detection has proved to be a challenging task owing to its...
research
08/01/2019

Evaluating an Immersive Space-Time Cube Geovisualization for Intuitive Trajectory Data Exploration

A Space-Time Cube enables analysts to clearly observe spatio-temporal fe...
research
07/11/2016

Efficient Activity Detection in Untrimmed Video with Max-Subgraph Search

We propose an efficient approach for activity detection in video that un...

Please sign up or login with your details

Forgot password? Click here to reset