Residual Recurrent CRNN for End-to-End Optical Music Recognition on Monophonic Scores

by   Aozhi Liu, et al.

Optical Music Recognition is a field that attempts to extract digital information from images of either the printed music scores or the handwritten music scores. One of the challenges of the Optical Music Recognition task is to transcript the symbols of the camera-captured images into digital music notations. Previous end-to-end model, based on deep learning, was developed as a Convolutional Recurrent Neural Network. However, it does not explore sufficient contextual information from full scales and there is still a large room for improvement. In this paper, we propose an innovative end-to-end framework that combines a block of Residual Recurrent Convolutional Neural Network with a recurrent Encoder-Decoder network to map a sequence of monophonic music symbols corresponding to the notations present in the image. The Residual Recurrent Convolutional block can improve the ability of the model to enrich the context information while the number of parameter will not be increasing. The experiment results were benchmarked against a publicly available dataset called CAMERA-PRIMUS. We evaluate the performances of our model on both the images with ideal conditions and that with non-ideal conditions. The experiments show that our approach surpass the state-of-the-art end-to-end method using Convolutional Recurrent Neural Network.



There are no comments yet.


page 1

page 2

page 3

page 4


An Empirical Evaluation of End-to-End Polyphonic Optical Music Recognition

Previous work has shown that neural architectures are able to perform op...

Deep Watershed Detector for Music Object Recognition

Optical Music Recognition (OMR) is an important and challenging area wit...

Optical Music Recognition with Convolutional Sequence-to-Sequence Models

Optical Music Recognition (OMR) is an important technology within Music ...

A holistic approach to polyphonic music transcription with neural networks

We present a framework based on neural networks to extract music scores ...

Handwritten digit string recognition by combination of residual network and RNN-CTC

Recurrent neural network (RNN) and connectionist temporal classification...

DPDnet: A Robust People Detector using Deep Learning with an Overhead Depth Camera

In this paper we propose a method based on deep learning that detects mu...

Transcribing Content from Structural Images with Spotlight Mechanism

Transcribing content from structural images, e.g., writing notes from mu...
This week in AI

Get the week's most popular data science and artificial intelligence research sent straight to your inbox every Saturday.