IIIT-AR-13K: A New Dataset for Graphical Object Detection in Documents

08/06/2020
by   Ajoy Mondal, et al.
0

We introduce a new dataset for graphical object detection in business documents, more specifically annual reports. This dataset, IIIT-AR-13k, is created by manually annotating the bounding boxes of graphical or page objects in publicly available annual reports. This dataset contains a total of 13k annotated page images with objects in five different popular categories - table, figure, natural image, logo, and signature. It is the largest manually annotated dataset for graphical object detection. Annual reports created in multiple languages for several years from various companies bring high diversity into this dataset. We benchmark IIIT-AR-13K dataset with two state of the art graphical object detection techniques using Faster R-CNN [20] and Mask R-CNN [11] and establish high baselines for further research. Our dataset is highly effective as training data for developing practical solutions for graphical object detection in both business documents and technical articles. By training with IIIT-AR-13K, we demonstrate the feasibility of a single solution that can report superior performance compared to the equivalent ones trained with a much larger amount of data, for table detection. We hope that our dataset helps in advancing the research for detecting various types of graphical objects in business documents.

READ FULL TEXT

page 1

page 2

page 3

page 4

research
11/14/2022

Marine Microalgae Detection in Microscopy Images: A New Dataset

Marine microalgae are widespread in the ocean and play a crucial role in...
research
04/18/2018

Falling Things: A Synthetic Dataset for 3D Object Detection and Pose Estimation

We present a new dataset, called Falling Things (FAT), for advancing the...
research
11/02/2018

The Open Images Dataset V4: Unified image classification, object detection, and visual relationship detection at scale

We present Open Images V4, a dataset of 9.2M images with unified annotat...
research
07/21/2022

Omni3D: A Large Benchmark and Model for 3D Object Detection in the Wild

Recognizing scenes and objects in 3D from a single image is a longstandi...
research
04/07/2023

V3Det: Vast Vocabulary Visual Detection Dataset

Recent advances in detecting arbitrary objects in the real world are tra...
research
05/04/2023

Revisiting Table Detection Datasets for Visually Rich Documents

Table Detection has become a fundamental task for visually rich document...
research
03/06/2022

Detection of Parasitic Eggs from Microscopy Images and the emergence of a new dataset

Automatic detection of parasitic eggs in microscopy images has the poten...

Please sign up or login with your details

Forgot password? Click here to reset