ICDAR 2021 Competition on Scientific Table Image Recognition to LaTeX

05/30/2021
by   Pratik Kayal, et al.
0

Tables present important information concisely in many scientific documents. Visual features like mathematical symbols, equations, and spanning cells make structure and content extraction from tables embedded in research documents difficult. This paper discusses the dataset, tasks, participants' methods, and results of the ICDAR 2021 Competition on Scientific Table Image Recognition to LaTeX. Specifically, the task of the competition is to convert a tabular image to its corresponding LaTeX source code. We proposed two subtasks. In Subtask 1, we ask the participants to reconstruct the LaTeX structure code from an image. In Subtask 2, we ask the participants to reconstruct the LaTeX content code from an image. This report describes the datasets and ground truth specification, details the performance evaluation metrics used, presents the final results, and summarizes the participating methods. Submission by team VCGroup got the highest Exact Match accuracy score of 74 for Subtask 2, beating previous baselines by 5 improvements can still be made to the recognition capabilities of models, this competition contributes to the development of fully automated table recognition systems by challenging practitioners to solve problems under specific constraints and sharing their approaches; the platform will remain available for post-challenge submissions at https://competitions.codalab.org/competitions/26979 .

READ FULL TEXT

page 1

page 2

page 3

page 4

research
10/31/2022

Tables to LaTeX: structure and content extraction from scientific tables

Scientific documents contain tables that list important information in a...
research
05/05/2021

PingAn-VCGroup's Solution for ICDAR 2021 Competition on Scientific Table Image Recognition to Latex

This paper presents our solution for the ICDAR 2021 Competition on Scien...
research
06/08/2021

ICDAR 2021 Competition on Scientific Literature Parsing

Scientific literature contain important information related to cutting-e...
research
05/12/2021

TabLeX: A Benchmark Dataset for Structure and Content Information Extraction from Scientific Tables

Information Extraction (IE) from the tables present in scientific articl...
research
05/05/2021

PingAn-VCGroup's Solution for ICDAR 2021 Competition on Scientific Literature Parsing Task B: Table Recognition to HTML

This paper presents our solution for ICDAR 2021 competition on scientifi...
research
12/15/2022

The First IEEE UV2022 Mathematical Modelling Competition: Backgrounds and Problems

Economic growth, people's health, and urban development face challenges ...
research
09/30/2021

Scientific evidence extraction

Recently, interest has grown in applying machine learning to the problem...

Please sign up or login with your details

Forgot password? Click here to reset