Modernizing Historical Documents: a User Study

07/01/2019
by   Miguel Domingo, et al.
0

Accessibility to historical documents is mostly limited to scholars. This is due to the language barrier inherent in human language and the linguistic properties of these documents. Given a historical document, modernization aims to generate a new version of it, written in the modern version of the document's language. Its goal is to tackle the language barrier, decreasing the comprehension difficulty and making historical documents accessible to a broader audience. In this work, we proposed a new neural machine translation approach that profits from modern documents to enrich its systems. We tested this approach with both automatic and human evaluation, and conducted a user study. Results showed that modernization is successfully reaching its goal, although it still has room for improvement.

READ FULL TEXT

page 1

page 2

page 3

page 4

research
10/08/2019

An Interactive Machine Translation Framework for Modernizing Historical Documents

Due to the nature of human language, historical documents are hard to co...
research
02/02/2021

Two Demonstrations of the Machine Translation Applications to Historical Documents

We present our demonstration of two machine translation applications to ...
research
05/20/2022

Translating Hanja historical documents to understandable Korean and English

The Annals of Joseon Dynasty (AJD) contain the daily records of the King...
research
06/12/2021

Predicting the Ordering of Characters in Japanese Historical Documents

Japan is a unique country with a distinct cultural heritage, which is re...
research
02/26/2023

User-Centric Evaluation of OCR Systems for Kwak'wala

There has been recent interest in improving optical character recognitio...
research
04/13/2021

Restoring and Mining the Records of the Joseon Dynasty via Neural Language Modeling and Machine Translation

Understanding voluminous historical records provides clues on the past i...
research
09/04/2020

Externalizing Transformations of Historical Documents: Opportunities for Provenance-Driven Visualization

Transcription, annotation, digitization and/or visualization are common ...

Please sign up or login with your details

Forgot password? Click here to reset