Unsupervised Part-of-Speech Induction

01/10/2018
by   Omid Kashefi, et al.
0

Part-of-Speech (POS) tagging is an old and fundamental task in natural language processing. While supervised POS taggers have shown promising accuracy, it is not always feasible to use supervised methods due to lack of labeled data. In this project, we attempt to unsurprisingly induce POS tags by iteratively looking for a recurring pattern of words through a hierarchical agglomerative clustering process. Our approach shows promising results when compared to the tagging results of the state-of-the-art unsupervised POS taggers.

READ FULL TEXT

page 1

page 2

page 3

page 4

research
03/09/2017

Turkish PoS Tagging by Reducing Sparsity with Morpheme Tags in Small Datasets

Sparsity is one of the major problems in natural language processing. Th...
research
05/24/2017

Joint PoS Tagging and Stemming for Agglutinative Languages

The number of word forms in agglutinative languages is theoretically inf...
research
01/22/2016

GeoTextTagger: High-Precision Location Tagging of Textual Documents using a Natural Language Processing Approach

Location tagging, also known as geotagging or geolocation, is the proces...
research
08/04/2020

Reliable Part-of-Speech Tagging of Historical Corpora through Set-Valued Prediction

Syntactic annotation of corpora in the form of part-of-speech (POS) tags...
research
09/20/2020

Persian Ezafe Recognition Using Transformers and Its Role in Part-Of-Speech Tagging

Ezafe is a grammatical particle in some Iranian languages that links two...
research
03/27/2023

ACO-tagger: A Novel Method for Part-of-Speech Tagging using Ant Colony Optimization

Swarm Intelligence algorithms have gained significant attention in recen...
research
06/11/2020

RTEX: A novel methodology for Ranking, Tagging, and Explanatory diagnostic captioning of radiography exams

This paper introduces RTEx, a novel methodology for a) ranking radiograp...

Please sign up or login with your details

Forgot password? Click here to reset