HTMOT : Hierarchical Topic Modelling Over Time

11/22/2021
by   Judicael Poumay, et al.
0

Over the years, topic models have provided an efficient way of extracting insights from text. However, while many models have been proposed, none are able to model topic temporality and hierarchy jointly. Modelling time provide more precise topics by separating lexically close but temporally distinct topics while modelling hierarchy provides a more detailed view of the content of a document corpus. In this study, we therefore propose a novel method, HTMOT, to perform Hierarchical Topic Modelling Over Time. We train HTMOT using a new implementation of Gibbs sampling, which is more efficient. Specifically, we show that only applying time modelling to deep sub-topics provides a way to extract specific stories or events while high level topics extract larger themes in the corpus. Our results show that our training procedure is fast and can extract accurate high-level topics and temporally precise sub-topics. We measured our model's performance using the Word Intrusion task and outlined some limitations of this evaluation method, especially for hierarchical models. As a case study, we focused on the various developments in the space industry in 2020.

READ FULL TEXT
research
08/12/2020

Neural Sinkhorn Topic Model

In this paper, we present a new topic modelling approach via the theory ...
research
02/03/2016

"Draw My Topics": Find Desired Topics fast from large scale of Corpus

We develop the "Draw My Topics" toolkit, which provides a fast way to in...
research
02/23/2017

Scalable Inference for Nested Chinese Restaurant Process Topic Models

Nested Chinese Restaurant Process (nCRP) topic models are powerful nonpa...
research
05/23/2022

Artificial intelligence for topic modelling in Hindu philosophy: mapping themes between the Upanishads and the Bhagavad Gita

A distinct feature of Hindu religious and philosophical text is that the...
research
10/03/2014

Probit Normal Correlated Topic Models

The logistic normal distribution has recently been adapted via the trans...
research
05/16/2023

HyHTM: Hyperbolic Geometry based Hierarchical Topic Models

Hierarchical Topic Models (HTMs) are useful for discovering topic hierar...
research
06/30/2018

A Constrained Coupled Matrix-Tensor Factorization for Learning Time-evolving and Emerging Topics

Topic discovery has witnessed a significant growth as a field of data mi...

Please sign up or login with your details

Forgot password? Click here to reset