Learning Social Networks from Text Data using Covariate Information

10/16/2020
by   Xiaoyi Yang, et al.
0

Describing and characterizing the impact of historical figures can be challenging, but unraveling their social structures perhaps even more so. Historical social network analysis methods can help and may also illuminate people who have been overlooked by historians but turn out to be influential social connection points. Text data, such as biographies, can be a useful source of information about the structure of historical social networks but can also introduce challenges in identifying links. The Local Poisson Graphical Lasso model leverages the number of co-mentions in the text to measure relationships between people and uses a conditional independence structure to model a social network. This structure will reduce the tendency to overstate the relationship between "friends of friends", but given the historical high frequency of common names, without additional distinguishing information, we can still introduce incorrect links. In this work, we extend the Local Poisson Graphical Lasso model with a (multiple) penalty structure that incorporates covariates giving increased link probabilities to people with shared covariate information. We propose both greedy and Bayesian approaches to estimate the penalty parameters. We present results on data simulated with characteristics of historical networks and show that this type of penalty structure can improve network recovery as measured by precision and recall. We also illustrate the approach on biographical data of individuals who lived in early modern Britain, targeting the period from 1500 to 1575.

READ FULL TEXT

page 1

page 2

page 3

page 4

research
01/23/2022

Perceptual Effects of Hierarchy in Art Historical Social Networks

Network representation is a crucial topic in historical social network a...
research
08/25/2020

Complicating the Social Networks for Better Storytelling: An Empirical Study of Chinese Historical Text and Novel

Digital humanities is an important subject because it enables developmen...
research
12/20/2018

A Survey of Hierarchy Identification in Social Networks

Humans are social by nature. Throughout history, people have formed comm...
research
04/08/2015

Mining and discovering biographical information in Difangzhi with a language-model-based approach

We present results of expanding the contents of the China Biographical D...
research
04/01/2022

Detecting changes in dynamic social networks using multiply-labeled movement data

The social structure of an animal population can often influence movemen...
research
04/25/2022

A new preferential model with homophily for recommender systems

"Rich-get-richer" and "homophily" are two important phenomena in evolvin...
research
03/10/2020

Disrupting Resilient Criminal Networks through Data Analysis: The case of Sicilian Mafia

Compared to other types of social networks, criminal networks present ha...

Please sign up or login with your details

Forgot password? Click here to reset