Title
Natural language model for automatic identification of Intimate Partner Violence reports from Twitter
Abstract
Intimate partner violence (IPV) is a preventable public health problem that affects millions of people worldwide. Approximately one in four women are estimated to be or have been victims of severe violence at some point in their lives, irrespective of age, ethnicity, and economic status. Victims often report IPV experiences on social media, and automatic detection of such reports via machine learning may enable improved surveillance and targeted distribution of support and/or interventions for those in need. However, no artificial intelligence systems for automatic detection currently exists, and we attempted to address this research gap. We collected posts from Twitter using a list of IPV-related keywords, manually reviewed subsets of retrieved posts, and prepared annotation guidelines to categorize tweets into IPV-report or non-IPV-report. We annotated 6,348 tweets in total, with the inter-annotator agreement (IAA) of 0.86 (Cohen's kappa) among 1,834 double-annotated tweets. The class distribution in the annotated dataset was highly imbalanced, with only 668 posts (∼11%) labeled as IPV-report. We then developed an effective natural language processing model to identify IPV-reporting tweets automatically. The developed model achieved classification F1-scores of 0.76 for the IPV-report class and 0.97 for the non-IPV-report class. We conducted post-classification analyses to determine the causes of system errors and to ensure that the system did not exhibit biases in its decision making, particularly with respect to race and gender. Our automatic model can be an essential component for a proactive social media-based intervention and support framework, while also aiding population-level surveillance and large-scale cohort studies.
Year
DOI
Venue
2022
10.1016/j.array.2022.100217
Array
Keywords
DocType
Volume
Intimate partner violence,Domestic violence,Natural language processing,Machine learning,Social media
Journal
15
ISSN
Citations 
PageRank 
2590-0056
0
0.34
References 
Authors
0
7
Name
Order
Citations
PageRank
Mohammed Ali Al-Garadi122.41
Sangmi Kim200.34
Yuting Guo300.34
Elise Warren400.34
Yuan-Chi Yang511.72
Sahithi Lakamana600.34
Abeed Sarker7607.38