arXiv · 1701.08796
Learning from various labeling strategies for suicide-related messages on social media: An experimental study
Abstract
Suicide is an important but often misunderstood problem, one that researchers are now seeking to better understand through social media. Due in large part to the fuzzy nature of what constitutes suicidal risks, most supervised approaches for learning to automatically detect suicide-related activity in social media require a great deal of human labor to train. However, humans themselves have diverse or conflicting views on what constitutes suicidal thoughts. So how to obtain reliable gold standard labels is fundamentally challenging and, we hypothesize, depends largely on what is asked of the annotators and what slice of the data they label. We conducted multiple rounds of data labeling and collected annotations from crowdsourcing workers and domain experts. We aggregated the resulting labels in various ways to train a series of supervised models. Our preliminary evaluations show that using unanimously agreed labels from multiple annotators is helpful to achieve robust machine models.
Explore related subjects
Keep this discovery
Tong Liu, Qijin Cheng, Christopher M. Homan, Vincent M. B. Silenzio. 2017-01-30. Learning from various labeling strategies for suicide-related messages on social media: An experimental study. https://arxiv.org/abs/1701.08796
Cite the original work for its findings. Save a collection to share your selection of sources.