Potential use of text classification tools as signatures of suicidal behavior: A proof-of-concept study using Virginia Woolf's personal writings

利用文本分类工具识别自杀行为特征的潜在用途:以弗吉尼亚·伍尔夫的个人作品为例的概念验证研究

阅读:2

Abstract

BACKGROUND: The present study analyzes the feasibility of text classification to predict individual suicidal behavior. Entries from Virginia Woolf's diaries and letters were used to assess whether a text classification algorithm could identify written patterns associated with suicide. METHODS: This is a text classification study. We compared 46 text entries from the two months before Virginia Woolf's suicide with 54 texts randomly selected from Virginia Woolf's work during other periods of her life. Letters and diaries were included, while books, novels, short stories, and article fragments were excluded. The data was analyzed using a Naïve-Bayes machine-learning algorithm. RESULTS: The model showed a balanced accuracy of 80.45%, sensitivity of 69%, and specificity of 91%. The Kappa statistic was 0.6, which means a good agreement, and the p-value of the model was 0.003. The area under the ROC curve (AUC) was 0.80. In other words, the model exhibited good performance when used for classifying Virginia Woolf's diaries and letters. DISCUSSION: The present study showed the feasibility of a machine-learning model coupled with text to identify individual written patterns associated with suicidal behavior. Our text signature was able to identify the period of two months preceding suicide with a high accuracy. This technique may be applied to subjects with psychiatric disorders by means of data captured from social media, e-mail, among others. The algorithm may then predict a specific outcome and enable early intervention by clinicians.

特别声明

1、本页面内容包含部分的内容是基于公开信息的合理引用;引用内容仅为补充信息,不代表本站立场。

2、若认为本页面引用内容涉及侵权,请及时与本站联系,我们将第一时间处理。

3、其他媒体/个人如需使用本页面原创内容,需注明“来源:[生知库]”并获得授权;使用引用内容的,需自行联系原作者获得许可。

4、投稿及合作请联系:info@biocloudy.com。