Can natural language processing models extract and classify instances of interpersonal violence in mental healthcare electronic records: an applied evaluative study. Issue 2 (16th February 2022)
- Record Type:
- Journal Article
- Title:
- Can natural language processing models extract and classify instances of interpersonal violence in mental healthcare electronic records: an applied evaluative study. Issue 2 (16th February 2022)
- Main Title:
- Can natural language processing models extract and classify instances of interpersonal violence in mental healthcare electronic records: an applied evaluative study
- Authors:
- Botelle, Riley
Bhavsar, Vishal
Kadra-Scalzo, Giouliana
Mascio, Aurelie
Williams, Marcus V
Roberts, Angus
Velupillai, Sumithra
Stewart, Robert - Abstract:
- Abstract : Objective: This paper evaluates the application of a natural language processing (NLP) model for extracting clinical text referring to interpersonal violence using electronic health records (EHRs) from a large mental healthcare provider. Design: A multidisciplinary team iteratively developed guidelines for annotating clinical text referring to violence. Keywords were used to generate a dataset which was annotated (ie, classified as affirmed, negated or irrelevant) for: presence of violence, patient status (ie, as perpetrator, witness and/or victim of violence) and violence type (domestic, physical and/or sexual). An NLP approach using a pretrained transformer model, BioBERT (Bidirectional Encoder Representations from Transformers for Biomedical Text Mining) was fine-tuned on the annotated dataset and evaluated using 10-fold cross-validation. Setting: We used the Clinical Records Interactive Search (CRIS) database, comprising over 500 000 de-identified EHRs of patients within the South London and Maudsley NHS Foundation Trust, a specialist mental healthcare provider serving an urban catchment area. Participants: Searches of CRIS were carried out based on 17 predefined keywords. Randomly selected text fragments were taken from the results for each keyword, amounting to 3771 text fragments from the records of 2832 patients. Outcome measures: We estimated precision, recall and F1 score for each NLP model. We examined sociodemographic and clinical variables in patientsAbstract : Objective: This paper evaluates the application of a natural language processing (NLP) model for extracting clinical text referring to interpersonal violence using electronic health records (EHRs) from a large mental healthcare provider. Design: A multidisciplinary team iteratively developed guidelines for annotating clinical text referring to violence. Keywords were used to generate a dataset which was annotated (ie, classified as affirmed, negated or irrelevant) for: presence of violence, patient status (ie, as perpetrator, witness and/or victim of violence) and violence type (domestic, physical and/or sexual). An NLP approach using a pretrained transformer model, BioBERT (Bidirectional Encoder Representations from Transformers for Biomedical Text Mining) was fine-tuned on the annotated dataset and evaluated using 10-fold cross-validation. Setting: We used the Clinical Records Interactive Search (CRIS) database, comprising over 500 000 de-identified EHRs of patients within the South London and Maudsley NHS Foundation Trust, a specialist mental healthcare provider serving an urban catchment area. Participants: Searches of CRIS were carried out based on 17 predefined keywords. Randomly selected text fragments were taken from the results for each keyword, amounting to 3771 text fragments from the records of 2832 patients. Outcome measures: We estimated precision, recall and F1 score for each NLP model. We examined sociodemographic and clinical variables in patients giving rise to the text data, and frequencies for each annotated violence characteristic. Results: Binary classification models were developed for six labels (violence presence, perpetrator, victim, domestic, physical and sexual). Among annotations affirmed for the presence of any violence, 78% (1724) referred to physical violence, 61% (1350) referred to patients as perpetrator and 33% (731) to domestic violence. NLP models' precision ranged from 89% (perpetrator) to 98% (sexual); recall ranged from 89% (victim, perpetrator) to 97% (sexual). Conclusions: State of the art NLP models can extract and classify clinical text on violence from EHRs at acceptable levels of scale, efficiency and accuracy. … (more)
- Is Part Of:
- BMJ open. Volume 12:Issue 2(2022)
- Journal:
- BMJ open
- Issue:
- Volume 12:Issue 2(2022)
- Issue Display:
- Volume 12, Issue 2 (2022)
- Year:
- 2022
- Volume:
- 12
- Issue:
- 2
- Issue Sort Value:
- 2022-0012-0002-0000
- Page Start:
- Page End:
- Publication Date:
- 2022-02-16
- Subjects:
- health informatics -- mental health -- psychiatry -- public health
Medicine -- Research -- Periodicals
610.72 - Journal URLs:
- http://www.bmj.com/archive ↗
http://bmjopen.bmj.com/ ↗ - DOI:
- 10.1136/bmjopen-2021-052911 ↗
- Languages:
- English
- ISSNs:
- 2044-6055
- Deposit Type:
- Legaldeposit
- View Content:
- Available online (eLD content is only available in our Reading Rooms) ↗
- Physical Locations:
- British Library DSC - BLDSS-3PM
British Library HMNTS - ELD Digital store - Ingest File:
- 20750.xml