Generating SOAP Notes from Doctor-Patient Conversations

05/04/2020
by   Kundan Krishna, et al.
29

Following each patient visit, physicians must draft detailed clinical summaries called SOAP notes. Moreover, with electronic health records, these notes must be digitized. For all the benefits of this documentation the process remains onerous, contributing to increasing physician burnout. In a parallel development, patients increasingly record audio from their visits (with consent), often through dedicated apps. In this paper, we present the first study to evaluate complete pipelines for leveraging these transcripts to train machine learning model to generate these notes. We first describe a unique dataset of patient visit records, consisting of transcripts, paired SOAP notes, and annotations marking noteworthy utterances that support each summary sentence. We decompose the problem into extractive and abstractive subtasks, exploring a spectrum of approaches according to how much they demand from each component. Our best performing method first (i) extracts noteworthy utterances via multi-label classification assigns them to summary section(s); (ii) clusters noteworthy utterances on a per-section basis; and (iii) generates the summary sentences by conditioning on the corresponding cluster and the subsection of the SOAP sentence to be generated. Compared to an end-to-end approach that generates the full SOAP note from the full conversation, our approach improves by 7 ROUGE-1 points. Oracle experiments indicate that fixing our generative capabilities, improvements in extraction alone could provide (up to) a further 9 ROUGE point gain.

READ FULL TEXT

page 12

page 14

research
08/08/2018

Learning to Write Notes in Electronic Health Records

Clinicians spend a significant amount of time inputting free-form textua...
research
07/14/2020

Extracting Structured Data from Physician-Patient Conversations By Predicting Noteworthy Utterances

Despite diverse efforts to mine various modalities of medical data, the ...
research
02/28/2022

PMC-Patients: A Large-scale Dataset of Patient Notes and Relations Extracted from Case Reports in PubMed Central

We present PMC-Patients, a dataset consisting of 167k patient notes with...
research
09/27/2017

Multi-Label Classification of Patient Notes a Case Study on ICD Code Assignment

In the context of the Electronic Health Record, automated diagnosis codi...
research
04/01/2022

Human Evaluation and Correlation with Automatic Metrics in Consultation Note Generation

In recent years, machine learning models have rapidly become better at g...
research
05/10/2023

A Method to Automate the Discharge Summary Hospital Course for Neurology Patients

Generation of automated clinical notes have been posited as a strategy t...
research
04/09/2020

Query-Focused EHR Summarization to Aid Imaging Diagnosis

Electronic Health Records (EHRs) provide vital contextual information to...

Please sign up or login with your details

Forgot password? Click here to reset