DeepAI AI Chat
Log In Sign Up

EHRKit: A Python Natural Language Processing Toolkit for Electronic Health Record Texts

by   Irene Li, et al.

The Electronic Health Record (EHR) is an essential part of the modern medical system and impacts healthcare delivery, operations, and research. Unstructured text is attracting much attention despite structured information in the EHRs and has become an exciting research field. The success of the recent neural Natural Language Processing (NLP) method has led to a new direction for processing unstructured clinical notes. In this work, we create a python library for clinical texts, EHRKit. This library contains two main parts: MIMIC-III-specific functions and tasks specific functions. The first part introduces a list of interfaces for accessing MIMIC-III NOTEEVENTS data, including basic search, information retrieval, and information extraction. The second part integrates many third-party libraries for up to 12 off-shelf NLP tasks such as named entity recognition, summarization, machine translation, etc.


page 1

page 2

page 3

page 4


Two-stage Federated Phenotyping and Patient Representation Learning

A large percentage of medical information is in unstructured text format...

Neural Natural Language Processing for Unstructured Data in Electronic Health Records: a Review

Electronic health records (EHRs), digital collections of patient healthc...

Multilingual Medical Question Answering and Information Retrieval for Rural Health Intelligence Access

In rural regions of several developing countries, access to quality heal...

CREATE: Cohort Retrieval Enhanced by Analysis of Text from Electronic Health Records using OMOP Common Data Model

Background: Widespread adoption of electronic health records (EHRs) has ...

Adapting Pretrained Language Models for Solving Tabular Prediction Problems in the Electronic Health Record

We propose an approach for adapting the DeBERTa model for electronic hea...