Back to search

Article

Building a Best-in-Class Automated De-identification Tool for Electronic Health Records Through Ensemble Learning

2020-12-23

Abstract excerpt

The natural language portions of electronic health records (EHRs) communicate critical information about disease and treatment progression. However, the presence of personally identifiable information (PII) in this data constrains its broad reuse. Despite continuous improvements in methods for the automated detection of PII, the presence of residual identifiers in clinical notes requires manual validation and corr...

Topics

Open a Topic to create a Post that cites this publication.

Identifiers and source

Literature Corpus work
52b9f22d-07ad-5cef-9d3c-41a8a9b24f8c
DOI
10.1101/2020.12.22.20248270
Open publication

Related research

Semantic proximity does not establish scientific evidence.

Click a neighbor to travelStep 1 · 12 closest
Interactive article relationship graphSelect a related publication card to move it into the centre and load its closest explainable connections. Solid lines are source-backed structured connections. Dashed lines are semantic discovery signals and are not scientific evidence.
Building a Best-in-Class Automated De-identification Tool for Electronic Health Records Through Ensemble LearningDOI 10.1101/2020.12.22.20248270
Select a neighboring publication to make it the new centre.