Article
ToKSA - Tokenized Key Sentence Annotation - a Novel Method for Rapid Approximation of Ground Truth for Natural Language Processing
2021-10-07
Abstract excerpt
<h4>ABSTRACT</h4> <h4>Objective</h4> Identifying phenotypes and pathology from free text is an essential task for clinical work and research. Natural language processing (NLP) is a key tool for processing free text at scale. Developing and validating NLP models requires labelled data. Labels are generated through time-consuming and repetitive manual annotation and are hard to obtain for sensitive clinical data. Th...
Topics
Open a Topic to create a Post that cites this publication.
Identifiers and source
- Literature Corpus work
- 0f09e96f-e4d3-5e04-ac6a-b8469eddffe5
- DOI
- 10.1101/2021.10.06.21264629
