4.4 Article Proceedings Paper

Time-sensitive clinical concept embeddings learned from large electronic health records

Journal

Publisher

BMC
DOI: 10.1186/s12911-019-0766-3

Keywords

Clinical concept embedding; Distributional representation; Time sensitive concept embedding; Electronic medical records; Concept similarity; Predictive modeling

Funding

  1. National Library of Medicine of the National Institutes of Health [U01TR002062, 1R01AI130460, 1R01LM011829]
  2. Cancer Prevention and Research Institute of Texas [U01TR002062, RP170668]

Ask authors/readers for more resources

BackgroundLearning distributional representation of clinical concepts (e.g., diseases, drugs, and labs) is an important research area of deep learning in the medical domain. However, many existing relevant methods do not consider temporal dependencies along the longitudinal sequence of a patient's records, which may lead to incorrect selection of contexts.MethodsTo address this issue, we extended three popular concept embedding learning methods: word2vec, positive pointwise mutual information (PPMI) and FastText, to consider time-sensitive information. We then trained them on a large electronic health records (EHR) database containing about 50 million patients to generate concept embeddings and evaluated them for both intrinsic evaluations focusing on concept similarity measure and an extrinsic evaluation to assess the use of generated concept embeddings in the task of predicting disease onset.ResultsOur experiments show that embeddings learned from information within one visit (time window zero) improve performance on the concept similarity measure and the FastText algorithm usually had better performance than the other two algorithms. For the predictive modeling task, the optimal result was achieved by word2vec embeddings with a 30-day sliding window.ConclusionsConsidering time constraints are important in training clinical concept embeddings. We expect they can benefit a series of downstream applications.

Authors

I am an author on this paper
Click your name to claim this paper and add it to your profile.

Reviews

Primary Rating

4.4
Not enough ratings

Secondary Ratings

Novelty
-
Significance
-
Scientific rigor
-
Rate this paper

Recommended

No Data Available
No Data Available