• FluTrackers.com Inc. does not provide medical advice. Information on this web site is collected from various internet resources, and the FluTrackers board of directors makes no warranty to the safety, efficacy, correctness or completeness of the information posted on this site by any author or poster. The information collated here is for instructional and/or discussion purposes only and is NOT intended to diagnose or treat any disease, illness, or other medical condition. Every individual reader or poster should seek advice from their personal physician/healthcare practitioner before considering or using any interventions that are discussed on this website. By continuing to access this website you agree to consult your personal physican before using any interventions posted on this website, and you agree to hold harmless FluTrackers.com Inc., the board of directors, the members, and all authors and posters for any effects from use of any medication, supplement, vitamin or other substance, device, intervention, etc. mentioned in posts on this website, or other internet venues referenced in posts on this website.
  • We are not asking for any donations. Do not donate to any entity who says they are raising funds for us.

Comparison of machine learning classifiers for influenza detection from emergency department free-text reports

tetano

Editor, Senior Moderator
J Biomed Inform. 2015 Sep 16. pii: S1532-0464(15)00187-2. doi: 10.1016/j.jbi.2015.08.019. [Epub ahead of print]
[h=1]Comparison of machine learning classifiers for influenza detection from emergency department free-text reports.[/h] L?pez Pineda A[SUP]1[/SUP], Ye Y[SUP]1[/SUP], Visweswaran S[SUP]1[/SUP], Cooper GF[SUP]1[/SUP], Wagner MM[SUP]1[/SUP], Rich Tsui F[SUP]2[/SUP].
[h=3]Author information[/h]

[h=3]Abstract[/h] Influenza is a yearly recurrent disease that has the potential to become a pandemic. An effective biosurveillance system is required for early detection of the disease. In our previous studies, we have shown that electronic Emergency Department (ED) free-text reports can be of value to improve influenza detection in real time. This paper studies seven machine learning (ML) classifiers for influenza detection, compares their diagnostic capabilities against an expert-built influenza Bayesian classifier, and evaluates different ways of handling missing clinical information from the free-text reports. We randomly identified 31,268 ED reports from 4 hospitals between 2008 and 2011 to form two different datasets: training (468 cases, 29,004 controls), and test (176 cases and 1620 controls). We employed Topaz, a natural language processing (NLP) tool, to extract influenza-related findings and to encode them into one of three values: Acute, non-acute, and missing. Results show that all ML classifiers had areas under ROCs (AUC) ranging from 0.88 to 0.93, and performed significantly better than the expert-built Bayesian model. Missing clinical information marked as a value of missing (not missing at random) had a consistently improved performance among 3 (out of 4) ML classifiers when it was compared with the configuration of not assigning a value of missing (missing completely at random). The case/control ratios did not affect the classification performance given the large number of training cases. Our study demonstrates ED reports in conjunction with the use of ML and NLP with the handling of missing value information have a great potential for the detection of infectious diseases.
Copyright ? 2015 Elsevier Inc. All rights reserved.


[h=4]KEYWORDS:[/h] Bayesian; Case detection; Emergency department reports; Influenza; Machine learning

PMID: 26385375 [PubMed - as supplied by publisher]
 
Back
Top Bottom