Unsupervised Discovery of Clinical Disease Signatures Using Probabilistic Independence

التفاصيل البيبلوغرافية
العنوان: Unsupervised Discovery of Clinical Disease Signatures Using Probabilistic Independence
المؤلفون: Lasko, Thomas A., Still, John M., Li, Thomas Z., Mota, Marco Barbero, Stead, William W., Strobl, Eric V., Landman, Bennett A., Maldonado, Fabien
سنة النشر: 2024
المجموعة: Computer Science
Statistics
مصطلحات موضوعية: Computer Science - Machine Learning, Statistics - Applications, Statistics - Machine Learning, I.2.6, I.2.1, J.3
الوصف: Insufficiently precise diagnosis of clinical disease is likely responsible for many treatment failures, even for common conditions and treatments. With a large enough dataset, it may be possible to use unsupervised machine learning to define clinical disease patterns more precisely. We present an approach to learning these patterns by using probabilistic independence to disentangle the imprint on the medical record of causal latent sources of disease. We inferred a broad set of 2000 clinical signatures of latent sources from 9195 variables in 269,099 Electronic Health Records. The learned signatures produced better discrimination than the original variables in a lung cancer prediction task unknown to the inference algorithm, predicting 3-year malignancy in patients with no history of cancer before a solitary lung nodule was discovered. More importantly, the signatures' greater explanatory power identified pre-nodule signatures of apparently undiagnosed cancer in many of those patients.
Comment: 29 Pages, 8 figures
نوع الوثيقة: Working Paper
الوصول الحر: http://arxiv.org/abs/2402.05802Test
رقم الانضمام: edsarx.2402.05802
قاعدة البيانات: arXiv