-
Predicting Emergency Department Visits for Patients with Type II Diabetes
Authors:
Javad M Alizadeh,
Jay S Patel,
Gabriel Tajeu,
Yuzhou Chen,
Ilene L Hollin,
Mukesh K Patel,
Junchao Fei,
Huanmei Wu
Abstract:
Over 30 million Americans are affected by Type II diabetes (T2D), a treatable condition with significant health risks. This study aims to develop and validate predictive models using machine learning (ML) techniques to estimate emergency department (ED) visits among patients with T2D. Data for these patients was obtained from the HealthShare Exchange (HSX), focusing on demographic details, diagnos…
▽ More
Over 30 million Americans are affected by Type II diabetes (T2D), a treatable condition with significant health risks. This study aims to develop and validate predictive models using machine learning (ML) techniques to estimate emergency department (ED) visits among patients with T2D. Data for these patients was obtained from the HealthShare Exchange (HSX), focusing on demographic details, diagnoses, and vital signs. Our sample contained 34,151 patients diagnosed with T2D which resulted in 703,065 visits overall between 2017 and 2021. A workflow integrated EMR data with SDoH for ML predictions. A total of 87 out of 2,555 features were selected for model construction. Various machine learning algorithms, including CatBoost, Ensemble Learning, K-nearest Neighbors (KNN), Support Vector Classification (SVC), Random Forest, and Extreme Gradient Boosting (XGBoost), were employed with tenfold cross-validation to predict whether a patient is at risk of an ED visit. The ROC curves for Random Forest, XGBoost, Ensemble Learning, CatBoost, KNN, and SVC, were 0.82, 0.82, 0.82, 0.81, 0.72, 0.68, respectively. Ensemble Learning and Random Forest models demonstrated superior predictive performance in terms of discrimination, calibration, and clinical applicability. These models are reliable tools for predicting risk of ED visits among patients with T2D. They can estimate future ED demand and assist clinicians in identifying critical factors associated with ED utilization, enabling early interventions to reduce such visits. The top five important features were age, the difference between visitation gaps, visitation gaps, R10 or abdominal and pelvic pain, and the Index of Concentration at the Extremes (ICE) for income.
△ Less
Submitted 12 December, 2024;
originally announced December 2024.
-
Self-supervised inter-intra period-aware ECG representation learning for detecting atrial fibrillation
Authors:
Xiangqian Zhu,
Mengnan Shi,
Xuexin Yu,
Chang Liu,
Xiaocong Lian,
Jintao Fei,
Jiangying Luo,
Xin Jin,
Ping Zhang,
Xiangyang Ji
Abstract:
Atrial fibrillation is a commonly encountered clinical arrhythmia associated with stroke and increased mortality. Since professional medical knowledge is required for annotation, exploiting a large corpus of ECGs to develop accurate supervised learning-based atrial fibrillation algorithms remains challenging. Self-supervised learning (SSL) is a promising recipe for generalized ECG representation l…
▽ More
Atrial fibrillation is a commonly encountered clinical arrhythmia associated with stroke and increased mortality. Since professional medical knowledge is required for annotation, exploiting a large corpus of ECGs to develop accurate supervised learning-based atrial fibrillation algorithms remains challenging. Self-supervised learning (SSL) is a promising recipe for generalized ECG representation learning, eliminating the dependence on expensive labeling. However, without well-designed incorporations of knowledge related to atrial fibrillation, existing SSL approaches typically suffer from unsatisfactory capture of robust ECG representations. In this paper, we propose an inter-intra period-aware ECG representation learning approach. Considering ECGs of atrial fibrillation patients exhibit the irregularity in RR intervals and the absence of P-waves, we develop specific pre-training tasks for interperiod and intraperiod representations, aiming to learn the single-period stable morphology representation while retaining crucial interperiod features. After further fine-tuning, our approach demonstrates remarkable AUC performances on the BTCH dataset, \textit{i.e.}, 0.953/0.996 for paroxysmal/persistent atrial fibrillation detection. On commonly used benchmarks of CinC2017 and CPSC2021, the generalization capability and effectiveness of our methodology are substantiated with competitive results.
△ Less
Submitted 8 October, 2024;
originally announced October 2024.
-
Graphical models for inferring single molecule dynamics
Authors:
Jonathan E. Bronson,
Jake M. Hofman,
Jingyi Fei,
Ruben L. Gonzalez Jr.,
Chris H. Wiggins
Abstract:
Background: The recent explosion of experimental techniques in single molecule biophysics has generated a variety of novel time series data requiring equally novel computational tools for analysis and inference. This article describes in general terms how graphical modeling may be used to learn from biophysical time series data using the variational Bayesian expectation maximization algorithm (VBE…
▽ More
Background: The recent explosion of experimental techniques in single molecule biophysics has generated a variety of novel time series data requiring equally novel computational tools for analysis and inference. This article describes in general terms how graphical modeling may be used to learn from biophysical time series data using the variational Bayesian expectation maximization algorithm (VBEM). The discussion is illustrated by the example of single-molecule fluorescence resonance energy transfer (smFRET) versus time data, where the smFRET time series is modeled as a hidden Markov model (HMM) with Gaussian observables. A detailed description of smFRET is provided as well. Results: The VBEM algorithm returns the model's evidence and an approximating posterior parameter distribution given the data. The former provides a metric for model selection via maximum evidence (ME), and the latter a description of the model's parameters learned from the data. ME/VBEM provide several advantages over the more commonly used approach of maximum likelihood (ML) optimized by the expectation maximization (EM) algorithm, the most important being a natural form of model selection and a well-posed (non-divergent) optimization problem. Conclusions: The results demonstrate the utility of graphical modeling for inference of dynamic processes in single molecule biophysics.
△ Less
Submitted 4 September, 2010;
originally announced September 2010.
-
Allosteric collaboration between elongation factor G and the ribosomal L1 stalk directs tRNA movements during translation
Authors:
Jingyi Fei,
Jonathan E. Bronson,
Jake M. Hofman,
Rathi L. Srinivas,
Chris H. Wiggins,
Ruben L. Gonzalez, Jr
Abstract:
Determining the mechanism by which transfer RNAs (tRNAs) rapidly and precisely transit through the ribosomal A, P and E sites during translation remains a major goal in the study of protein synthesis. Here, we report the real-time dynamics of the L1 stalk, a structural element of the large ribosomal subunit that is implicated in directing tRNA movements during translation. Within pre-translocati…
▽ More
Determining the mechanism by which transfer RNAs (tRNAs) rapidly and precisely transit through the ribosomal A, P and E sites during translation remains a major goal in the study of protein synthesis. Here, we report the real-time dynamics of the L1 stalk, a structural element of the large ribosomal subunit that is implicated in directing tRNA movements during translation. Within pre-translocation ribosomal complexes, the L1 stalk exists in a dynamic equilibrium between open and closed conformations. Binding of elongation factor G (EF-G) shifts this equilibrium towards the closed conformation through one of at least two distinct kinetic mechanisms, where the identity of the P-site tRNA dictates the kinetic route that is taken. Within post-translocation complexes, L1 stalk dynamics are dependent on the presence and identity of the E-site tRNA. Collectively, our data demonstrate that EF-G and the L1 stalk allosterically collaborate to direct tRNA translocation from the P to the E sites, and suggest a model for the release of E-site tRNA.
△ Less
Submitted 2 September, 2009;
originally announced September 2009.
-
Learning Rates and States from Biophysical Time Series: A Bayesian Approach to Model Selection and Single-Molecule FRET Data
Authors:
Jonathan E. Bronson,
Jingyi Fei,
Jake M. Hofman,
Ruben L. Gonzalez, Jr.,
Chris H. Wiggins
Abstract:
Time series data provided by single-molecule Forster resonance energy transfer (sm-FRET) experiments offer the opportunity to infer not only model parameters describing molecular complexes, e.g. rate constants, but also information about the model itself, e.g. the number of conformational states. Resolving whether or how many of such states exist requires a careful approach to the problem of mod…
▽ More
Time series data provided by single-molecule Forster resonance energy transfer (sm-FRET) experiments offer the opportunity to infer not only model parameters describing molecular complexes, e.g. rate constants, but also information about the model itself, e.g. the number of conformational states. Resolving whether or how many of such states exist requires a careful approach to the problem of model selection, here meaning discriminating among models with differing numbers of states. The most straightforward approach to model selection generalizes the common idea of maximum likelihood-selecting the most likely parameter values-to maximum evidence: selecting the most likely model. In either case, such inference presents a tremendous computational challenge, which we here address by exploiting an approximation technique termed variational Bayes. We demonstrate how this technique can be applied to temporal data such as smFRET time series; show superior statistical consistency relative to the maximum likelihood approach; and illustrate how model selection in such probabilistic or generative modeling can facilitate analysis of closely related temporal data currently prevalent in biophysics. Source code used in this analysis, including a graphical user interface, is available open source via http://vbFRET.sourceforge.net
△ Less
Submitted 20 July, 2009;
originally announced July 2009.