Modeling disease progression in irregularly observed patients.
problem Irregular patient observation in healthcare databases.
method Continuous-time hidden Markov model with generalized linear model.
result Interpretable model of healthcare utilization events.
New method infers causality from short memory-less transition data.
problem Inferring causality from short time series data.
method Composition of Transitions (COT) and machine learning models.
result Highly accurate in inferring causal relationships from short data.
Study predicts which patients will benefit from digital health interventions.
problem Unclear targeting of patients for digital care management programs.
method Analyzed claims data, combined with sociodemographic and app-generated data. Created two models: cost prediction and impactability classification.
result Random forest model accurately categorized patients as impactable or not, achieving 71.9% accuracy.
Machine learning in healthcare faces challenges due to complex data attributes.
problem Complex data attributes hinder accurate insights from machine learning models.
method Discusses preprocessing, model building, and interpretation challenges.
result Understanding data attributes is crucial for successful machine learning in healthcare.
This paper surveys deep learning applications in EHR analysis.
problem Leveraging EHR data for clinical informatics tasks.
method Reviews deep learning architectures and techniques applied to EHR data.
result Identifies current limitations and future research directions.
LMM predicts healthcare costs and risks with improved accuracy.
problem Wasteful healthcare spending and inefficiencies in risk prediction.
method Generative pre-trained transformer trained on patient event sequences.
result Improves cost prediction by 14.1% and chronic conditions prediction by 1.9%.
Germany's tax admin costs likely exceed 20% of total revenue, requiring system improvement.
problem High tax administrative costs in Germany and other jurisdictions.
method Statistical data, surveys, and a novel approach to measure total administrative cost as a percentage of total tax revenue.
result Germany's 2021 tax administrative costs likely exceeded 20% of total tax revenue.
Neural networks outperform logistic regression for predicting HF readmission.
problem Predicting 30-day all-cause readmission in heart failure patients.
method Used a large administrative claims dataset to compare neural network models (RNNCRF) with logistic regression models (LASSO) for predicting readmission.
result RNNCRF model achieved best performance with 0.642 AUC, while logistic regression with LASSO had equal performance.
New model maps malaria prevalence across Kenya's changing administrative boundaries.
problem Mapping disease prevalence with changing administrative boundaries.
method Combines deep learning and MCMC with aggVAE for disease mapping.
result Solves the change-of-support problem in disease surveillance.
This paper provides a guide to using machine learning in public administration.
problem Lack of clarity in proper use and potential pitfalls of machine learning methods.
method Provides a foundational view of machine learning and demonstrates its use in public administration research.
result Machine learning techniques can enrich public administration research and practice.
Study analyzes costs of managing research funds, developing a model for optimal administration.
problem High variability in administration costs among research funding agencies.
method Identified standard agency activities, developed a model estimating optimum portfolio success rate and administration ratio.
result Model estimates optimum portfolio success rate and administration ratio based on input variables.
Randomized machine learning methods improve suicide risk prediction from administrative data.
problem Accurately predicting suicide risk in mental health patients.
method Three randomized machine learning techniques: random forests, gradient boosting machines, and deep neural nets with dropout.
result Randomized methods outperform traditional approaches in predicting suicide risk with robustness against data redundancies.
Paper analyzes AI's impact on job tasks, predicting future demands.
problem AI's impact on job tasks and potential technological unemployment.
method Dynamic task shares analysis using ARIMA model on large job postings dataset.
result AI has risen in high wage occupations, predicting future task demands.
This paper addresses issues with the Brier score in administrative censoring scenarios.
problem Problems with the Brier score in administrative censoring scenarios.
method Proposes an alternative Brier score for administratively censored data.
result The administrative Brier score is valid even when censoring times can be identified from covariates.
System automates identification of cancer drug repurposing from PubMed.
problem Manual extraction of cancer drug repurposing evidence from scientific publications is infeasible.
method NLP pipeline including querying, filtering, entity extraction, classification, and study type classification.
result Automated system extracts cancer drug repurposing evidence from PubMed abstracts.
The paper analyzes fairness of compensation-based risk-sharing schemes for fund payouts.
problem Fair allocation of payouts in an endowment contingency fund.
method Analyzes two types of administrators and general non-negative loss distributions.
result General conditions for actuarial fairness are provided.
New method simplifies data analysis.
problem Complex data analysis challenges.
method Innovative algorithm for data simplification.
result Significant reduction in analysis time.
Dataset of Italian municipalities' income taxes from 2007-2011.
problem Understanding the economic structure of Italian municipalities.
method Annual aggregated income taxes of all Italian municipalities, clustered by regions and provinces.
result Data useful for economic comparisons and understanding municipal structures.
Study improves risk evaluation timing with right-censored reporting delays.
problem Improving risk evaluation under short observation windows due to administrative censoring.
method Jointly models parametric hazards for event and reporting processes, uses Monte Carlo expectation-maximization algorithm, and proposes transfer-learning procedure.
result Improves accuracy of timely risk evaluation under administrative censoring.
Probabilistic ML improves healthcare data analysis.
problem Insufficient understanding and incomplete data in healthcare.
method Examination of probabilistic machine learning models for healthcare challenges.
result Probabilistic models enhance healthcare data analysis and model building.
Big data analytics improves healthcare through early detection and quality life.
problem Limited access to healthcare data hinders evidence-based decision-making.
method Analysis of healthcare data using various tools and techniques.
result Big data analytics enhances healthcare quality and patient outcomes.
Deep learning improves CVD risk prediction from health records.
problem Predicting cardiovascular disease risk from administrative health data.
method Combined survival analysis and deep learning models.
result Deep learning models outperform traditional Cox models in accuracy and explained time-to-event occurrence.
Study on tax administration issues and their impact on Georgia's budget revenues.
problem Problems in revenue administration and tax rates in Georgia.
method Analyzed foreign experience and proposed a progressive tax system.
result A progressive tax system would benefit Georgia's business and economy.
MPVAA learns holistic patient representations from mixed healthcare data.
problem Learning personalized patient representations from heterogeneous healthcare data.
method Mixed Pooling Multi-View Attention Autoencoder (MPVAA) that integrates non-linear relationships among multiple data modalities.
result MPVAA generates more effective patient representations than state-of-the-art methods.
Modeling student course choices using latent variables.
problem Understanding student enrollment patterns in large universities.
method Probabilistic approach based on multilabel classification and mixture models.
result Demonstrated the model's ability to infer student interests guiding enrollment decisions.
Study uses data to analyze COPD patients' impact on hospital systems.
problem Understanding and quantifying resource requirements for COPD patients.
method Combines segmentation, queuing theory, and data recovery techniques.
result Finding useful operational results from incomplete administrative data.
Broad learning integrates diverse healthcare data for diagnostics and precision medicine.
problem Integrating various types of healthcare data for better diagnostics and personalized medicine.
method Fusing multi-view data including scalar, tensor, graph, and sequence data for knowledge discovery and machine learning tasks.
result Accurate user profiles and brain connectivity patterns can be created for improved diagnostics and personalized medicine.
This article was withdrawn by the arXiv.org administrators since it plagiarizes math.AT/0401211.
Machine learning faces challenges in healthcare data, but offers opportunities.
problem Poorly labeled data, multiple endotypes, and underrepresented healthy individuals.
method Review of existing machine learning methods and challenges in healthcare.
result Opportunities for machine learning in healthcare identified.
This article was withdrawn by the arXiv.org administrators since it plagiarizes math.GT/0011056.
Paper proposes a method to estimate confidence bands for survival random forests.
problem No statistically valid and computationally feasible approach for estimating confidence bands for survival random forests.
method Extending recent developments in infinite-order incomplete U-statistics, the paper proposes an unbiased confidence band estimation.
result The proposed method accurately estimates the confidence band and achieves desired coverage rate.
VHGM-MAE generates synthetic humans from healthcare data.
problem Handling high-dimensional, sparse healthcare data with missing values.
method Masked autoencoder (MAE) tailored for healthcare data, addressing heterogeneity, missingness, and high-dimensionality.
result VHGM-MAE outperforms existing methods in missing value imputation and synthetic data generation.
Study shows Lula's Zero Hunger program reduced income inequality in Brazil.
problem Income inequality in Brazil during Lula's administration.
method Breakpoint regression analysis using detailed descriptive statistics.
result The Zero Hunger program substantially reduced income inequality and provided income security for the poor.
Unsupervised model detects healthcare fraud from patient visit data.
problem Detecting fraudulent healthcare bills from patient visit data.
method Uses LSTM and seq2seq models for anomaly detection, normalizes scores with EDF.
result Improves anomaly detection for high class imbalance problems.
MCRAGE generates synthetic data to balance healthcare datasets.
problem Imbalanced datasets in healthcare lead to biased model performance for minority groups.
method Generative modeling to create synthetic data for underrepresented classes.
result MCRAGE improves model performance on minority groups.
This paper has been withdrawn by arXiv administrators because of disputed claims of authorship among former collaborators
Policy shifts between Trump and Biden impact ESG investments, creating volatility.
problem Dramatic policy shifts between Trump and Biden administrations affect ESG investments.
method Analyzes contrasting policies of Trump and Biden administrations and their impacts on ESG investments.
result Policy changes significantly influence ESG investments, leading to volatility and portfolio reassessment.
Deep generative model for healthcare data identifies coherent substructures and mutational clusters.
problem Analytical challenges in healthcare data, including sparsity, missingness, and small sample sizes.
method Proposes a deep generative Bayesian model with collapsed Gibbs sampling for multinomial count data.
result Identifies coherent substructures and biologically meaningful mutational clusters in cancer data.
Method fuses low and high-resolution data for better health estimates.
problem Improving high-resolution health estimates from mixed data sources.
method Fusion of unbiased low-resolution and potentially biased high-resolution data, learning a distribution consistent with sampling bias.
result Significant reduction in bias in high-resolution estimates.
Paper proposes a recursive PLS model for optimal response to security threats.
problem Optimal response to security threats after violations have occurred.
method Recursive Partial Least Squares (PLS) model with factorial analysis of security events.
result The model optimally estimates security administrators' responses to threats.
Study improves healthcare time series imputation by considering structured missingness.
problem Structured missingness in clinical data impacts time series imputation models.
method Analysis of different masking strategies on imputation methods using PhysioNet Challenge 2012 dataset.
result Masking choices significantly affect imputation accuracy and clinical prediction.
MiME learns EHR data structure for predictive healthcare tasks.
problem Data insufficiency in EHR for predictive healthcare tasks.
method Leverages multilevel structure of EHR data and learns multilevel embedding.
result MiME outperforms baseline methods in diverse evaluation settings.
Isthmus platform simplifies ML/AI integration in healthcare.
problem Challenges in deploying ML models in healthcare due to data quality, regulatory, and security issues.
method Turnkey, cloud-based platform addressing data quality, clinical relevance, and regulatory compliance.
result Reduces time to market for operationalizing ML/AI in healthcare.
Equity-Directed Bootstrapping improves model performance across groups in imbalanced datasets.
problem Improving model performance across different groups in imbalanced datasets.
method Equity-Directed Bootstrapping to balance training data with respect to both labels and group identity.
result The equity-directed bootstrap brings test set sensitivities and specificities closer to satisfying the equal odds criterion.
Research predicts healthcare index movements using historical OHLC data.
problem Predicting the directional movement of healthcare indices based on historical data.
method Supervised classification task with a one-step-ahead rolling window, using a diverse feature set including OHLC ratios.
result Robust predictive performance with accuracy exceeding 0.8 and Matthews correlation coefficients above 0.6, highlighting the importance of nowcasting features.
Paper presents a new method for clustering patient records using tensor decomposition.
problem Clustering high-dimensional binary data, especially in healthcare records.
method Tensor decomposition for an efficient and robust heuristic.
result Clinically meaningful results obtained on two healthcare datasets.
New method improves classification of healthcare data with missing values.
problem Predictive analytics on noisy, missing data with class imbalance.
method Multilevel Weighted Support Vector Machine (SVM) with imputation.
result Multilevel SVM produces more accurate and robust results.
A new method uses Hamiltonian Monte Carlo for imputation and augmentation of healthcare data.
problem Missing values in clinical studies lead to biased results and loss of statistical power.
method Folded Hamiltonian Monte Carlo (F-HMC) with Bayesian inference to handle high-dimensional, small sample size datasets.
result The method enriches the quality of data in precision, accuracy, recall, F1 score, and propensity metric.