Study uses machine learning to predict future health from various health data types.
problem Predicting future health using diverse health data types.
method Applied machine learning (neural networks and XGBoost) to longitudinal data from 6830 individuals.
result Health-related measures were the strongest predictors of future health status, while genetic data performed poorly.
Although there are millions of transgender people in the world, a lack of information exists about their health issues. This issue has consequences for the medical field, which only has a nascent understanding of how to identify and meet this population's health-related needs. Social media sites like Twitter provide ne…
Twitter has been a prominent social media platform for mining population-level health data and accurate clustering of health-related tweets into topics is important for extracting relevant health insights. In this work, we propose deep convolutional autoencoders for learning compact representations of health-related tw…
Social media based digital epidemiology has the potential to support faster response and deeper understanding of public health related threats. This study proposes a new framework to analyze unstructured health related textual data via Twitter users' post (tweets) to characterize the negative health sentiments and non-…
With the expeditious advancement of information technologies, health-related data presented unprecedented potentials for medical and health discoveries but at the same time significant challenges for machine learning techniques both in terms of size and complexity. Those challenges include: the structured data with var…
A lack of information exists about the health issues of lesbian, gay, bisexual, transgender, and queer (LGBTQ) people who are often excluded from national demographic assessments, health studies, and clinical trials. As a result, medical experts and researchers lack a holistic understanding of the health disparities fa…
Community-based Question Answering (CQA) sites play an important role in addressing health information needs. However, a significant number of posted questions remain unanswered. Automatically answering the posted questions can provide a useful source of information for online health communities. In this study, we deve…
Stable health predictions need deconfounding test set features.
problem Stability of predictions in health machine learning is compromised by selection biases.
method Deconfounding the test set features improves prediction stability across different environments.
result Improved stability achieved by deconfounding test set features.
Study analyzes misinformation on social media during COVID-19.
problem Misinformation spreads on social media during the COVID-19 pandemic, affecting public health adherence.
method Analysis of model-labeled data, random forest classifier, sentiment analysis.
result Misinformation tweets show more negative sentiment and evolve over time, incorporating details from unrelated theories.
dCMF models evolving patterns in multiway data with temporal dynamics.
problem Capturing evolving patterns in multiway datasets with temporal dependencies.
method Time-aware coupled factorization model constrained by LDS structure.
result dCMF outperforms alternatives in capturing complex dynamics.
BayesMR estimates causal effects and directionality from genetic data.
problem Challenges in finding good genetic instruments and estimating causal effects.
method Bayesian Mendelian randomization approach that accounts for pleiotropy and reverse causation.
result BayesMR provides a posterior distribution over causal effects and uncertainty.
This study seeks to validate a search protocol of ill health-related terms using Twitter data which can later be used to understand if, and how, Twitter can reveal information on the current health situation. We extracted conversations related to health and disease postings on Twitter using a set of pre-defined keyword…
Multiple cause-of-death data provides a valuable source of information that can be used to enhance health standards by predicting health related trajectories in societies with large populations. These data are often available in large quantities across U.S. states and require Big Data techniques to uncover complex hidd…
Automated fatigue assessment using ECG and actigraphy sensors.
problem Fatigue assessment based on self-reporting suffers from recall bias.
method Wearable sensing, machine learning, feature selection, self-attention model, consistency self-attention mechanism.
result Very promising results achieved in fatigue assessment.
Satellite images predict U.S. county mortality rates.
problem Predicting mortality rates in U.S. counties using satellite imagery.
method Convolutional neural network trained on crude mortality rates, learned features interpreted using Shapley Additive Feature Explanations.
result Predicted mortality from satellite images correlated strongly with true mortality rates (Pearson r=0.72).
Enhances medical code predictions for multi-morbidity patients using text classification.
problem Improving accuracy in predicting medical codes for patients with multiple illnesses.
method Used machine learning techniques, including multi-label medical text classification, to enhance predictions.
result High dimensional embeddings pre-trained on health data significantly improve multi-label classification performance.
Recent years have witnessed a significant increase in the online sharing of medical information, with videos representing a large fraction of such online sources. Previous studies have however shown that more than half of the health-related videos on platforms such as YouTube contain misleading information and biases. …
Unified RL survey for healthcare AI interventions.
problem Limited real-life application of RL in healthcare.
method Unified technical survey and case studies.
result Bridge between dynamic treatment regimes and mobile health.
Weather2vec learns representations to adjust for non-local confounding in air pollution studies.
problem Non-local confounding in evaluating environmental policies and climate events on health outcomes.
method weather2vec framework using balancing scores to learn representations of non-local information.
result The framework effectively adjusts for confounding in air pollution studies.
Automatically assesses the quality of online health articles.
problem Lack of automated tools to evaluate the quality of online health information.
method Data mining approach using 10 quality criteria and feature selection.
result Classifier achieved 84%-90% accuracy on 10 criteria.
The introduction of data analytics into medicine has changed the nature of patient treatment. In this, patients are asked to disclose personal information such as genetic markers, lifestyle habits, and clinical history. This data is then used by statistical models to predict personalized treatments. However, due to pri…
Application of intelligent systems especially in smart homes and health-related topics has been drawing more attention in the last decades. Training Human Activity Recognition (HAR) models -- as a major module -- requires a fair amount of labeled data. Despite training with large datasets, most of the existing models w…
Interpretability of ML models improves healthcare decisions.
problem Ensuring machine learning models are understandable for healthcare users.
method Classifying interpretability into local and global approaches, and model-specific vs. model-agnostic methods.
result Examples of practical interpretability in healthcare, including prediction and treatment optimization.
Novel approach for robust domain generalization in health studies.
problem Challenges in making statistical inferences about underrepresented minority groups.
method Structured tensor completion for multi-dimensional domain generalization in linear regression models.
result Established rigorous theoretical guarantees and demonstrated minimax optimality.
Objectives: Electronic health records (EHRs) are only a first step in capturing and utilizing health-related data - the challenge is turning that data into useful information. Furthermore, EHRs are increasingly likely to include data relating to patient outcomes, functionality such as clinical decision support, and gen…
Bayesian model tackles spatial count data issues with flexible non-parametric techniques.
problem Challenges in traditional parametric models for spatial count data with unbalanced distributions and complex dependencies.
method Bayesian semi-parametric spatial dispersed count model combining non-parametric techniques and adapted count models.
result Demonstrates superior performance in managing dispersion and capturing intricate spatial patterns.
Dynamic topic model improves mental health note analysis for children.
problem Lack of longitudinal topic models for psychiatric clinical notes.
method Developed a dynamic topic model with consistent topics and individualized temporal dependencies.
result Achieved a 38% increase in topic coherence.
Depression and anxiety are critical public health issues affecting millions of people around the world. To identify individuals who are vulnerable to depression and anxiety, predictive models have been built that typically utilize data from one source. Unlike these traditional models, in this study, we leverage a rich …
Much recent research aims to identify evidence for Drug-Drug Interactions (DDI) and Adverse Drug reactions (ADR) from the biomedical scientific literature. In addition to this "Bibliome", the universe of social media provides a very promising source of large-scale data that can help identify DDI and ADR in ways that ha…
Bayesian model predicts mental health symptoms from IAT data, improving accuracy over D-score.
problem Limited predictive performance of D-score method for mental health assessment.
method Sparse hierarchical Bayesian model leveraging multi-modal data.
result AUCs of 0.73 (E-IAT) and 0.76 (PSY-IAT) in best modality configurations, significant after FDR correction.
Study invariant measures on measured laminations for subgroups of mapping class group.
problem Classify invariant Radon measures on space of measured laminations for subgroups of mapping class group.
method Geometric approach, focusing on recurrent measured laminations, explicitly constructing ergodic measures.
result Show uniquely ergodic for divergence-type subgroups, generalize results for full mapping class group.
New set-valued star-shaped risk measures introduced for better risk assessment.
problem Improving risk assessment in financial contexts.
method Developed new set-valued star-shaped risk measures and proved their representation theorems.
result Set-valued star-shaped risk measures can be represented as unions of set-valued convex risk measures.
The Bergman measure converges to the Zhang measure on a hybrid space.
problem Proving convergence of Bergman measures to Zhang measure.
method Analyzing convergence on a hybrid space and metrized curve complex.
result Bergman measure converges to Zhang measure on a hybrid space.
Bayesian approach to robust risk measures under model uncertainty.
problem Representing robust risk measures as a single probability measure.
method Introducing two types of risk measures and analyzing their relation to robust risk measures.
result Robust risk measures can be represented by a mixture probability measure, a Bayesian approach.
The paper studies dynamic star-shaped risk measures and their representation.
problem Representing dynamic star-shaped risk measures and their properties.
method Representation theorems for dynamic monetary and star-shaped risk measures.
result Dynamic star-shaped risk measures can be represented as the lower envelope of a family of dynamic convex risk measures.
Introduces Star-Shaped deviation measures for risk analysis.
problem Risk measurement and analysis in finance.
method Characterizes Star-Shaped deviation measures through acceptance sets and convex deviation measures.
result Exposes the relationship between Star-Shaped risk measures and deviation measures.
Transformers can interpolate between arbitrary measures.
problem Understanding the expressive power of Transformers as measure-to-measure maps.
method Provided an explicit choice of parameters for a single Transformer to match N arbitrary input measures to N arbitrary target measures.
result A single Transformer can interpolate between arbitrary measures.
Classifies invariant measures on specific character varieties.
problem Classifying invariant probability measures on character varieties.
method Measure disintegration along transverse Lagrangian tori fibrations.
result Ergodic measures are either counting measures on finite orbits or Liouville measures.
Paper characterizes star-shaped risk measures and their properties.
problem Characterizing risk measures in the presence of liquidity risk and competitive delegation.
method Characterization of star-shaped risk measures, study of their properties.
result Star-shaped risk measures include all practically used risk measures.
Paper compares fairness measures and feature importance measures using SHAP.
problem Comparing fairness measures and feature importance measures.
method Focus on SHAP, a game-theoretic measure of feature importance.
result Results for unfairness-prone datasets.
Paper introduces quasi-logconvex risk measures and their properties.
problem Characterizing and understanding new risk measures.
method Characterization through dual representation and properties of acceptance sets.
result Established dual representation and taxonomy of quasi-logconvex risk measures.
Submodularity is studied for convex risk measures, including Expected Shortfall.
problem Characterizing submodularity in convex risk measures.
method Analyzing submodularity properties of law-invariant coherent risk measures, including Expected Shortfall and Value-at-Risk.
result AES is submodular only when it reduces to ES, and empirical analysis shows AES violations are less frequent than VaR and ES violations.
New geometric measure simplifies complex analysis.
problem Complex geometric analysis challenges.
method Geometric integration and convergence methods.
result Smallest measure satisfying Area Formula.
The paper explores non-convex risk measures and their characterizations.
problem Characterizing non-convex risk measures without convexity or weak convexity.
method Characterizes monetary risk measures as lower envelopes of families of convex or coherent risk measures, considering law-invariance and SSD-consistency.
result Unified representation theorems for law-invariant risk measures, including VaR.
The paper calculates extreme measures in continuous time conic finance.
problem Determining valuation bounds for financial claims.
method Using dynamic spectral risk measures and estimating extreme measures from market data.
result Explicit formulas for extreme measures' Radon-Nykodim derivatives and estimation methods.
Paper characterizes monotonic mean-deviation risk measures.
problem Developing consistent risk measures from mean-deviation models.
method Applying a risk-weighting function to the deviation part of a mean-deviation model.
result Characterizes monotonic mean-deviation measures as consistent risk measures.
Introduces factor risk measures to assess risk relative to multiple factors.
problem Measuring risk relative to multiple factors.
method Introduces a double-argument mapping as a risk measure to assess risk relative to a vector of factors.
result Characterizes various types of factor risk measures including distortion, quantile, linear, and coherent measures.
Dual representations for robust risk measures and uncertainty sets.
problem Characterizing continuity of robust risk measures and their uncertainty sets.
method Develop dual representations for robust risk measures and uncertainty sets based on distinct geometric assumptions.
result Two dual frameworks for consolidated uncertainty sets are complementary, not interchangeable.